| 2024 | Interspeech | Signal processing algorithm effective for sound quality of hearing loss simulators. | Toshio Irino, Shintaro Doan, Minami Ishikawa |
| 2023 | Interspeech | Impact of Residual Noise and Artifacts in Speech Enhancement Errors on Intelligibility of Human and Machine. | Shoko Araki, Ayako Yamamoto, Tsubasa Ochiai, Kenichi Arai, Atsunori Ogawa, Tomohiro Nakatani, Toshio Irino |
| 2022 | Interspeech | Speech intelligibility of simulated hearing loss sounds and its prediction using the Gammachirp Envelope Similarity Index (GESI). | Toshio Irino, Honoka Tamaru, Ayako Yamamoto |
| 2021 | Interspeech | Mixture of Orthogonal Sequences Made from Extended Time-Stretched Pulses Enables Measurement of Involuntary Voice Fundamental Frequency Response to Pitch Perturbation. | Hideki Kawahara, Toshie Matsui, Kohei Yatabe, Ken-Ichi Sakakibara, Minoru Tsuzaki, Masanori Morise, Toshio Irino |
| 2021 | Interspeech | Interactive and Real-Time Acoustic Measurement Tools for Speech Data Acquisition and Presentation: Application of an Extended Member of Time Stretched Pulses. | Hideki Kawahara, Kohei Yatabe, Ken-Ichi Sakakibara, Mitsunori Mizumachi, Masanori Morise, Hideki Banno, Toshio Irino |
| 2021 | Interspeech | Comparison of Remote Experiments Using Crowdsourcing and Laboratory Experiments on Speech Intelligibility. | Ayako Yamamoto, Toshio Irino, Kenichi Arai, Shoko Araki, Atsunori Ogawa, Keisuke Kinoshita, Tomohiro Nakatani |
| 2020 | Interspeech | Predicting Intelligibility of Enhanced Speech Using Posteriors Derived from DNN-Based ASR System. | Kenichi Arai, Shoko Araki, Atsunori Ogawa, Keisuke Kinoshita, Tomohiro Nakatani, Toshio Irino |
| 2020 | Interspeech | Speech Clarity Improvement by Vocal Self-Training Using a Hearing Impairment Simulator and its Correlation with an Auditory Modulation Index. | Toshio Irino, Soichi Higashiyama, Hanako Yoshigi |
| 2019 | Interspeech | Predicting Speech Intelligibility of Enhanced Speech Using Phone Accuracy of DNN-Based ASR System. | Kenichi Arai, Shoko Araki, Atsunori Ogawa, Keisuke Kinoshita, Tomohiro Nakatani, Katsuhiko Yamamoto, Toshio Irino |
| 2018 | Interspeech | Frequency Domain Variants of Velvet Noise and Their Application to Speech Processing and Synthesis. | Hideki Kawahara, Ken-Ichi Sakakibara, Masanori Morise, Hideki Banno, Tomoki Toda, Toshio Irino |
| 2018 | Interspeech | Multi-resolution Gammachirp Envelope Distortion Index for Intelligibility Prediction of Noisy Speech. | Katsuhiko Yamamoto, Toshio Irino, Narumi Ohashi, Shoko Araki, Keisuke Kinoshita, Tomohiro Nakatani |
| 2017 | Interspeech | An Auditory Model of Speaker Size Perception for Voiced Speech Sounds. | Toshio Irino, Eri Takimoto, Toshie Matsui, Roy D. Patterson |
| 2017 | Interspeech | A New Cosine Series Antialiasing Function and its Application to Aliasing-Free Glottal Source Models for Speech and Singing Synthesis. | Hideki Kawahara, Ken-Ichi Sakakibara, Masanori Morise, Hideki Banno, Tomoki Toda, Toshio Irino |
| 2017 | Interspeech | The Effect of Spectral Tilt on Size Discrimination of Voiced Speech Sounds. | Toshie Matsui, Toshio Irino, Kodai Yamamoto, Hideki Kawahara, Roy D. Patterson |
| 2017 | Interspeech | Predicting Speech Intelligibility Using a Gammachirp Envelope Distortion Index Based on the Signal-to-Distortion Ratio. | Katsuhiko Yamamoto, Toshio Irino, Toshie Matsui, Shoko Araki, Keisuke Kinoshita, Tomohiro Nakatani |
| 2016 | Interspeech | Speech Intelligibility Prediction Based on the Envelope Power Spectrum Model with the Dynamic Compressive Gammachirp Auditory Filterbank. | Katsuhiko Yamamoto, Toshio Irino, Toshie Matsui, Shoko Araki, Keisuke Kinoshita, Tomohiro Nakatani |
| 2015 | Interspeech | How the slope of the speech spectrum affects the perception of speaker size. | Kodai Yamamoto, Toshio Irino, Ryuichi Nisimura, Hideki Kawahara, Roy D. Patterson |
| 2014 | HCI | Development of a Mobile Application for Crowdsourcing the Data Collection of Environmental Sounds. | Minori Matsuyama, Ryuichi Nisimura, Hideki Kawahara, Junnosuke Yamada, Toshio Irino |
| 2014 | HCI | Proposal for an Interactive 3D Sound Playback Interface Controlled by User behavior. | Ryuichi Nisimura, Kazuki Hashimoto, Hideki Kawahara, Toshio Irino |
| 2014 | Interspeech | Vocal tract length estimation based on vowels using a database consisting of 385 speakers and a database with MRI-based vocal tract shape information. | Hideki Kawahara, Tatsuya Kitamura, Hironori Takemoto, Ryuichi Nisimura, Toshio Irino |
| 2014 | Interspeech | Excitation source analysis for high-quality speech manipulation systems based on an interference-free representation of group delay with minimum phase response compensation. | Hideki Kawahara, Masanori Morise, Tomoki Toda, Hideki Banno, Ryuichi Nisimura, Toshio Irino |
| 2013 | ICASSP | Higher order waveform symmetry measure and its application to periodicity detectors for speech and singing with fine temporal resolution. | Hideki Kawahara, Masanori Morise, Ryuichi Nisimura, Toshio Irino |
| 2013 | Interspeech | Beyond bandlimited sampling of speech spectral envelope imposed by the harmonic structure of voiced sounds. | Hideki Kawahara, Masanori Morise, Tomoki Toda, Ryuichi Nisimura, Toshio Irino |
| 2013 | Interspeech | Controlling "shout" expression in a Japanese POP singing performance: analysis and suppression study. | Yuri Nishigaki, Ken-Ichi Sakakibara, Masanori Morise, Ryuichi Nisimura, Toshio Irino, Hideki Kawahara |
| 2012 | Interspeech | Deviation measure of waveform symmetry and its application to high-speed and temporally-fine F0 extraction for vocal sound texture manipulation. | Hideki Kawahara, Masanori Morise, Ryuichi Nisimura, Toshio Irino |
| 2011 | HCI | Manual and Accelerometer Analysis of Head Nodding Patterns in Goal-oriented Dialogues. | Masashi Inoue, Toshio Irino, Nobuhiro Furuyama, Ryoko Hanada, Takako Ichinomiya, Hiroyasu Massaki |
| 2011 | HCI | Development of Web-Based Voice Interface to Identify Child Users Based on Automatic Speech Recognition System. | Ryuichi Nisimura, Shoko Miyamori, Lisa Kurihara, Hideki Kawahara, Toshio Irino |
| 2011 | ICASSP | An interference-free representation of instantaneous frequency of periodic signals and its application to F0 extraction. | Hideki Kawahara, Toshio Irino, Masanori Morise |
| 2011 | Interspeech | Auditory Filterbank Improves Voice Morphing. | Erika Okamoto, Toshio Irino, Ryuichi Nisimura, Hideki Kawahara |
| 2010 | ICASSP | High-quality and light-weight voice transformation enabling extrapolation without perceptual and objective breakdown. | Hideki Kawahara, Ryuichi Nisimura, Toshio Irino, Masanori Morise, Toru Takahashi, Hideki Banno |
| 2010 | Interspeech | Simplification and extension of non-periodic excitation source representations for high-quality speech manipulation systems. | Hideki Kawahara, Masanori Morise, Toru Takahashi, Hideki Banno, Ryuichi Nisimura, Toshio Irino |
| 2010 | ISCAS | Auditory speech processing for scale-shift covariance and its evaluation in automatic speech recognition. | Roy D. Patterson, Thomas C. Walters, Jessica Monaghan, Christian Feldbauer, Toshio Irino |
| 2009 | HCI | Development of Speech Input Method for Interactive VoiceWeb Systems. | Ryuichi Nisimura, Jumpei Miyake, Hideki Kawahara, Toshio Irino |
| 2009 | ICASSP | Temporally variable multi-aspect auditory morphing enabling extrapolation without objective and perceptual breakdown. | Hideki Kawahara, Ryuichi Nisimura, Toshio Irino, Masanori Morise, Toru Takahashi, Hideki Banno |
| 2009 | Interspeech | Observation of empirical cumulative distribution of vowel spectral distances and its application to vowel based voice conversion. | Hideki Kawahara, Masanori Morise, Toru Takahashi, Hideki Banno, Ryuichi Nisimura, Toshio Irino |
| 2009 | Interspeech | Influences of vowel duration on speaker-size estimation and discrimination. | Chihiro Takeshima, Minoru Tsuzaki, Toshio Irino |
| 2008 | ICASSP | Tandem-STRAIGHT: A temporally stable power spectral representation for periodic signals and applications to interference-free spectrum, F0, and aperiodicity estimation. | Hideki Kawahara, Masanori Morise, Toru Takahashi, Ryuichi Nisimura, Toshio Irino, Hideki Banno |
| 2008 | Interspeech | Spectral envelope recovery beyond the nyquist limit for high-quality manipulation of speech sounds. | Hideki Kawahara, Masanori Morise, Hideki Banno, Toru Takahashi, Ryuichi Nisimura, Toshio Irino |
| 2007 | Interspeech | Discrimination and recognition of scaled word sounds. | Toshio Irino, Yoshie Aoki, Yoshie Hayashi, Hideki Kawahara, Roy D. Patterson |
| 2006 | ICASSP | Dynamic, Compressive Gammachirp Auditory Filterbank for Perceptual Signal Processing. | Toshio Irino, Roy D. Patterson |
| 2006 | Interspeech | Analyzing dialogue data for real-world emotional speech classification. | Ryuichi Nisimura, Souji Omae, Hideki Kawahara, Toshio Irino |
| 2006 | Interspeech | Automatic assignment of anchoring points on vowel templates for defining correspondence between time-frequency representations of speech samples. | Toru Takahashi, Masashi Nishi, Toshio Irino, Hideki Kawahara |
| 2005 | Interspeech | Speech intelligibility derived from time-frequency and source smearing. | Toshio Irino, Satoru Satou, Shunsuke Nomura, Hideki Banno, Hideki Kawahara |
| 2005 | Interspeech | Nearly defect-free F0 trajectory extraction for expressive speech modifications based on STRAIGHT. | Hideki Kawahara, Alain de Cheveign, Hideki Banno, Toru Takahashi, Toshio Irino |
| 2005 | Interspeech | Voice and emotional expression transformation based on statistics of vowel parameters in an emotional speech database. | Toru Takahashi, Takeshi Fujii, Masashi Nishi, Hideki Banno, Toshio Irino, Hideki Kawahara |
| 2004 | ICASSP | Algorithm amalgam: morphing waveform based methods, sinusoidal models and STRAIGHT. | Hideki Kawahara, Hideki Banno, Toshio Irino, Parham Zolfaghari |
| 2004 | Interspeech | Intelligibility of degraded speech from smeared STRAIGHT spectrum. | Hideki Kawahara, Hideki Banno, Toshio Irino, Jiang Jin |
| 2004 | MMSP | A design of audio-visual talker tracking system based on CSP analysis and frame difference in real noisy environments. | Nishiura Denda, Takanobu Nishiura, Hideki Kawahara, Toshio Irino |
| 2003 | ICASSP | Speech segregation using event synchronous auditory vocoder. | Toshio Irino, Roy D. Patterson, Hideki Kawahara |
| 2003 | Interspeech | Speech segregation based on fundamental event information using an auditory vocoder. | Toshio Irino, Roy D. Patterson, Hideki Kawahara |
| 2003 | Interspeech | Dominance spectrum based v/UV classification and f_0 estimation. | Tomohiro Nakatani, Toshio Irino, Parham Zolfaghari |
| 2003 | Interspeech | Glottal closure instant synchronous sinusoidal model for high quality speech analysis/synthesis. | Parham Zolfaghari, Tomohiro Nakatani, Toshio Irino, Hideki Kawahara, Fumitada Itakura |
| 2002 | ICASSP | Auditory VOCODER: Speech resynthesis from an auditory Mellin representation. | Toshio Irino, Roy D. Patterson, Hideki Kawahara |
| 2002 | Interspeech | Evaluation of a speech recognition / generation method based on HMM and straight. | Toshio Irino, Yasuhiro Minami, Tomohiro Nakatani, Minoru Tsuzaki, H. Tagawa |
| 2002 | Interspeech | Robust fundamental frequency estimation against background noise and spectral distortion. | Tomohiro Nakatani, Toshio Irino |
| 2000 | Interspeech | Robust fundamental frequency estimation using instantaneous frequencies of harmonic components. | Yoshinori Atake, Toshio Irino, Hideki Kawahara, Jinlin Lu, Satoshi Nakamura, Kiyohiro Shikano |
| 1999 | ICASSP | Noise suppression using a time-varying, analysis/synthesis gamma chirp filterbank. | Toshio Irino |
| 1999 | Interspeech | Stabilised wavelet mellin transform: an auditory strategy for normalising sound-source size. | Toshio Irino, Roy D. Patterson |
| 1998 | ICASSP | A time-varying, analysis/synthesis auditory filterbank using the gammachirp. | Toshio Irino, Masashi Unoki |
| 1998 | ICONIP | The Gammachirp for Optimal Auditory Filtering. | Toshio Irino, Roy D. Patterson |
| 1996 | ICASSP | A 'gammachirp' function as an optimal auditory filter with the Mellin transform. | Toshio Irino |
| 1994 | Interspeech | A theory of asymmetric intensity enhancement around acoustic transients. | Toshio Irino, Roy D. Patterson |
| 1992 | ICASSP | Signal reconstruction from modified wavelet transform-An application to auditory signal processing. | Toshio Irino, Hideki Kawahara |