| 2022 | Interspeech | An objective test tool for pitch extractors' response attributes. | Hideki Kawahara, Kohei Yatabe, Ken-Ichi Sakakibara, Tatsuya Kitamura, Hideki Banno, Masanori Morise |
| 2022 | Interspeech | Perceptual Evaluation of Penetrating Voices through a Semantic Differential Method. | Tatsuya Kitamura, Naoki Kunimoto, Hideki Kawahara, Shigeaki Amano |
| 2021 | ICASSP | Cascaded All-Pass Filters with Randomized Center Frequencies and Phase Polarity for Acoustic and Speech Measurement and Data Augmentation. | Hideki Kawahara, Kohei Yatabe |
| 2021 | Interspeech | Mixture of Orthogonal Sequences Made from Extended Time-Stretched Pulses Enables Measurement of Involuntary Voice Fundamental Frequency Response to Pitch Perturbation. | Hideki Kawahara, Toshie Matsui, Kohei Yatabe, Ken-Ichi Sakakibara, Minoru Tsuzaki, Masanori Morise, Toshio Irino |
| 2021 | Interspeech | Interactive and Real-Time Acoustic Measurement Tools for Speech Data Acquisition and Presentation: Application of an Extended Member of Time Stretched Pulses. | Hideki Kawahara, Kohei Yatabe, Ken-Ichi Sakakibara, Mitsunori Mizumachi, Masanori Morise, Hideki Banno, Toshio Irino |
| 2019 | Interspeech | Investigating the Physiological and Acoustic Contrasts Between Choral and Operatic Singing. | Hiroko Terasawa, Kenta Wakasa, Hideki Kawahara, Ken-Ichi Sakakibara |
| 2018 | Interspeech | Frequency Domain Variants of Velvet Noise and Their Application to Speech Processing and Synthesis. | Hideki Kawahara, Ken-Ichi Sakakibara, Masanori Morise, Hideki Banno, Tomoki Toda, Toshio Irino |
| 2017 | Interspeech | A Modulation Property of Time-Frequency Derivatives of Filtered Phase and its Application to Aperiodicity and f | Hideki Kawahara, Ken-Ichi Sakakibara, Masanori Morise, Hideki Banno, Tomoki Toda |
| 2017 | Interspeech | A New Cosine Series Antialiasing Function and its Application to Aliasing-Free Glottal Source Models for Speech and Singing Synthesis. | Hideki Kawahara, Ken-Ichi Sakakibara, Masanori Morise, Hideki Banno, Tomoki Toda, Toshio Irino |
| 2017 | Interspeech | The Effect of Spectral Tilt on Size Discrimination of Voiced Speech Sounds. | Toshie Matsui, Toshio Irino, Kodai Yamamoto, Hideki Kawahara, Roy D. Patterson |
| 2016 | Interspeech | SparkNG: Interactive MATLAB Tools for Introduction to Speech Production, Perception and Processing Fundamentals and Application of the Aliasing-Free L-F Model Component. | Hideki Kawahara |
| 2016 | Interspeech | TUSK: A Framework for Overviewing the Performance of F0 Estimators. | Masanori Morise, Hideki Kawahara |
| 2015 | Interspeech | How the slope of the speech spectrum affects the perception of speaker size. | Kodai Yamamoto, Toshio Irino, Ryuichi Nisimura, Hideki Kawahara, Roy D. Patterson |
| 2014 | HCI | Development of a Mobile Application for Crowdsourcing the Data Collection of Environmental Sounds. | Minori Matsuyama, Ryuichi Nisimura, Hideki Kawahara, Junnosuke Yamada, Toshio Irino |
| 2014 | HCI | Proposal for an Interactive 3D Sound Playback Interface Controlled by User behavior. | Ryuichi Nisimura, Kazuki Hashimoto, Hideki Kawahara, Toshio Irino |
| 2014 | Interspeech | Vocal tract length estimation based on vowels using a database consisting of 385 speakers and a database with MRI-based vocal tract shape information. | Hideki Kawahara, Tatsuya Kitamura, Hironori Takemoto, Ryuichi Nisimura, Toshio Irino |
| 2014 | Interspeech | Excitation source analysis for high-quality speech manipulation systems based on an interference-free representation of group delay with minimum phase response compensation. | Hideki Kawahara, Masanori Morise, Tomoki Toda, Hideki Banno, Ryuichi Nisimura, Toshio Irino |
| 2013 | ICASSP | Higher order waveform symmetry measure and its application to periodicity detectors for speech and singing with fine temporal resolution. | Hideki Kawahara, Masanori Morise, Ryuichi Nisimura, Toshio Irino |
| 2013 | Interspeech | Beyond bandlimited sampling of speech spectral envelope imposed by the harmonic structure of voiced sounds. | Hideki Kawahara, Masanori Morise, Tomoki Toda, Ryuichi Nisimura, Toshio Irino |
| 2013 | Interspeech | Periodicity extraction for voiced sounds with multiple periodicity. | Masanori Morise, Hideki Kawahara, Kenji Ozawa |
| 2013 | Interspeech | Controlling "shout" expression in a Japanese POP singing performance: analysis and suppression study. | Yuri Nishigaki, Ken-Ichi Sakakibara, Masanori Morise, Ryuichi Nisimura, Toshio Irino, Hideki Kawahara |
| 2012 | ICASSP | Analysis and synthesis of strong vocal expressions: Extension and application of audio texture features to singing voice. | Hideki Kawahara, Masanori Morise |
| 2012 | Interspeech | Deviation measure of waveform symmetry and its application to high-speed and temporally-fine F0 extraction for vocal sound texture manipulation. | Hideki Kawahara, Masanori Morise, Ryuichi Nisimura, Toshio Irino |
| 2012 | Interspeech | Inharmonic speech: a tool for the study of speech perception and separation. | Josh H. McDermott, Daniel P. W. Ellis, Hideki Kawahara |
| 2012 | Interspeech | Pitch-Scaled Analysis based Residual Reconstruction for Speech Analysis and Synthesis. | Zhengqi Wen, Hideki Kawahara, Jianhua Tao |
| 2011 | HCI | Development of Web-Based Voice Interface to Identify Child Users Based on Automatic Speech Recognition System. | Ryuichi Nisimura, Shoko Miyamori, Lisa Kurihara, Hideki Kawahara, Toshio Irino |
| 2011 | ICASSP | An interference-free representation of instantaneous frequency of periodic signals and its application to F0 extraction. | Hideki Kawahara, Toshio Irino, Masanori Morise |
| 2011 | Interspeech | Auditory Filterbank Improves Voice Morphing. | Erika Okamoto, Toshio Irino, Ryuichi Nisimura, Hideki Kawahara |
| 2010 | ICASSP | High quality voice manipulation method based on the vocal tract area function obtained from sub-band LSP of straight spectrum. | Ayanori Arakawa, Yoshinori Uchimura, Hideki Banno, Fumitada Itakura, Hideki Kawahara |
| 2010 | ICASSP | High-quality and light-weight voice transformation enabling extrapolation without perceptual and objective breakdown. | Hideki Kawahara, Ryuichi Nisimura, Toshio Irino, Masanori Morise, Toru Takahashi, Hideki Banno |
| 2010 | Interspeech | Simplification and extension of non-periodic excitation source representations for high-quality speech manipulation systems. | Hideki Kawahara, Masanori Morise, Toru Takahashi, Hideki Banno, Ryuichi Nisimura, Toshio Irino |
| 2009 | HCI | Development of Speech Input Method for Interactive VoiceWeb Systems. | Ryuichi Nisimura, Jumpei Miyake, Hideki Kawahara, Toshio Irino |
| 2009 | ICASSP | Temporally variable multi-aspect auditory morphing enabling extrapolation without objective and perceptual breakdown. | Hideki Kawahara, Ryuichi Nisimura, Toshio Irino, Masanori Morise, Toru Takahashi, Hideki Banno |
| 2009 | Interspeech | Observation of empirical cumulative distribution of vowel spectral distances and its application to vowel based voice conversion. | Hideki Kawahara, Masanori Morise, Toru Takahashi, Hideki Banno, Ryuichi Nisimura, Toshio Irino |
| 2008 | ICASSP | Tandem-STRAIGHT: A temporally stable power spectral representation for periodic signals and applications to interference-free spectrum, F0, and aperiodicity estimation. | Hideki Kawahara, Masanori Morise, Toru Takahashi, Ryuichi Nisimura, Toshio Irino, Hideki Banno |
| 2008 | Interspeech | Spectral envelope recovery beyond the nyquist limit for high-quality manipulation of speech sounds. | Hideki Kawahara, Masanori Morise, Hideki Banno, Toru Takahashi, Ryuichi Nisimura, Toshio Irino |
| 2008 | Interspeech | Study on manipulation method of voice quality based on the vocal tract area function. | Yoshinori Uchimura, Hideki Banno, Fumitada Itakura, Hideki Kawahara |
| 2007 | Interspeech | Discrimination and recognition of scaled word sounds. | Toshio Irino, Yoshie Aoki, Yoshie Hayashi, Hideki Kawahara, Roy D. Patterson |
| 2006 | Interspeech | Analyzing dialogue data for real-world emotional speech classification. | Ryuichi Nisimura, Souji Omae, Hideki Kawahara, Toshio Irino |
| 2006 | Interspeech | Automatic assignment of anchoring points on vowel templates for defining correspondence between time-frequency representations of speech samples. | Toru Takahashi, Masashi Nishi, Toshio Irino, Hideki Kawahara |
| 2005 | Interspeech | Speech intelligibility derived from time-frequency and source smearing. | Toshio Irino, Satoru Satou, Shunsuke Nomura, Hideki Banno, Hideki Kawahara |
| 2005 | Interspeech | Nearly defect-free F0 trajectory extraction for expressive speech modifications based on STRAIGHT. | Hideki Kawahara, Alain de Cheveign, Hideki Banno, Toru Takahashi, Toshio Irino |
| 2005 | Interspeech | Voice and emotional expression transformation based on statistics of vowel parameters in an emotional speech database. | Toru Takahashi, Takeshi Fujii, Masashi Nishi, Hideki Banno, Toshio Irino, Hideki Kawahara |
| 2004 | ICASSP | Algorithm amalgam: morphing waveform based methods, sinusoidal models and STRAIGHT. | Hideki Kawahara, Hideki Banno, Toshio Irino, Parham Zolfaghari |
| 2004 | Interspeech | Intelligibility of degraded speech from smeared STRAIGHT spectrum. | Hideki Kawahara, Hideki Banno, Toshio Irino, Jiang Jin |
| 2004 | Interspeech | Procedure "senza vibrato": a key component for morphing singing. | Hideki Kawahara, Yumi Hirachi, Masanori Morise, Hideki Banno |
| 2004 | MMSP | A design of audio-visual talker tracking system based on CSP analysis and frame difference in real noisy environments. | Nishiura Denda, Takanobu Nishiura, Hideki Kawahara, Toshio Irino |
| 2003 | ICASSP | Speech segregation using event synchronous auditory vocoder. | Toshio Irino, Roy D. Patterson, Hideki Kawahara |
| 2003 | ICASSP | Auditory morphing based on an elastic perceptual distance metric in an interference-free time-frequency representation. | Hideki Kawahara, Hisami Matsui |
| 2003 | Interspeech | Speech enhancement with microphone array and fourier / wavelet spectral subtraction in real noisy environments. | Yuki Denda, Takanobu Nishiura, Hideki Kawahara |
| 2003 | Interspeech | Speech segregation based on fundamental event information using an auditory vocoder. | Toshio Irino, Roy D. Patterson, Hideki Kawahara |
| 2003 | Interspeech | Influence of recording equipment on the identification of second language phoneme contrasts. | Hiroaki Kato, Masumi Nukinay, Hideki Kawahara, Reiko Akahane-Yamada |
| 2003 | Interspeech | Investigation of emotionally morphed speech perception and its structure using a high quality speech manipulation system. | Hisami Matsui, Hideki Kawahara |
| 2003 | Interspeech | Glottal closure instant synchronous sinusoidal model for high quality speech analysis/synthesis. | Parham Zolfaghari, Tomohiro Nakatani, Toshio Irino, Hideki Kawahara, Fumitada Itakura |
| 2002 | ICASSP | Auditory VOCODER: Speech resynthesis from an auditory Mellin representation. | Toshio Irino, Roy D. Patterson, Hideki Kawahara |
| 2002 | Interspeech | On F0 trajectory optimization for very high-quality speech manipulation. | Hideki Kawahara, Parham Zolfaghari, Alain de Cheveign |
| 2001 | Interspeech | Comparative evaluation of F0 estimation algorithms. | Alain de Cheveign, Hideki Kawahara |
| 2001 | Interspeech | Systematic F0 glitches around nasal-vowel transitions. | Hideki Kawahara, Parham Zolfaghari |
| 2000 | Interspeech | Robust fundamental frequency estimation using instantaneous frequencies of harmonic components. | Yoshinori Atake, Toshio Irino, Hideki Kawahara, Jinlin Lu, Satoshi Nakamura, Kiyohiro Shikano |
| 2000 | Interspeech | Accurate vocal event detection method based on a fixed-point analysis of mapping from time to weighted average group delay. | Hideki Kawahara, Yoshinori Atake, Parham Zolfaghari |
| 2000 | Interspeech | Investigation of analysis and synthesis parameters of straight by subjective evaluation. | Parham Zolfaghari, Yoshinori Atake, Kiyohiro Shikano, Hideki Kawahara |
| 2000 | Interspeech | A sinusoidal model based on frequency-to-instantaneous frequency mapping. | Parham Zolfaghari, Hideki Kawahara |
| 1999 | Interspeech | Fixed point analysis of frequency to instantaneous frequency mapping for accurate estimation of F0 and periodicity. | Hideki Kawahara, Haruhiro Katayose, Alain de Cheveign, Roy D. Patterson |
| 1998 | ICASSP | Efficient representation of short-time phase based on group delay. | Hideki Banno, Jinlin Lu, Satoshi Nakamura, Kiyohiro Shikano, Hideki Kawahara |
| 1998 | ICONIP | Brain Creators: Japanese Initiative to Create Computational Models of Brain Functions. | Yasuji Sawada, Hideki Kawahara |
| 1998 | Interspeech | Computer-based second language production training by using spectrographic representation and HMM-based speech recognition scores. | Reiko Akahane-Yamada, Erik McDermott, Takahiro Adachi, Hideki Kawahara, John S. Pruitt |
| 1998 | Interspeech | An instantaneous-frequency-based pitch extraction method for high-quality speech transformation: revised TEMPO in the STRAIGHT-suite. | Hideki Kawahara, Alain de Cheveign, Roy D. Patterson |
| 1997 | ICASSP | Speech representation and transformation using adaptive interpolation of weighted spectrum: vocoder revisited. | Hideki Kawahara |
| 1996 | Interspeech | A neural matrix model for active tracking of frequency-modulated tones. | Kiyoaki Aikawa, Hideki Kawahara, Minoru Tsuzaki |
| 1996 | Interspeech | Effects of auditory feedback on F0 trajectory generation. | Hideki Kawahara, Hiroko Kato, J. C. Williams |
| 1994 | Interspeech | Effects of natural auditory feedback on fundamental frequency control. | Hideki Kawahara |
| 1993 | ICASSP | A dynamic cepstrum incorporating time-frequency masking and its application to continuous speech recognition. | Kiyoaki Aikawa, Harald Singer, Hideki Kawahara, Yoh'ichi Tohkura |
| 1992 | ICASSP | Signal reconstruction from modified wavelet transform-An application to auditory signal processing. | Toshio Irino, Hideki Kawahara |