| 2026 | HRI | Estimation of Mobile Robot Waiting Locations using Fluid Simulation without Prior Observation. | Ryoya Sakaki, Yoshio Ishiguro, Kento Ohtani, Takanori Nishino, Kazuya Takeda |
| 2025 | WACV | AAT-DA: Accident Anticipation Transformer with Driver Attention. | Yuto Kumamoto, Kento Ohtani, Daiki Suzuki, Minori Yamataka, Kazuya Takeda |
| 2024 | CVPR | Pseudo-label based unsupervised fine-tuning of a monocular 3D pose estimation model for sports motions. | Tomohiro Suzuki, Ryota Tanaka, Kazuya Takeda, Keisuke Fujii |
| 2024 | ICASSP | Audio Difference Learning for Audio Captioning. | Tatsuya Komatsu, Yusuke Fujita, Kazuya Takeda, Tomoki Toda |
| 2022 | BMVC | Improving Dense Representation Learning by Superpixelization and Contrasting Cluster Assignment. | Robin Karlsson, Tomoki Hayashi, Keisuke Fujii, Alexander Carballo, Kento Ohtani, Kazuya Takeda |
| 2022 | ICPR | Emergence of Collaborative Hunting via Multi-Agent Deep Reinforcement Learning. | Kazushi Tsutsui, Kazuya Takeda, Keisuke Fujii |
| 2022 | VTC | Disentangled Bad Weather Removal GAN for Pedestrian Detection. | Hanting Yang, Alexander Carballo, Kazuya Takeda |
| 2022 | UIST | Methods of Gently Notifying Pedestrians of Approaching Objects when Listening to Music. | Yuki Sakashita, Yoshio Ishiguro, Kento Ohtani, Takanori Nishino, Kazuya Takeda |
| 2020 | ICASSP | Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to-Speech Toolkit. | Tomoki Hayashi, Ryuichi Yamamoto, Katsuki Inoue, Takenori Yoshimura, Shinji Watanabe, Tomoki Toda, Kazuya Takeda, Yu Zhang, Xu Tan |
| 2020 | ICASSP | Weakly-Supervised Sound Event Detection with Self-Attention. | Koichi Miyazaki, Tatsuya Komatsu, Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda |
| 2020 | ICASSP | End-to-End Automatic Speech Recognition Integrated with CTC-Based Voice Activity Detection. | Takenori Yoshimura, Tomoki Hayashi, Kazuya Takeda, Shinji Watanabe |
| 2020 | Interspeech | Intelligibility Enhancement Based on Speech Waveform Modification Using Hearing Impairment. | Shu Hikosaka, Shogo Seki, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Hideki Banno, Tomoki Toda |
| 2020 | SiggraphA | Generation of Origami Folding Animations from 3D Point Cloud Using Latent Space Interpolation. | Chiaki Nakagaito, Takanori Nishino, Kazuya Takeda |
| 2019 | ASRU | Attention-Based Speech Recognition Using Gaze Information. | Osamu Segawa, Tomoki Hayashi, Kazuya Takeda |
| 2019 | ICASSP | Scene-dependent Anomalous Acoustic-event Detection Based on Conditional Wavenet and I-vector. | Tatsuya Komatsu, Tomoki Hayashi, Reishi Kondo, Tomoki Toda, Kazuya Takeda |
| 2019 | ICRA | A Predictive Reward Function for Human-Like Driving Based on a Transition Model of Surrounding Environment. | Daiki Hayashi, Yunfei Xu, Takashi Bando, Kazuya Takeda |
| 2019 | Interspeech | Pre-Trained Text Embeddings for Enhanced Text-to-Speech Synthesis. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda, Shubham Toshniwal, Karen Livescu |
| 2019 | Interspeech | Robustness of Statistical Voice Conversion Based on Direct Waveform Modification Against Background Sounds. | Yusuke Kurita, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2019 | ICRA | Point Cloud Compression for 3D LiDAR Sensor using Recurrent Neural Network with Residual Blocks. | Chenxi Tu, Eijiro Takeuchi, Alexander Carballo, Kazuya Takeda |
| 2018 | Interspeech | Multi-Head Decoder for End-to-End Speech Recognition. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Kazuya Takeda |
| 2018 | SMC | Learning How to Drive in Blind Intersections from Human Data. | Kyle Sama, Yoichi Morales, Naoki Akai, Eijiro Takeuchi, Kazuya Takeda |
| 2017 | ASRU | An investigation of multi-speaker training for wavenet vocoder. | Tomoki Hayashi, Akira Tamamori, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2017 | ICASSP | BLSTM-HMM hybrid system combined with sound activity detection network for polyphonic Sound Event Detection. | Tomoki Hayashi, Shinji Watanabe, Tomoki Toda, Takaaki Hori, Jonathan Le Roux, Kazuya Takeda |
| 2017 | ICASSP | Music staging AI. | Kenta Niwa, Kento Ohtani, Kazuya Takeda |
| 2017 | Interspeech | Speaker-Dependent WaveNet Vocoder. | Akira Tamamori, Tomoki Hayashi, Kazuhiro Kobayashi, Kazuya Takeda, Tomoki Toda |
| 2016 | Interspeech | Robust Example Search Using Bottleneck Features for Example-Based Speech Enhancement. | Atsunori Ogawa, Shogo Seki, Keisuke Kinoshita, Marc Delcroix, Takuya Yoshioka, Tomohiro Nakatani, Kazuya Takeda |
| 2015 | ICASSP | Exploring multi-channel features for denoising-autoencoder-based speech enhancement. | Shoko Araki, Tomoki Hayashi, Marc Delcroix, Masakiyo Fujimoto, Kazuya Takeda, Tomohiro Nakatani |
| 2015 | Interspeech | Integration of deep bottleneck features for audio-visual speech recognition. | Hiroshi Ninomiya, Norihide Kitaoka, Satoshi Tamura, Yurie Iribe, Kazuya Takeda |
| 2014 | ICASSP | Stochastic modeling and disaggregation of energy-consumption behavior. | Panikos Heracleous, Pongtep Angkititrakul, Kazuya Takeda |
| 2013 | CogSci | A Discussion on the Consistency of Driving Behavior across Laboratory and Real Situational Studies. | Hitoshi Terai, Kazuhisa Miwa, Hiroyuki Okuda, Yuichi Tazaki, Tatsuya Suzuki, Kazuaki Kojima, Junya Morita, Akihiro Maehigashi, Kazuya Takeda |
| 2013 | ICASSP | Analysis and modeling of entrainment in chorus singing. | Motonari Kawagishi, Shota Kawabuchi, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2013 | ICASSP | Modeling head-related transfer functions via spatial-temporal Gaussian process. | Tatsuya Komatsu, Takanori Nishino, Gareth W. Peters, Tomoko Matsui, Kazuya Takeda |
| 2013 | ICASSP | Computationally efficient single channel dereverberation based on complementary wiener filter. | Kazunobu Kondo, Yu Takahashi, Tatsuya Komatsu, Takanori Nishino, Kazuya Takeda |
| 2013 | ICASSP | Estimation of vocal tract parameters for the classification of speech under stress. | Xiao Yao, Takatoshi Jitsuhiro, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2013 | Interspeech | Classification of speech under stress by modeling the aerodynamics of the laryngeal ventricle. | Xiao Yao, Takatoshi Jitsuhiro, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2012 | CogSci | Multi-platform Experiment to Discuss Behavioral Consistency across Laboratory and Real Situational Studies. | Hitoshi Terai, Kazuhisa Miwa, Hiroyuki Okuda, Yuichi Tazaki, Tatsuya Suzuki, Kazuaki Kojima, Junya Morita, Akihiro Maehigashi, Kazuya Takeda |
| 2012 | ICASSP | Estimating sound source depth using a small-size array. | Satoshi Esaki, Kenta Niwa, Takanori Nishino, Kazuya Takeda |
| 2012 | ICASSP | Physical characteristics of vocal folds during speech under stress. | Xiao Yao, Takatoshi Jitsuhiro, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2012 | Interspeech | Classification of Stressed Speech Using Physical Parameters Derived from Two-Mass Model. | Xiao Yao, Takatoshi Jitsuhiro, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2012 | LREC | Causal analysis of task completion errors in spoken music retrieval interactions. | Sunao Hara, Norihide Kitaoka, Kazuya Takeda |
| 2011 | ASRU | Robust seed model training for speaker adaptation using pseudo-speaker features generated by inverse CMLLR transformation. | Arata Itoh, Sunao Hara, Norihide Kitaoka, Kazuya Takeda |
| 2011 | ICASSP | Driver risk evaluation based on acceleration, deceleration, and steering behavior. | Chiyomi Miyajima, Hiroki Ukai, Atsumi Naito, Hideomi Amata, Norihide Kitaoka, Kazuya Takeda |
| 2011 | ICASSP | Improving head-related impulse response measured in noisy environments with spatio-temporal frequency analysis. | Takanori Nishino, Kazuya Takeda |
| 2011 | Interspeech | Detection of Task-Incomplete Dialogs Based on Utterance-and-Behavior Tag N-Gram for Spoken Dialog Systems. | Sunao Hara, Norihide Kitaoka, Kazuya Takeda |
| 2010 | ICASSP | A small dodecahedral microphone array for blind source separation. | Motoki Ogasawara, Takanori Nishino, Kazuya Takeda |
| 2010 | ICASSP | Analyzing grasping for inferring cognitive states of users. | Kotaro Ogino, Takatoshi Jitsuhiro, Chiyomi Miyajima, Kazuya Takeda |
| 2010 | Interspeech | Automatic detection of task-incompleted dialog for spoken dialog system based on dialog act n-gram. | Sunao Hara, Norihide Kitaoka, Kazuya Takeda |
| 2010 | LREC | Estimation Method of User Satisfaction Using N-gram-based Dialog History Model for Spoken Dialog System. | Sunao Hara, Norihide Kitaoka, Kazuya Takeda |
| 2009 | ICASSP | Spoken dialog strategy based on understanding graph search. | Yuji Kinoshita, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2009 | ICASSP | Stochastic modeling of vehicle trajectory during lane-changing. | Yoshihiro Nishiwaki, Chiyomi Miyajima, Hidenori Kitaoka, Kazuya Takeda |
| 2009 | ICASSP | Feature transformation based on discriminant analysis preserving local structure for speech recognition. | Makoto Sakai, Norihide Kitaoka, Kazuya Takeda |
| 2009 | MMSP | A multimedia corpus of driving behaviors. | Lucas Malta, Akira Ozaki, Chiyomi Miyajima, Norihide Kitaoka, Kazuya Takeda |
| 2008 | ICASSP | Encoding large array signals into a 3D sound field representation for selective listening point audio based on blind source separation. | Kenta Niwa, Takanori Nishino, Kazuya Takeda |
| 2008 | ICMI | An integrative recognition method for speech and gestures. | Madoka Miki, Chiyomi Miyajima, Takanori Nishino, Norihide Kitaoka, Kazuya Takeda |
| 2008 | Interspeech | CENSREC-4: development of evaluation framework for distant-talking speech recognition under reverberant environments. | Masato Nakayama, Takanobu Nishiura, Yuki Denda, Norihide Kitaoka, Kazumasa Yamamoto, Takeshi Yamada, Satoru Tsuge, Chiyomi Miyajima, Masakiyo Fujimoto, Tetsuya Takiguchi, Satoshi Tamura, Tetsuji Ogawa, Shigeki Matsuda, Shingo Kuroiwa, Kazuya Takeda, Satoshi Nakamura |
| 2008 | Interspeech | Parameter estimation method of F0 control model for singing voices. | Yasunori Ohishi, Hirokazu Kameoka, Kunio Kashino, Kazuya Takeda |
| 2008 | Interspeech | Building and combining document and music spaces for music query-by-webpage system. | Ryoei Takahashi, Yasunori Ohishi, Norihide Kitaoka, Kazuya Takeda |
| 2008 | LREC | Evaluation Framework for Distant-talking Speech Recognition under Reverberant Environments: newest Part of the CENSREC Series -. | Takanobu Nishiura, Masato Nakayama, Yuki Denda, Norihide Kitaoka, Kazumasa Yamamoto, Takeshi Yamada, Satoru Tsuge, Chiyomi Miyajima, Masakiyo Fujimoto, Tetsuya Takiguchi, Satoshi Tamura, Shingo Kuroiwa, Kazuya Takeda, Satoshi Nakamura |
| 2008 | LREC | In-car Speech Data Collection along with Various Multimodal Signals. | Akira Ozaki, Sunao Hara, Takashi Kusakawa, Chiyomi Miyajima, Takanori Nishino, Norihide Kitaoka, Katunobu Itou, Kazuya Takeda |
| 2008 | MMSP | 3DAV integrated system featuring arbitrary listening-point and viewpoint generation. | Mehrdad Panahpour Tehrani, Kenta Niwa, Norishige Fukushima, Yasushi Hirano, Toshiaki Fujii, Masayuki Tanimoto, Kazuya Takeda, Kenji Mase, Akio Ishikawa, Shigeyuki Sakazawa, Atsushi Koike |
| 2007 | ASRU | Development of VAD evaluation framework CENSREC-1-C and investigation of relationship between VAD and speech recognition performance. | Norihide Kitaoka, Kazumasa Yamamoto, Tomohiro Kusamizu, Seiichi Nakagawa, Takeshi Yamada, Satoru Tsuge, Chiyomi Miyajima, Takanobu Nishiura, Masato Nakayama, Yuki Denda, Masakiyo Fujimoto, Tetsuya Takiguchi, Satoshi Tamura, Shingo Kuroiwa, Kazuya Takeda, Satoshi Nakamura |
| 2007 | ICMI | Statistical segmentation and recognition of fingertip trajectories for a gesture interface. | Kazuhiro Morimoto, Chiyomi Miyajima, Norihide Kitaoka, Katunobu Itou, Kazuya Takeda |
| 2006 | ICASSP | Multichannel Speech Enhancement Based on Speech Spectral Magnitude Estimation Using Generalized Gamma Prior Distribution. | Tran Huy Dat, Kazuya Takeda, Fumitada Itakura |
| 2006 | ICASSP | Development of Micro-Dodecahedral Loudspeaker for Measuring Head-Related Transfer Functions in The Proximal region. | Seiichiro Hosoe, Takanori Nishino, Katsunobu Itou, Kazuya Takeda |
| 2006 | ICASSP | Adaptive Regression Based Framework for In-Car Speech Recognition. | Weifeng Li, Katsunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2006 | ICASSP | Cepstral Analysis of Driving Behavioral Signals for Driver Identification. | Chiyomi Miyajima, Yoshihiro Nishiwaki, Koji Ozawa, Toshihiro Wakita, Katsunobu Itou, Kazuya Takeda |
| 2006 | ICASSP | Arbitrary Listening-Point Generation Using Sub-Band Representation of Sound Wave Ray-Space. | Mehrdad Panahpour Tehrani, Yasushi Hirano, Toshiaki Fujii, Shoji Kajita, Kazuya Takeda, Kenji Mase |
| 2006 | Interspeech | CENSREC2: corpus and evaluation environments for in car continuous digit speech recognition. | Satoshi Nakamura, Masakiyo Fujimoto, Kazuya Takeda |
| 2006 | LREC | Statistical Analysis for Thesaurus Construction using an Encyclopedic Corpus. | Yasunori Ohishi, Katunobu Itou, Kazuya Takeda, Atsushi Fujii |
| 2005 | ICASSP | Generalized gamma modeling of speech and its online estimation for speech enhancement. | Tran Huy Dat, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICASSP | Analysis of a large in-car speech corpus and its application to the multimodel ASR. | Hiroshi Fujimura, Chiyomi Miyajima, Katsunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICASSP | Spatial coding based on the extraction of moving sound sources in wavefield synthesis. | Toshiyuki Kimura, Kazuhiko Kakehi, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICASSP | Two-stage Noise Spectra Estimation and Regression based In-car Speech Recognition using Single Distant Microphone. | Weifeng Li, Katunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICASSP | SNR and Local Noise Power Estimations Based on Gaussian Mixture Modeling on the Log-Power Domain. | Kazuya Takeda, Tran Huy Dat, Hiroshi Fujimura, Fumitada Itakura |
| 2005 | ICDE | A speech enhancement system based on data clustering and cumulative histogram equalization. | Tran Huy Dat, Kazuya Takeda, Fumitada Itakura |
| 2005 | ICDE | CENSREC-3: Data Collection for In-Car Speech Recognition and Its Common Evaluation Framework. | Masakiyo Fujimoto, Satoshi Nakamura, Toshiki Endo, Kazuya Takeda, Chiyomi Miyajima, Shingo Kuroiwa, Takeshi Yamada, Norihide Kitaoka, Kazumasa Yamamoto, Mitsunori Mizumachi, Takanobu Nishiura, Akira Sasou |
| 2005 | ICDE | Analysis of a large in-car speech corpus. | Hiroshi Fujimura, Nobuo Kawaguchi, Shigeki Matsubara, Katsunobu Itou, Kazuya Takeda |
| 2005 | ICDE | Improved Noise Spectra Estimation and Log-spectral Regression for In-car Speech Recognition. | Weifeng Li, Katunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | Interspeech | Subjective and objective quality assessment of regression-enhanced speech in real car environments. | Weifeng Li, Katunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2005 | Interspeech | Discrimination between singing and speaking voices. | Yasunori Ohishi, Masataka Goto, Katunobu Itou, Kazuya Takeda |
| 2005 | Interspeech | Data collection and evaluation of speech recognition for motorbike riders. | H. Tanaka, Hiroshi Fujimura, Chiyomi Miyajima, Takanori Nishino, Katunobu Itou, Kazuya Takeda |
| 2005 | Interspeech | Speaker verification using Gaussian mixture models within changing real car environments. | Xianxian Zhang, John H. L. Hansen, Pongtep Angkititrakul, Kazuya Takeda |
| 2005 | PIMRC | Performance Evaluation of H.264 Video Streaming over Inter-Vehicular 802.11 Ad Hoc Networks. | Paolo Bucciol, Enrico Masala, Nobuo Kawaguchi, Kazuya Takeda, Juan Carlos De Martin |
| 2004 | Interspeech | Speech recognition using synchronization between speech and finger tapping. | Hiromitsu Ban, Chiyomi Miyajima, Katsunobu Itou, Fumitada Itakura, Kazuya Takeda |
| 2004 | Interspeech | Analysis of in-car speech recognition experiments using a large-scale multi-mode dialogue corpus. | Hiroshi Fujimura, Katsunobu Itou, Kazuya Takeda, Fumitada Itakura |
| 2004 | Interspeech | CIAIR in-car speech database. | Nobuo Kawaguchi, Shigeki Matsubara, Yukiko Yamaguchi, Kazuya Takeda, Fumitada Itakura |
| 2004 | Interspeech | Recent progress of open-source LVCSR engine julius and Japanese model repository. | Tatsuya Kawahara, Akinobu Lee, Kazuya Takeda, Katsunobu Itou, Kiyohiro Shikano |
| 2004 | Interspeech | Optimizing regression for in-car speech recognition using multiple distributed microphones. | Weifeng Li, Fumitada Itakura, Kazuya Takeda |
| 2004 | Interspeech | Speech enhancement based on magnitude estimation using the gamma prior. | Weifeng Li, Kazuya Takeda, Fumitada Itakura, Tran Huy Dat |
| 2004 | Interspeech | Example-based spoken dialogue system with online example augmentation. | Hiroya Murao, Nobuo Kawaguchi, Shigeki Matsubara, Yukiko Yamaguchi, Kazuya Takeda, Yasuyoshi Inagaki |
| 2004 | Interspeech | Audio-visual SPeaker localization for car navigation systems. | Xianxian Zhang, Kazuya Takeda, John H. L. Hansen, Toshiki Maeno |
| 2003 | ICASSP | In-car speech recognition using distributed microphones-adapting to automatically detected driving conditions. | Hideki Banno, Tetsuya Shinde, Kazuya Takeda, Fumitada Itakura |
| 2003 | Interspeech | A study on domain recognition of spoken dialogue systems. | Toshihiro Isobe, Shoji Hayakawa, Hiroya Murao, Tatsuji Mizutani, Kazuya Takeda, Fumitada Itakura |
| 2003 | Interspeech | Integration of noise reduction algorithms for Aurora2 task. | Takeshi Yamada, Jiro Okada, Kazuya Takeda, Norihide Kitaoka, Masakiyo Fujimoto, Shingo Kuroiwa, Kazumasa Yamamoto, Takanobu Nishiura, Mitsunori Mizumachi, Satoshi Nakamura |
| 2002 | ICASSP | Synthesis of car noise based on a composition of engine noise and friction noise. | Yoshihide Ban, Hideki Banno, Kazuya Takeda, Fumitada Itakura |
| 2002 | ICASSP | Acoustic analysis and recognition of whispered speech. | Taisuke Ito, Kazuya Takeda, Fumitada Itakura |
| 2002 | ICASSP | Spatial compression of multi-channel audio signals using inverse filters. | Toshiyuki Kimura, Kazuhiko Kakehi, Kazuya Takeda, Fumitada Itakura |
| 2002 | Interspeech | Recognition of continuous speech segments of monophone units using support vector machines. | Weifeng Lee, C. Chandra Sekhar, Kazuya Takeda, Fumitada Itakura |
| 2002 | Interspeech | Multiple regression of log-spectra for in-car speech recognition. | Tetsuya Shinde, Kazuya Takeda, Fumitada Itakura |
| 2002 | Interspeech | Experiments on recognition of lavalier microphone speech and whispered speech in real world environments. | Kiyoshi Tatara, Taisuke Ito, Parham Zolfaghari, Kazuya Takeda, Fumitada Itakura |
| 2002 | LREC | Multi-Dimensional Data Acquisition for Integrated Acoustic Information Research. | Nobuo Kawaguchi, Shigeki Matsubara, Kazuya Takeda, Fumitada Itakura |
| 2002 | LREC | The Present Status of Speech Database in Japan: Development, Management, and Application to Speech Research. | Hisao Kuwabara, Shuichi Itahashi, Mikio Yamamoto, Toshiyuki Takezawa, Satoshi Nakamura, Kazuya Takeda |
| 2002 | LREC | Continuous Speech Recognition Consortium an Open Repository for CSR Tools and Models. | Akinobu Lee, Tatsuya Kawahara, Kazuya Takeda, Masato Mimura, Atsushi Yamada, Akinori Ito, Katsunobu Itou, Kiyohiro Shikano |
| 2001 | ESANN | Recognition of consonant-vowel utterances using Support Vector Machines. | Chellu Chandra Sekhar, Kazuya Takeda, Fumitada Itakura |
| 2001 | ICANN | Close-Class-Set Discrimination Method for Recognition of Stop_Consonant-Vowel Utterances Using Support Vector Machines. | Chellu Chandra Sekhar, Kazuya Takeda, Fumitada Itakura |
| 2001 | ICASSP | A study on perceptual distance measure for phase spectrum of stimuli. | Hideki Banno, Kazuya Takeda, Fumitada Itakura |
| 2001 | ICASSP | Direction of arrival estimation based on nonlinear microphone array. | Hidekazu Kamiyanagida, Hiroshi Saruwatari, Kazuya Takeda, Fumitada Itakura |
| 2001 | ICASSP | Blind source separation combining frequency-domain ICA and beamforming. | Hiroshi Saruwatari, Satoshi Kurita, Kazuya Takeda |
| 2001 | ICASSP | Continuous speech recognition without end-point detection. | Osamu Segawa, Kazuya Takeda, Fumitada Itakura |
| 2001 | Interspeech | Multimedia data collection of in-car speech communication. | Nobuo Kawaguchi, Shigeki Matsubara, Kazuya Takeda, Fumitada Itakura |
| 2001 | Interspeech | Robust speech recognition based on selective use of missing frequency band HMMs. | Takayoshi Kawamura, Kazuya Takeda, Fumitada Itakura |
| 2000 | ICASSP | Evaluation of blind signal separation method using directivity pattern under reverberant conditions. | Satoshi Kurita, Hiroshi Saruwatari, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 2000 | ICASSP | A new phonetic tied-mixture model for efficient decoding. | Akinobu Lee, Tatsuya Kawahara, Kazuya Takeda, Kiyohiro Shikano |
| 2000 | ICASSP | Speech enhancement using nonlinear microphone array with noise adaptive complementary beamforming. | Hiroshi Saruwatari, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 2000 | ICASSP | Speech recognition based on space diversity using distributed multi-microphone. | Yasuhiro Shimizu, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 2000 | ICASSP | An acoustic measure for predicting recognition performance degradation. | Kazuya Takeda, Masaaki Kondo, Fumitada Itakura |
| 2000 | Interspeech | Construction of speech corpus in moving car environment. | Nobuo Kawaguchi, Shigeki Matsubara, Hiroyuki Iwa, Shoji Kajita, Kazuya Takeda, Fumitada Itakura, Yasuyoshi Inagaki |
| 2000 | Interspeech | Free software toolkit for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Akinobu Lee, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Shigeki Sagayama, Katsunobu Itou, Akinori Ito, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 2000 | Interspeech | Blind source separation based on subband ICA and beamforming. | Hiroshi Saruwatari, Satoshi Kurita, Kazuya Takeda, Fumitada Itakura, Kiyohiro Shikano |
| 2000 | Interspeech | Vector space representation of language probabilities through SVD of n-gram matrix. | Shiro Terashima, Kazuya Takeda, Fumitada Itakura |
| 2000 | LREC | IPA Japanese Dictation Free Software Project. | Katsunobu Itou, Kiyohiro Shikano, Tatsuya Kawahara, Kazuya Takeda, Atsushi Yamada, Akinori Ito, Takehito Utsuro, Tetsunori Kobayashi, Nobuaki Minematsu, Mikio Yamamoto, Shigeki Sagayama, Akinobu Lee |
| 1999 | ICASSP | Audio data hiding by use of band-limited random sequences. | Mikio Ikeda, Kazuya Takeda, Fumitada Itakura |
| 1999 | ICASSP | Compensating of room acoustic transfer functions affected by change of room temperature. | Michiaki Omura, Motohiko Yada, Hiroshi Saruwatari, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1999 | ICASSP | Speech enhancement using nonlinear microphone array with complementary beamforming. | Hiroshi Saruwatari, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1999 | Interspeech | Speaker conversion through non-linear frequency warping of straight spectrum. | Noriyasu Maeda, Hideki Banno, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1999 | Interspeech | Speech enhancement using nonlinear microphone array under nonstationary noise conditions. | Hiroshi Saruwatari, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1998 | ICASSP | Spectral weighting of SBCOR for noise robust speech recognition. | Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1998 | ICASSP | Balancing acoustic and linguistic probabilities. | Atsunori Ogawa, Kazuya Takeda, Fumitada Itakura |
| 1998 | Interspeech | The design of the newspaper-based Japanese large vocabulary continuous speech recognition corpus. | Katunobu Itou, Mikio Yamamoto, Kazuya Takeda, Toshiyuki Takezawa, Tatsuo Matsuoka, Tetsunori Kobayashi, Kiyohiro Shikano, Shuichi Itahashi |
| 1998 | Interspeech | Sharable software repository for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Katsunobu Itou, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 1998 | Interspeech | Estimating entropy of a language from optimal word insertion penalty. | Kazuya Takeda, Atsunori Ogawa, Fumitada Itakura |
| 1997 | ICASSP | A binaural speech processing method using subband-cross correlation analysis for noise robust recognition. | Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1997 | Interspeech | Voice activity detection using source separation techniques. | Tomohiko Taniguchi, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1996 | Interspeech | Subband-crosscorrelation analysis for robust speech recognition. | Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1996 | Interspeech | Extracting speech features from human speech-like noise. | Daisuke Kobayashi, Shoji Kajita, Kazuya Takeda, Fumitada Itakura |
| 1996 | Interspeech | Variability of lombard effects under different noise conditions. | Atsushi Wakao, Kazuya Takeda, Fumitada Itakura |
| 1995 | Interspeech | A prototype of a Japanese-Korean realtime speech translation system. | Masami Suzuki, Naomi Inoue, Fumihiro Yato, Kazuya Takeda, Seiichi Yamamoto |
| 1995 | Interspeech | Top-down speech detection and n-best meaning search in a voice activated telephone extension system. | Kazuya Takeda, Shingo Kuroiwa, Masaki Naito, Seiichi Yamamoto |
| 1994 | Interspeech | A trellis-based implementation of minimum error rate training. | Kazuya Takeda, Tetsunori Murakami, Shingo Kuroiwa, Seiichi Yamamoto |
| 1993 | Interspeech | A voice-activated extension telephone exchange system. | Shingo Kuroiwa, Kazuya Takeda, Naomi Inoue, Izuru Nogaito, Seiichi Yamamoto, Makoto Shozakai, Kunihiko Owa, Masahiko Takahashi, Ryuuji Matsumoto |
| 1993 | Interspeech | Improving robustness of network grammar by using class HMM. | Kazuya Takeda, Naomi Inoue, Shingo Kuroiwa, Tomohiro Konuma, Seiichi Yamamoto |
| 1992 | Interspeech | Architecture and algorithms of a real-time word recognizer for telephone input. | Shingo Kuroiwa, Kazuya Takeda, Fumihiro Yato, Seiichi Yamamoto, Kunihiko Owa, Makoto Shozakai, Ryuuji Matsumoto |
| 1990 | Interspeech | Statistical analysis for segmental duration rules in Japanese speech synthesis. | Nobuyoshi Kaiki, Kazuya Takeda, Yoshinori Sagisaka |
| 1990 | Interspeech | A large-scale Japanese speech database. | Yoshinori Sagisaka, Kazuya Takeda, M. Abel, Shigeru Katagiri, T. Umeda, Hisao Kuwabara |
| 1990 | Interspeech | On the unit search criteria and algorithms for speech synthesis using non-uniform units. | Kazuya Takeda, Katsuo Abe, Yoshinori Sagisaka |
| 1989 | ICASSP | Construction of a large-scale Japanese speech database and its management system. | Hisao Kuwabara, Kazuya Takeda, Yoshinori Sagisaka, Shigeru Katagiri, S. Morikawa, T. Watanabe |
| 1989 | Interspeech | Adaptive manipulation of non-uniform synthesis units using multi-level unit transcription. | Kazuya Takeda, Katsuo Abe, Yoshinori Sagisaka, Hisao Kuwabara |
| 1987 | Interspeech | Acoustic-phonetic labels in a Japanese speech database. | Kazuya Takeda, Yoshinori Sagisaka, Shigeru Katagiri |