| 2022 | ICASSP | Entrainment Analysis for Assessment of Autistic Speech Prosody Using Bottleneck Features of Deep Neural Network. | Keiko Ochi, Nobutaka Ono, Keiho Owada, Miho Kuroda, Shigeki Sagayama, Hidenori Yamasue |
| 2022 | Interspeech | Use of Nods Less Synchronized with Turn-Taking and Prosody During Conversations in Adults with Autism. | Keiko Ochi, Nobutaka Ono, Keiho Owada, Miho Kuroda, Shigeki Sagayama, Hidenori Yamasue |
| 2019 | CISS | Piano Practice Evaluation and Visualization by HMM for Arbitrary Jumps and Mistakes. | Matsuto Hori, Christoph M. Wilk, Shigeki Sagayama |
| 2019 | CISS | Automatic Piano Reduction of Orchestral Music Based on Musical Entropy. | You Li, Christoph M. Wilk, Takeshi Hori, Shigeki Sagayama |
| 2019 | CISS | Autism Spectrum Disorder Discrimination Based on Voice Activities Related to Fillers and Laughter. | Daiki Mitsumoto, Takeshi Hori, Shigeki Sagayama, Hidenori Yamasue, Keiho Owada, Masaki Kojima, Keiko Ochi, Nobutaka Ono |
| 2013 | Interspeech | Probabilistic speech F | Tatsuma Ishihara, Hirokazu Kameoka, Kota Yoshizato, Daisuke Saito, Shigeki Sagayama |
| 2013 | Interspeech | Generative modeling of speech F | Hirokazu Kameoka, Kota Yoshizato, Tatsuma Ishihara, Yasunori Ohishi, Kunio Kashino, Shigeki Sagayama |
| 2012 | ICASSP | A tandem connectionist model using combination of multi-scale spectro-temporal features for acoustic event detection. | Miquel Espi, Masakiyo Fujimoto, Daisuke Saito, Nobutaka Ono, Shigeki Sagayama |
| 2012 | ICASSP | Constrained and regularized variants of non-negative matrix factorization incorporating music-specific constraints. | Hirokazu Kameoka, Masahiro Nakano, Kazuki Ochiai, Yutaka Imoto, Kunio Kashino, Shigeki Sagayama |
| 2012 | ICASSP | Explicit beat structure modeling for non-negative matrix factorization-based multipitch analysis. | Kazuki Ochiai, Hirokazu Kameoka, Shigeki Sagayama |
| 2012 | ICASSP | User-guided independent vector analysis with source activity tuning. | Takuma Ono, Nobutaka Ono, Shigeki Sagayama |
| 2012 | ICASSP | Comparative evaluations of various harmonic/percussive sound separation algorithms based on anisotropic continuity of spectrogram. | Hideyuki Tachibana, Hirokazu Kameoka, Nobutaka Ono, Shigeki Sagayama |
| 2012 | Interspeech | Speaker-Dependent Voice Activity Detection Robust to Background Speech Noise. | Shigeki Matsuda, Naoya Ito, Kosuke Tsujino, Hideki Kashioka, Shigeki Sagayama |
| 2012 | Interspeech | Hidden Markov Convolutive Mixture Model for Pitch Contour Analysis of Speech. | Kota Yoshizato, Hirokazu Kameoka, Daisuke Saito, Shigeki Sagayama |
| 2011 | CogSci | Templatic features for modeling phoneme acquisition. | Emmanuel Dupoux, Guillaume Beraud-Sudreau, Shigeki Sagayama |
| 2011 | ICASSP | Multichannel harmonic and percussive component separation by joint modeling of spatial and spectral continuity. | Ngoc Q. K. Duong, Hideyuki Tachibana, Emmanuel Vincent, Nobutaka Ono, Rmi Gribonval, Shigeki Sagayama |
| 2011 | ICASSP | Automatic video annotation via Hierarchical Topic Trajectory Model considering cross-modal correlations. | Takuho Nakano, Akisato Kimura, Hirokazu Kameoka, Shigeki Miyabe, Shigeki Sagayama, Nobutaka Ono, Kunio Kashino, Takuya Nishimoto |
| 2011 | ICASSP | Infinite-state spectrum model for music signal analysis. | Masahiro Nakano, Jonathan Le Roux, Hirokazu Kameoka, Nobutaka Ono, Shigeki Sagayama |
| 2011 | ICASSP | Multipitch estimation by joint modeling of harmonic and transient sounds. | Jun Wu, Emmanuel Vincent, Stanislaw Andrzej Raczynski, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2011 | ICDAR | Concurrent Optimization of Context Clustering and GMM for Offline Handwritten Word Recognition Using HMM. | Tomoyuki Hamamura, Bunpei Irie, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2011 | Interspeech | Using Spectral Fluctuation of Speech in Multi-Feature HMM-Based Voice Activity Detection. | Miquel Espi, Shigeki Miyabe, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2010 | ICASSP | Designing the Wiener post-filter for diffuse noise suppression using imaginary parts of inter-channel cross-spectra. | Nobutaka Ito, Nobutaka Ono, Emmanuel Vincent, Shigeki Sagayama |
| 2010 | ICASSP | A sparse component model of source signals and its application to blind source separation. | Yu Kitano, Hirokazu Kameoka, Yosuke Izumi, Nobutaka Ono, Shigeki Sagayama |
| 2010 | ICASSP | R-means localization: A simple iterative algorithm for range-difference-based source localization. | Nobutaka Ono, Shigeki Sagayama |
| 2010 | ICASSP | Melody line estimation in homophonic music audio signals based on temporal-variability of melodic source. | Hideyuki Tachibana, Takuma Ono, Nobutaka Ono, Shigeki Sagayama |
| 2010 | ICASSP | Music mood classification by rhythm and bass-line unit pattern analysis. | Emiru Tsunoo, Taichi Akase, Nobutaka Ono, Shigeki Sagayama |
| 2010 | ICASSP | HMM-based approach for automatic chord detection using refined acoustic features. | Yushi Ueda, Yuuki Uchiyama, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2010 | Interspeech | Musical instrument identification based on harmonic temporal timbre features. | Jun Wu, Yu Kitano, Stanislaw Andrzej Raczynski, Shigeki Miyabe, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2009 | ICASSP | Complex NMF: A new sparse representation for acoustic signals. | Hirokazu Kameoka, Nobutaka Ono, Kunio Kashino, Shigeki Sagayama |
| 2009 | ICASSP | Rhythm map: Extraction of unit rhythmic patterns and analysis of rhythmic structure from music acoustic signals. | Emiru Tsunoo, Nobutaka Ono, Shigeki Sagayama |
| 2009 | Interspeech | Stereo-input speech recognition using sparseness-based time-frequency masking in a reverberant environment. | Yosuke Izumi, Kenta Nishiki, Shinji Watanabe, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2008 | ICASSP | A blind noise decorrelation approach with crystal arrays on designing post-filters for diffuse noise suppression. | Nobutaka Ito, Nobutaka Ono, Shigeki Sagayama |
| 2008 | ICASSP | Auxiliary function approach to parameter estimation of constrained sinusoidal model for monaural speech separation. | Hirokazu Kameoka, Nobutaka Ono, Shigeki Sagayama |
| 2008 | ICASSP | Harmonic-Temporal-Timbral Clustering (HTTC) for the analysis of multi-instrument polyphonic music signals. | Kenichi Miyamoto, Hirokazu Kameoka, Takuya Nishimoto, Nobutaka Ono, Shigeki Sagayama |
| 2008 | ICASSP | Modulation analysis of speech through orthogonal FIR filterbank optimization. | Jonathan Le Roux, Hirokazu Kameoka, Nobutaka Ono, Shigeki Sagayama, Alain de Cheveign |
| 2008 | ICPR | On-line handwritten Kanji string recognition based on grammar description of character structures. | Ikumi Ota, Ryo Yamamoto, Takuya Nishimoto, Shigeki Sagayama |
| 2008 | Interspeech | Computational auditory induction by missing-data non-negative matrix factorization. | Jonathan Le Roux, Hirokazu Kameoka, Nobutaka Ono, Alain de Cheveign, Shigeki Sagayama |
| 2008 | Interspeech | Explicit consistency constraints for STFT spectrograms and their application to phase reconstruction. | Jonathan Le Roux, Nobutaka Ono, Shigeki Sagayama |
| 2007 | ICASSP | Probabilistic Approach to Automatic Music Transcription from Audio Signals. | Kenichi Miyamoto, Hirokazu Kameoka, Haruto Takeda, Takuya Nishimoto, Shigeki Sagayama |
| 2007 | ICASSP | Harmonic-Temporal Clustering of Speech for Single and Multiple F0 Contour Estimation in Noisy Environments. | Jonathan Le Roux, Hirokazu Kameoka, Nobutaka Ono, Alain de Cheveign, Shigeki Sagayama |
| 2007 | ICASSP | Rhythm and Tempo Analysis Toward Automatic Music Transcription. | Haruto Takeda, Takuya Nishimoto, Shigeki Sagayama |
| 2007 | ICDAR | Online Handwritten Kanji Recognition Based on Inter-stroke Grammar. | Ikumi Ota, Ryo Yamamoto, Shinji Sako, Shigeki Sagayama |
| 2007 | IJCAI | Automatic Decision of Piano Fingering Based on a Hidden Markov Models. | Yuichiro Yonebayashi, Hirokazu Kameoka, Shigeki Sagayama |
| 2007 | MMSP | Sound Source Localization by Asymmetrically Arrayed 2ch Microphones on a Sphere. | Nobutaka Ono, Souichiro Fukamachi, Takuya Nishimoto, Shigeki Sagayama |
| 2006 | ICASSP | Model Adaptation for Long Convolutional Distortion by Maximum Likelihood Based State Filtering Approach. | Chandra Kant Raut, Takuya Nishimoto, Shigeki Sagayama |
| 2006 | Interspeech | Speech analyzer using a joint estimation model of spectral envelope and fine structure. | Hirokazu Kameoka, Jonathan Le Roux, Nobutaka Ono, Shigeki Sagayama |
| 2005 | ICASSP | Audio stream segregation of multi-pitch music signal based on time-space clustering using Gaussian kernel 2-dimensional model. | Hirokazu Kameoka, Takuya Nishimoto, Shigeki Sagayama |
| 2005 | Interspeech | Model adaptation by state splitting of HMM for long reverberation. | Chandra Kant Raut, Takuya Nishimoto, Shigeki Sagayama |
| 2004 | ICASSP | Separation of harmonic structures based on tied Gaussian mixture model and information criterion for concurrent sounds. | Hirokazu Kameoka, Takuya Nishimoto, Shigeki Sagayama |
| 2004 | Interspeech | Multi-pitch trajectory estimation of concurrent speech based on harmonic GMM and nonlinear kalman filtering. | Takuya Nishimoto, Shigeki Sagayama, Hirokazu Kameoka |
| 2004 | Interspeech | Model composition by lagrange polynomial approximation for robust speech recognition in noisy environment. | Chandra Kant Raut, Takuya Nishimoto, Shigeki Sagayama |
| 2004 | Interspeech | Specmurt anasylis: a piano-roll-visualization of polyphonic music signal by deconvolution of log-frequency spectrum. | Shigeki Sagayama, Keigo Takahashi, Hirokazu Kameoka, Takuya Nishimoto |
| 2004 | Interspeech | Complex spectrum circle centroid for microphone-array-based noisy speech recognition. | Shigeki Sagayama, Okajima Takashi, Yutaka Kamamoto, Takuya Nishimoto |
| 2003 | ICDAR | Generation of Hierarchical Dictionary for Stroke-order Free Kanji Handwriting Recognition Based on Substroke HMM. | Mitsuru Nakai, Hiroshi Shimodaira, Shigeki Sagayama |
| 2003 | ICDAR | On-line Overlaid-Handwriting Recognition Based on Substroke HMMs. | Hiroshi Shimodaira, Takashi Sudo, Mitsuru Nakai, Shigeki Sagayama |
| 2002 | ICASSP | Jacobian joint adaptation to noise, channel and vocal tract length. | Hiroshi Shimodaira, Nobuyoshi Sakai, Mitsuru Nakai, Shigeki Sagayama |
| 2002 | ICFHR | Context-dependent substroke model for HMM-based on-line handwriting recognition. | Junko Tokuno, Nobuhito Inami, Shigeki Matsuda, Mitsuru Nakai, Hiroshi Shimodaira, Shigeki Sagayama |
| 2002 | ICPR | Pen Pressure Features for Writer-Independent On-Line Handwriting Recognition Based on Substroke HMM. | Mitsuru Nakai, Takashi Sudo, Hiroshi Shimodaira, Shigeki Sagayama |
| 2001 | ICASSP | Multiple-regression hidden Markov model. | Katsuhisa Fujinaga, Mitsuru Nakai, Hiroshi Shimodaira, Shigeki Sagayama |
| 2001 | ICDAR | Substroke Approach to HMM-Based On-line Kanji Handwriting Recognition. | Mitsuru Nakai, Naoto Akira, Hiroshi Shimodaira, Shigeki Sagayama |
| 2001 | Interspeech | Support vector machine with dynamic time-alignment kernel for speech recognition. | Hiroshi Shimodaira, Ken-ichi Noma, Mitsuru Nakai, Shigeki Sagayama |
| 2000 | ICASSP | Asynchronous-transition HMM. | Shigeki Matsuda, Mitsuru Nakai, Hiroshi Shimodaira, Shigeki Sagayama |
| 2000 | Interspeech | Free software toolkit for Japanese large vocabulary continuous speech recognition. | Tatsuya Kawahara, Akinobu Lee, Tetsunori Kobayashi, Kazuya Takeda, Nobuaki Minematsu, Shigeki Sagayama, Katsunobu Itou, Akinori Ito, Mikio Yamamoto, Atsushi Yamada, Takehito Utsuro, Kiyohiro Shikano |
| 2000 | Interspeech | Feature-dependent allophone clustering. | Shigeki Matsuda, Mitsuru Nakai, Hiroshi Shimodaira, Shigeki Sagayama |
| 2000 | Interspeech | Jacobian adaptation of HMM with initial model selection for noisy speech recognition. | Hiroshi Shimodaira, Yutaka Kato, Toshihiko Akae, Mitsuru Nakai, Shigeki Sagayama |
| 2000 | LREC | IPA Japanese Dictation Free Software Project. | Katsunobu Itou, Kiyohiro Shikano, Tatsuya Kawahara, Kazuya Takeda, Atsushi Yamada, Akinori Ito, Takehito Utsuro, Tetsunori Kobayashi, Nobuaki Minematsu, Mikio Yamamoto, Shigeki Sagayama, Akinobu Lee |
| 1998 | ICASSP | Two-step generation of variable-word-length language model integrating local and global constraints. | Shoichi Matsunaga, Shigeki Sagayama |
| 1997 | ICASSP | Improved estimation of supervision in unsupervised speaker adaptation. | Shigeru Homma, Kiyoaki Aikawa, Shigeki Sagayama |
| 1997 | ICASSP | Jacobian approach to fast acoustic model adaptation. | Shigeki Sagayama, Yoshikazu Yamaguchi, Satoshi Takahashi, Jun-ichi Takahashi |
| 1997 | ICASSP | Discrete mixture HMM. | Satoshi Takahashi, Kiyoaki Aikawa, Shigeki Sagayama |
| 1997 | Interspeech | Variable-length language modeling integrating global constraints. | Shoichi Matsunaga, Shigeki Sagayama |
| 1997 | Interspeech | Fast adaptation of acoustic models to environmental noise using jacobian adaptation algorithm. | Yoshikazu Yamaguchi, Satoshi Takahashi, Shigeki Sagayama |
| 1996 | ICASSP | Tied-structure HMM based on parameter correlation for efficient model training. | Satoshi Takahashi, Shigeki Sagayama |
| 1996 | ICASSP | Minimum classification error training for a small amount of data enhanced by vector-field-smoothed Bayesian learning. | Jun-ichi Takahashi, Shigeki Sagayama |
| 1996 | Interspeech | Iterative unsupervised speaker adaptation for batch dictation. | Shigeru Homma, Jun-ichi Takahashi, Shigeki Sagayama |
| 1996 | Interspeech | LR-parser-driven viterbi search with hypotheses merging mechanism using context-dependent phone models. | Tomokazu Yamada, Shigeki Sagayama |
| 1995 | ICASSP | On the use of scalar quantization for fast HMM computation. | Shigeki Sagayama, Satoshi Takahashi |
| 1995 | ICASSP | Four-level tied-structure for efficient representation of acoustic modeling. | Satoshi Takahashi, Shigeki Sagayama |
| 1995 | ICASSP | Vector-field-smoothed Bayesian learning for incremental speaker adaptation. | Jun-ichi Takahashi, Shigeki Sagayama |
| 1995 | Interspeech | Syllabic duration control for vocabulary-free speech recognition. | Takatoshi Jitsuhiro, Tomokazu Yamada, Shigeki Sagayama |
| 1995 | Interspeech | Fast and accurate beam search using forward heuristic functions in HMM-LR speech recognition. | Yoshiaki Noda, Shigeki Sagayama |
| 1994 | ICASSP | Tree-structured speaker clustering for fast speaker adaptation. | Tetsuo Kosaka, Shigeki Sagayama |
| 1994 | ICASSP | All-phoneme ergodic hidden Markov network for unsupervised speaker adaptation. | Yasunaga Miyazawa, Jun-ichi Takami, Shigeki Sagayama, Shoichi Matsunaga |
| 1994 | Interspeech | Tree-structured speaker clustering for speaker-independent continuous speech recognition. | Tetsuo Kosaka, Shoichi Matsunaga, Shigeki Sagayama |
| 1994 | Interspeech | Telephone line characteristic adaptation using vector field smoothing technique. | Jun-ichi Takahashi, Shigeki Sagayama |
| 1994 | Interspeech | Speaker-consistent parsing for speaker-independent continuous speech recognition. | Kouichi Yamaguchi, Harald Singer, Shoichi Matsunaga, Shigeki Sagayama |
| 1993 | ICASSP | Rapid speaker adaptation using speaker-mixture allophone models applied to speaker-independent speech recognition. | Tetsuo Kosaka, Jun-ichi Takami, Shigeki Sagayama |
| 1993 | ICASSP | ATREUS: a comparative study of continuous speech recognition systems at ATR. | Akito Nagai, Kouichi Yamaguchi, Shigeki Sagayama, Akira Kurematsu |
| 1993 | ICASSP | Matrix parser and its application to HMM-based speech recognition. | Harald Singer, Shigeki Sagayama |
| 1993 | IJCAI | Spoken Language Translation System. | Gen-ichiro Kikui, Mark Seligman, Toshiyuki Takezawa, Masami Suzuki, Kenji Kita, Tsuyoshi Morimoto, Masaaki Nagata, Toshihisa Tashiro, Herbert S. Tropf, Shigeki Sagayama, Jun-ichi Takami, Kazumi Ohkura, Akira Kurematsu |
| 1993 | Interspeech | Speech recognition using particle n-grams and content-word n-grams. | Ryosuke Isotani, Shigeki Sagayama |
| 1993 | Interspeech | A dynamic approach to speaker adaptation of hidden Markov networks for speech recognition. | Tetsuo Kosaka, Edward Willems, Jun-ichi Takami, Shigeki Sagayama |
| 1993 | Interspeech | ATR's speech translation system: ASURA. | Tsuyoshi Morimoto, Toshiyuki Takezawa, Fumihiro Yato, Shigeki Sagayama, Toshihisa Tashiro, Masaaki Nagata, Akira Kurematsu |
| 1993 | Interspeech | The possibility for acquisition of statistical network grammar using ergodic HMM. | Jin'ichi Murakami, Hiroki Yamatomo, Shigeki Sagayama |
| 1993 | Interspeech | ATREUS: a speech recognition front-end for a speech translation system. | Shigeki Sagayama, Jun-ichi Takami, Akito Nagai, Harald Singer, Kouichi Yamaguchi, Kazumi Ohkura, Kenji Kita, Akira Kurematsu |
| 1992 | ICASSP | Pitch dependent phone modelling for HMM based speech recognition. | Harald Singer, Shigeki Sagayama |
| 1992 | ICASSP | A successive state splitting algorithm for efficient allophone modeling. | Jun-ichi Takami, Shigeki Sagayama |
| 1992 | Interspeech | Vector field smoothing principle for speaker adaptation. | Hiroaki Hattori, Shigeki Sagayama |
| 1992 | Interspeech | Continuously spoken sentence recognition by HMM-LR. | Kenji Kita, Tsuyoshi Morimoto, Kazumi Ohkura, Shigeki Sagayama |
| 1992 | Interspeech | Enhancement of ATR's spoken language translation system: SL-TRANS2. | Tsuyoshi Morimoto, Toshiyuki Takezawa, Kazumi Ohkura, Masaaki Nagata, Fumihiro Yato, Shigeki Sagayama, Akira Kurematsu |
| 1992 | Interspeech | Hardware implementation of realtime 1000-word HMM-LR continuous speech recognition. | Akito Nagai, Kenji Kita, Toshiyuki Hanazawa, Tadashi Suzuki, Tomohiro Iwasaki, Tsuyoshi Kawabata, Kunio Nakajima, Kiyohiro Shikano, Tsuyoshi Morimoto, Shigeki Sagayama, Akira Kurematsu |
| 1992 | Interspeech | The SSS-LR continuous speech recognition system: integrating SSS-derived allophone models and a phoneme-context-dependent LR parser. | Akito Nagai, Jun-ichi Takami, Shigeki Sagayama |
| 1992 | Interspeech | Speaker adaptation based on transfer vector field smoothing with continuous mixture density HMMs. | Kazumi Ohkura, Masahide Sugiyama, Shigeki Sagayama |
| 1992 | Interspeech | Appropriate error criterion selection for continuous speech HMM minimum error training. | David Rainion, Shigeki Sagayama |
| 1992 | Interspeech | Continuous mixture HMM-LR using the a* algorithm for continuous speech recognition. | Kouichi Yamaguchi, Shigeki Sagayama, Kenji Kita, Frank K. Soong |
| 1991 | ICASSP | Phoneme recognition by phoneme filter neural networks. | Masami Nakamura, Shinichi Tamura, Shigeki Sagayama |
| 1991 | ICASSP | A pairwise discriminant approach to robust phoneme recognition by time-delay neural networks. | Jun-ichi Takami, Shigeki Sagayama |
| 1991 | Interspeech | Phoneme-context-dependent LR parsing algorithms for HMM-based continuous speech recognition. | Akito Nagai, Shigeki Sagayama, Kenji Kita |
| 1991 | Interspeech | A matrix representation of HMM-based speech recognition algorithms. | Shigeki Sagayama |
| 1990 | ICASSP | A continuous speech recognition system based on a two-level grammar approach. | Shoichi Matsunaga, Shigeki Sagayama, Shigeru Homma, Sadaoki Furui |
| 1990 | Interspeech | Statistical study on voice individuality conversion across different languages. | Masanobu Abe, Shigeki Sagayama |
| 1990 | Interspeech | Line spectrum pair frequency - based distance measures for speech recognition. | Fikret S. Grgen, Shigeki Sagayama, Sadaoki Furui |
| 1990 | Interspeech | Speaker weighted training of HMM using multiple reference speakers. | Hiroaki Hattori, Satoshi Nakamura, Kiyohiro Shikano, Shigeki Sagayama |
| 1990 | Interspeech | Sentence speech recognition using semantic dependency analysis. | Shoichi Matsunaga, Shigeki Sagayama |
| 1990 | Interspeech | Estimation of unknown context using a phoneme environment clustering algorithm. | Shigeki Sagayama, Shigeru Honrna |
| 1990 | Interspeech | Isolated word recognition using pitch pattern information. | Satoshi Takahashi, Shoichi Matsunaga, Shigeki Sagayama |
| 1990 | Interspeech | Phoneme recognition by pairwise discriminant TDNNs. | Jun-ichi Takami, Shigeki Sagayama |
| 1989 | ICASSP | Phoneme environment clustering for speech recognition. | Shigeki Sagayama |
| 1986 | ICASSP | Duality theory of composite sinusoidal modeling and linear prediction. | Shigeki Sagayama, Fumitada Itakura |