| 2021 | ICASSP | Construction of a Large-Scale Japanese ASR Corpus on TV Recordings. | Shintaro Ando, Hiromasa Fujihara |
| 2014 | ICASSP | Cultivating vocal activity detection for music audio signals in a circulation-type crowdsourcing ecosystem. | Kazuyoshi Yoshii, Hiromasa Fujihara, Tomoyasu Nakano, Masataka Goto |
| 2012 | ICASSP | Instrumentation-based music similarity using sparse representations. | Hiromasa Fujihara, Anssi Klapuri, Mark D. Plumbley |
| 2012 | ITA | PodCastle and songle: Crowdsourcing-based web services for spoken document retrieval and active music listening. | Masataka Goto, Jun Ogata, Kazuyoshi Yoshii, Hiromasa Fujihara, Matthias Mauch, Tomoyasu Nakano |
| 2012 | WWW | PodCastle and Songle: Crowdsourcing-Based Web Services for Retrieval and Browsing of Speech and Music Content. | Masataka Goto, Jun Ogata, Kazuyoshi Yoshii, Hiromasa Fujihara, Matthias Mauch, Tomoyasu Nakano |
| 2011 | ICASSP | Concurrent estimation of singing voice F0 and phonemes by using spectral envelopes estimated from polyphonic music. | Hiromasa Fujihara, Masataka Goto |
| 2010 | ICASSP | Singing information processing based on singing voice modeling. | Masataka Goto, Takeshi Saitou, Tomoyasu Nakano, Hiromasa Fujihara |
| 2008 | ICASSP | Three techniques for improving automatic synchronization between music and lyrics: Fricative detection, filler model, and novel feature vectors for vocal activity detection. | Hiromasa Fujihara, Masataka Goto |
| 2006 | ICASSP | F0 Estimation Method for Singing Voice in Polyphonic Audio Signal Based on Statistical Vocal Model and Viterbi Search. | Hiromasa Fujihara, Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | Speaker identification under noisy environments by using harmonic structure extraction and reliable frame weighting. | Hiromasa Fujihara, Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ISM | Automatic Synchronization between Lyrics and Music CD Recordings Based on Viterbi Alignment of Segregated Vocal Signals. | Hiromasa Fujihara, Masataka Goto, Jun Ogata, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |