| 2026 | MMM | A Case Study of a Transparent and Controllable Music Recommender System with Multi-relational Layers. | Kosetsu Tsukuda, Keisuke Ishida, Takumi Takahashi, Masahiro Hamasaki, Masataka Goto |
| 2025 | IJCAI | Constrained Preferential Bayesian Optimization and Its Application in Banner Ad Design. | Koki Iwai, Yusuke Kumagae, Yuki Koyama, Masahiro Hamasaki, Masataka Goto |
| 2025 | MMM | Kiite World: Socializing Map-Based Music Exploration Through Playlist Sharing and Synchronized Listening. | Kosetsu Tsukuda, Takumi Takahashi, Keisuke Ishida, Masahiro Hamasaki, Masataka Goto |
| 2025 | NAACL | A Data-Driven Method for Analyzing and Quantifying Lyrics-Dance Motion Relationships. | Kento Watanabe, Masataka Goto |
| 2023 | CHI | Lyric App Framework: A Web-based Framework for Developing Interactive Lyric-driven Musical Applications. | Jun Kato, Masataka Goto |
| 2023 | CHI | CatAlyst: Domain-Extensible Intervention for Preventing Task Procrastination Using Large Generative Models. | Riku Arakawa, Hiromu Yakura, Masataka Goto |
| 2023 | ICASSP | U-Beat: A Multi-Scale Beat Tracking Model Based on Wave-U-Net. | Tian Cheng, Masataka Goto |
| 2023 | WACV | Content-Based Music-Image Retrieval Using Self- and Cross-Modal Feature Embedding Memory. | Takayuki Nakatsuka, Masahiro Hamasaki, Masataka Goto |
| 2022 | IUI | BeParrot: Efficient Interface for Transcribing Unclear Speech via Respeaking. | Riku Arakawa, Hiromu Yakura, Masataka Goto |
| 2022 | UIST | BO as Assistant: Using Bayesian Optimization for Asynchronously Generating Design Suggestions. | Yuki Koyama, Masataka Goto |
| 2021 | IJCAI | Tool- and Domain-Agnostic Parameterization of Style Transfer Effects Leveraging Pretrained Perceptual Metrics. | Hiromu Yakura, Yuki Koyama, Masataka Goto |
| 2021 | IUI | Interactive Exploration-Exploitation Balancing for Generative Melody Composition. | Yijun Zhou, Yuki Koyama, Masataka Goto, Takeo Igarashi |
| 2021 | MMM | Atypical Lyrics Completion Considering Musical Audio Signals. | Kento Watanabe, Masataka Goto |
| 2020 | DAS | Lyric Video Analysis Using Text Detection and Tracking. | Shota Sakaguchi, Jun Kato, Masataka Goto, Seiichi Uchida |
| 2020 | ISMAR | Enhancing Participation Experience in VR Live Concerts by Improving Motions of Virtual Audience Avatars. | Hiromu Yakura, Masataka Goto |
| 2020 | IUI | Interactive deep singing-voice separation based on human-in-the-loop adaptation. | Tomoyasu Nakano, Yuki Koyama, Masahiro Hamasaki, Masataka Goto |
| 2020 | RecSys | Explainable Recommendation for Repeat Consumption. | Kosetsu Tsukuda, Masataka Goto |
| 2020 | SIGIR | Query/Task Satisfaction and Grid-based Evaluation Metrics Under Different Image Search Intents. | Kosetsu Tsukuda, Masataka Goto |
| 2019 | ICASSP | Zero-mean Convolutional Network with Data Augmentation for Sound Level Invariant Singing Voice Separation. | Kin Wah Edward Lin, Masataka Goto |
| 2019 | ICASSP | Automatic Singing Transcription Based on Encoder-decoder Recurrent Neural Networks with a Weakly-supervised Attention Mechanism. | Ryo Nishikimi, Eita Nakamura, Satoru Fukayama, Masataka Goto, Kazuyoshi Yoshii |
| 2019 | ICASSP | Transdrums: A Drum Pattern Transfer System Preserving Global Pattern Structure. | Shun Sawada, Satoru Fukayama, Masataka Goto, Keiji Hirata |
| 2019 | ICASSP | Joint Transcription of Lead, Bass, and Rhythm Guitars Based on a Factorial Hidden Semi-Markov Model. | Kentaro Shibata, Ryo Nishikimi, Satoru Fukayama, Masataka Goto, Eita Nakamura, Katsutoshi Itoyama, Kazuyoshi Yoshii |
| 2019 | IUI | Autocomplete vocal- | Tomoyasu Nakano, Yuki Koyama, Masahiro Hamasaki, Masataka Goto |
| 2019 | MMM | Audio-Based Automatic Generation of a Piano Reduction Score by Considering the Musical Structure. | Hirofumi Takamori, Takayuki Nakatsuka, Satoru Fukayama, Masataka Goto, Shigeo Morishima |
| 2019 | MMM | Query-by-Dancing: A Dance Music Retrieval System Based on Body-Motion Similarity. | Shuhei Tsuchida, Satoru Fukayama, Masataka Goto |
| 2019 | RecSys | DualDiv: diversifying items and explanation styles in explainable hybrid recommendation. | Kosetsu Tsukuda, Masataka Goto |
| 2019 | SIGIR | ABCPRec: Adaptively Bridging Consumer and Producer Roles for User-Generated Content Recommendation. | Kosetsu Tsukuda, Satoru Fukayama, Masataka Goto |
| 2018 | CHI | OptiMo: Optimization-Guided Motion Editing for Keyframe Character Animation. | Yuki Koyama, Masataka Goto |
| 2018 | ICASSP | Music Structure Boundary Detection and Labelling by a Deconvolution of Path-Enhanced Self-Similarity Matrix. | Tian Cheng, Jordan B. L. Smith, Masataka Goto |
| 2018 | ICASSP | Retrieval of Song Lyrics from Sung Queries. | Anna M. Kruspe, Masataka Goto |
| 2018 | ICASSP | Instlistener: An Expressive Parameter Estimation System Imitating Human Performances of Monophonic Musical Instruments. | Zhengshan Shi, Tomoyasu Nakano, Masataka Goto |
| 2018 | ICASSP | Nonnegative Tensor Factorization for Source Separation of Loops in Audio. | Jordan B. L. Smith, Masataka Goto |
| 2018 | ICWSM | Collaboration in N-th Order Derivative Creation. | Shiori Hironaka, Kosetsu Tsukuda, Masahiro Hamasaki, Masataka Goto |
| 2018 | IUI | Intelligent Music Interfaces. | Masataka Goto |
| 2018 | IUI | FocusMusicRecommender: A System for Recommending Music to Listen to While Working. | Hiromu Yakura, Tomoyasu Nakano, Masataka Goto |
| 2018 | NAACL | A Melody-Conditioned Lyrics Language Model. | Kento Watanabe, Yuichiroh Matsubayashi, Satoru Fukayama, Masataka Goto, Kentaro Inui, Tomoyasu Nakano |
| 2017 | HRI | A Robotic Framework for Video Recording and Authoring. | Jun Kato, Masataka Goto |
| 2017 | IUI | LyriSys: An Interactive Support System for Writing Lyrics Based on Topic Transition. | Kento Watanabe, Yuichiroh Matsubayashi, Kentaro Inui, Tomoyasu Nakano, Satoru Fukayama, Masataka Goto |
| 2017 | PAKDD | Taste or Addiction?: Using Play Logs to Infer Song Selection Motivation. | Kosetsu Tsukuda, Masataka Goto |
| 2016 | CIKM | Why Did You Cover That Song?: Modeling N-th Order Derivative Creation with Content Popularity. | Kosetsu Tsukuda, Masahiro Hamasaki, Masataka Goto |
| 2016 | COLING | Modeling Discourse Segments in Lyrics Using Repeated Patterns. | Kento Watanabe, Yuichiroh Matsubayashi, Naho Orita, Naoaki Okazaki, Kentaro Inui, Satoru Fukayama, Tomoyasu Nakano, Jordan B. L. Smith, Masataka Goto |
| 2016 | ICASSP | Music emotion recognition with adaptive aggregation of Gaussian process regressors. | Satoru Fukayama, Masataka Goto |
| 2016 | ICASSP | An estimation method of voice timbre evaluation values using feature extraction with Gaussian mixture model based on reference singer. | Soichi Yamane, Kazuhiro Kobayashi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Satoshi Nakamura |
| 2016 | ICASSP | Student's T nonnegative matrix factorization and positive semidefinite tensor factorization for single-channel audio source separation. | Kazuyoshi Yoshii, Katsutoshi Itoyama, Masataka Goto |
| 2016 | ICDM | SmartVideoRanking: Video Search by Mining Emotions from Time-Synchronized Comments. | Kosetsu Tsukuda, Masahiro Hamasaki, Masataka Goto |
| 2016 | IUI | PlaylistPlayer: An Interface Using Multiple Criteria to Change the Playback Order of a Music Playlist. | Tomoyasu Nakano, Jun Kato, Masahiro Hamasaki, Masataka Goto |
| 2016 | SAC | Plate: persistent memory management for nonvolatile main memory. | Toshihiro Yamauchi, Yuta Yamamoto, Kengo Nagai, Tsukasa Matono, Shinji Inamoto, Masaya Ichikawa, Masataka Goto, Hideo Taniguchi |
| 2016 | SCA | A choreographic authoring system for character dance animation reflecting a user's preference. | Ryo Kakitsuka, Kosetsu Tsukuda, Satoru Fukayama, Naoya Iwamoto, Masataka Goto, Shigeo Morishima |
| 2015 | CHI | TextAlive: Integrated Design Environment for Kinetic Typography. | Jun Kato, Tomoyasu Nakano, Masataka Goto |
| 2015 | ICASSP | A feedback framework for improved chord recognition based on NMF-based approximate note transcription. | Satoshi Maruo, Kazuyoshi Yoshii, Katsutoshi Itoyama, Matthias Mauch, Masataka Goto |
| 2015 | ISM | Songle Widget: Making Animation and Physical Devices Synchronized with Music Videos on the Web. | Masataka Goto, Kazuyoshi Yoshii, Tomoyasu Nakano |
| 2015 | ISM | Musical Similarity and Commonness Estimation Based on Probabilistic Generative Models. | Tomoyasu Nakano, Kazuyoshi Yoshii, Masataka Goto |
| 2015 | ISM | ExploratoryVideoSearch: A Music Video Search System Based on Coordinate Terms and Diversification. | Kosetsu Tsukuda, Masataka Goto |
| 2015 | SIGGRAPH | A music video authoring system synchronizing climax of video clips and music via rearrangement of musical bars. | Haruki Sato, Tatsunori Hirai, Tomoyasu Nakano, Masataka Goto, Shigeo Morishima |
| 2015 | UIST | Form Follows Function(): An IDE to Create Laser-cut Interfaces and Microcontroller Programs from Single Code Base. | Jun Kato, Masataka Goto |
| 2014 | HAI | Sharedo: to-do list interface for human-agent task sharing. | Jun Kato, Daisuke Sakamoto, Takeo Igarashi, Masataka Goto |
| 2014 | ICASSP | Regression approaches to perceptual age control in singing voice conversion. | Kazuhiro Kobayashi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2014 | ICASSP | Leveraging repetition for improved automatic lyric transcription in popular music. | Matt McVicar, Daniel P. W. Ellis, Masataka Goto |
| 2014 | ICASSP | Timbre replacement of harmonic and drum components for music audio signals. | Tomohiko Nakamura, Hirokazu Kameoka, Kazuyoshi Yoshii, Masataka Goto |
| 2014 | ICASSP | Vocal timbre analysis using latent Dirichlet allocation and cross-gender vocal timbre similarity. | Tomoyasu Nakano, Kazuyoshi Yoshii, Masataka Goto |
| 2014 | ICASSP | Cultivating vocal activity detection for music audio signals in a circulation-type crowdsourcing ecosystem. | Kazuyoshi Yoshii, Hiromasa Fujihara, Tomoyasu Nakano, Masataka Goto |
| 2014 | PACLIC | Modeling Structural Topic Transitions for Automatic Lyrics Generation. | Kento Watanabe, Yuichiroh Matsubayashi, Kentaro Inui, Masataka Goto |
| 2014 | WWW | Songrium: a music browsing assistance service with interactive visualization and exploration of protect a web of music. | Masahiro Hamasaki, Masataka Goto, Tomoyasu Nakano |
| 2013 | ICASSP | Infinite kernel linear prediction for joint estimation of spectral envelope and fundamental frequency. | Kazuyoshi Yoshii, Masataka Goto |
| 2013 | ICML | Infinite Positive Semidefinite Tensor Factorization for Source Separation of Mixture Signals. | Kazuyoshi Yoshii, Ryota Tomioka, Daichi Mochihashi, Masataka Goto |
| 2013 | Interspeech | Evaluation of a singing voice conversion method based on many-to-many eigenvoice conversion. | Hironori Doi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Satoshi Nakamura |
| 2013 | Interspeech | An investigation of acoustic features for singing voice conversion based on perceptual age. | Kazuhiro Kobayashi, Hironori Doi, Tomoki Toda, Tomoyasu Nakano, Masataka Goto, Graham Neubig, Sakriani Sakti, Satoshi Nakamura |
| 2012 | ICASSP | VocaListener and VocaWatcher: Imitating a human singer by using signal processing. | Masataka Goto, Tomoyasu Nakano, Shuuji Kajita, Yosuke Matsusaka, Shinichiro Nakaoka, Kazuhito Yokoi |
| 2012 | ICASSP | Unsupervised music understanding based on nonparametric Bayesian models. | Kazuyoshi Yoshii, Masataka Goto |
| 2012 | Interspeech | A spectral envelope estimation method based on F0-adaptive multi-frame integration analysis. | Tomoyasu Nakano, Masataka Goto |
| 2012 | Interspeech | PodCastle: Collaborative Training of Language Models on the Basis of Wisdom of Crowds. | Jun Ogata, Masataka Goto |
| 2012 | ITA | PodCastle and songle: Crowdsourcing-based web services for spoken document retrieval and active music listening. | Masataka Goto, Jun Ogata, Kazuyoshi Yoshii, Hiromasa Fujihara, Matthias Mauch, Tomoyasu Nakano |
| 2012 | WWW | PodCastle and Songle: Crowdsourcing-Based Web Services for Retrieval and Browsing of Speech and Music Content. | Masataka Goto, Jun Ogata, Kazuyoshi Yoshii, Hiromasa Fujihara, Matthias Mauch, Tomoyasu Nakano |
| 2011 | CSCW | Social Infobox: collaborative knowledge construction by social property tagging. | Masahiro Hamasaki, Masataka Goto, Hideaki Takeda |
| 2011 | ICASSP | Concurrent estimation of singing voice F0 and phonemes by using spectral envelopes estimated from polyphonic music. | Hiromasa Fujihara, Masataka Goto |
| 2011 | ICASSP | Simultaneous processing of sound source separation and musical instrument identification using Bayesian spectral modeling. | Katsutoshi Itoyama, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2011 | ICASSP | Polyphonic audio-to-score alignment based on Bayesian Latent Harmonic Allocation Hidden Markov Model. | Akira Maezawa, Hiroshi G. Okuno, Tetsuya Ogata, Masataka Goto |
| 2011 | ICASSP | Vocalistener2: A singing synthesis system able to mimic a user's singing in terms of voice timbre changes as well as pitch and dynamics. | Tomoyasu Nakano, Masataka Goto |
| 2011 | Interspeech | PodCastle: Recent Advances of a Spoken Document Retrieval Service Improved by Anonymous User Contributions. | Masataka Goto, Jun Ogata |
| 2011 | IROS | VocaWatcher: Natural singing motion generator for a humanoid robot. | Shuuji Kajita, Tomoyasu Nakano, Masataka Goto, Yosuke Matsusaka, Shinichiro Nakaoka, Kazuhito Yokoi |
| 2010 | ICASSP | Singing information processing based on singing voice modeling. | Masataka Goto, Takeshi Saitou, Tomoyasu Nakano, Hiromasa Fujihara |
| 2010 | PACLIC | PodCastle: A Spoken Document Retrieval Service Improved by Anonymous User Contributions. | Masataka Goto, Jun Ogata |
| 2009 | ICASSP | The use of acoustically detected filled and silent pauses in spontaneous speech recognition. | Jun Ogata, Masataka Goto, Katunobu Itou |
| 2009 | Interspeech | Podcastle: collaborative training of acoustic models on the basis of wisdom of crowds for podcast transcription. | Jun Ogata, Masataka Goto |
| 2009 | Interspeech | Acoustic and perceptual effects of vocal training in amateur male singing. | Takeshi Saitou, Masataka Goto |
| 2009 | Interspeech | Acoustic event detection for spotting "hot spots" in podcasts. | Kouhei Sumi, Tatsuya Kawahara, Jun Ogata, Masataka Goto |
| 2008 | ICASSP | Three techniques for improving automatic synchronization between music and lyrics: Fricative detection, filler model, and novel feature vectors for vocal activity detection. | Hiromasa Fujihara, Masataka Goto |
| 2007 | ICASSP | Active Music Listening Interfaces Based on Signal Processing. | Masataka Goto |
| 2007 | ICASSP | Integration and Adaptation of Harmonic and Inharmonic Models for Separating Polyphonic Musical Signals. | Katsutoshi Itoyama, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2007 | ICMI | Presentation sensei: a presentation training system using speech and image processing. | Kazutaka Kurihara, Masataka Goto, Jun Ogata, Yosuke Matsusaka, Takeo Igarashi |
| 2007 | Interspeech | Podcastle: a web 2.0 approach to speech recognition research. | Masataka Goto, Jun Ogata, Kouichirou Eto |
| 2007 | Interspeech | Automatic transcription for a web 2.0 service to search podcasts. | Jun Ogata, Masataka Goto, Kouichirou Eto |
| 2007 | Interspeech | Vocal conversion from speaking voice to singing voice using STRAIGHT. | Takeshi Saitou, Masataka Goto, Masashi Unoki, Masato Akagi |
| 2006 | CHI | Speech pen: predictive handwriting based on ambient multimodal recognition. | Kazutaka Kurihara, Masataka Goto, Jun Ogata, Takeo Igarashi |
| 2006 | ICASSP | F0 Estimation Method for Singing Voice in Polyphonic Audio Signal Based on Statistical Vocal Model and Viterbi Search. | Hiromasa Fujihara, Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ICASSP | Instrogram: A New Musical Instrument Recognition Technique Without Using Onset Detection NOR F0 Estimation. | Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ICASSP | An Error Correction Framework Based on Drum Pattern Periodicity for Improving Drum Sound Detection. | Kazuyoshi Yoshii, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | Speaker identification under noisy environments by using harmonic structure extraction and reliable frame weighting. | Hiromasa Fujihara, Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | Interspeech | An automatic singing skill evaluation method for unknown melodies using pitch interval accuracy and vibrato features. | Tomoyasu Nakano, Masataka Goto, Yuzuru Hiraga |
| 2006 | ISM | Automatic Synchronization between Lyrics and Music CD Recordings Based on Viterbi Alignment of Segregated Vocal Signals. | Hiromasa Fujihara, Masataka Goto, Jun Ogata, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2006 | ISM | Musical Instrument Recognizer "Instrogram" and Its Application to Music Retrieval Based on Instrumentation Similarity. | Tetsuro Kitahara, Masataka Goto, Kazunori Komatani, Tetsuya Ogata, Hiroshi G. Okuno |
| 2005 | ICASSP | An Auto-Regressive, Non-Stationary Excited Signal Parameter Estimation Method and an Evaluation of a Singing-Voice Recognition. | Akira Sasou, Masataka Goto, Satoru Hayamizu, Kazuyo Tanaka |
| 2005 | Interspeech | Speech repair: quick error correction just by using selection operation for speech input interfaces. | Jun Ogata, Masataka Goto |
| 2005 | Interspeech | Discrimination between singing and speaking voices. | Yasunori Ohishi, Masataka Goto, Katunobu Itou, Kazuya Takeda |
| 2005 | WiMob | A wireless LAN architecture using PANA for secure network selection. | Yoshimichi Tanizawa, Masataka Goto, Victor Fajardo, Yoshihiro Ohba |
| 2004 | ICASSP | Category-level identification of non-registered musical instrument sounds. | Tetsuro Kitahara, Masataka Goto, Hiroshi G. Okuno |
| 2004 | Interspeech | Speech spotter: on-demand speech recognition in human-human conversation on the telephone or in face-to-face situations. | Masataka Goto, Koji Kitayama, Katsunobu Itou, Tetsunori Kobayashi |
| 2004 | Interspeech | Drum sound identification for polyphonic music using template adaptation and matching methods. | Kazuyoshi Yoshii, Masataka Goto, Hiroshi G. Okuno |
| 2003 | ICASSP | A chorus-section detecting method for musical audio signals. | Masataka Goto |
| 2003 | ICASSP | Musical instrument identification based on F0-dependent multivariate normal distribution. | Tetsuro Kitahara, Masataka Goto, Hiroshi G. Okuno |
| 2003 | IJCAI | A Learning-Based Jam Session System that Imitates a Player's Personality Model. | Masatoshi Hamanaka, Masataka Goto, Hideki Asoh, Nobuyuki Otsu |
| 2003 | Interspeech | Speech shift: direct speech-input-mode switching through intentional control of voice pitch. | Masataka Goto, Yukihiro Omoto, Katunobu Itou, Tetsunori Kobayashi |
| 2003 | Interspeech | Speech starter: noise-robust endpoint detection by using filled pauses. | Koji Kitayama, Masataka Goto, Katunobu Itou, Tetsunori Kobayashi |
| 2003 | UIST | SmartMusicKIOSK: music listening station with chorus-search function. | Masataka Goto |
| 2002 | Interspeech | Speech completion: on-demand completion assistance using filled pauses for speech input interfaces. | Masataka Goto, Katunobu Itou, Satoru Hayamizu |
| 2001 | ICASSP | A predominant-F | Masataka Goto |
| 2001 | Interspeech | Real-time sound source localization and separation system and its application to automatic speech recognition. | Futoshi Asano, Masataka Goto, Katunobu Itou, Hideki Asoh |
| 2000 | ICASSP | A robust predominant-F0 estimation method for real-time detection of melody and bass lines in CD recordings. | Masataka Goto |
| 1999 | Interspeech | A real-time filled pause detection system for spontaneous speech recognition. | Masataka Goto, Katunobu Itou, Satoru Hayamizu |
| 1996 | ICASSP | Localization by harmonic structure and its application to harmonic sound stream segregation. | Tomohiro Nakatani, Masataka Goto, Hiroshi G. Okuno |