| 2022 | UIST | Record Once, Post Everywhere: Automatic Shortening of Audio Stories for Social Media. | Bryan Wang, Zeyu Jin, Gautham J. Mysore |
| 2020 | ICASSP | F0-Consistent Many-To-Many Non-Parallel Voice Conversion Via Conditional Autoencoder. | Kaizhi Qian, Zeyu Jin, Mark Hasegawa-Johnson, Gautham J. Mysore |
| 2020 | Interspeech | A Differentiable Perceptual Audio Metric Learned from Just Noticeable Differences. | Pranay Manocha, Adam Finkelstein, Richard Zhang, Nicholas J. Bryan, Gautham J. Mysore, Zeyu Jin |
| 2020 | Interspeech | Controllable Neural Prosody Synthesis. | Max Morrison, Zeyu Jin, Justin Salamon, Nicholas J. Bryan, Gautham J. Mysore |
| 2019 | CHI | B-Script: Transcript-based B-roll Video Editing with Recommendations. | Bernd Huber, Hijung Valentina Shin, Bryan C. Russell, Oliver Wang, Gautham J. Mysore |
| 2019 | CHI | VoiceAssist: Guiding Users to High-Quality Voice Recordings. | Prem Seetharaman, Gautham J. Mysore, Bryan Pardo, Paris Smaragdis, Celso Gomes |
| 2018 | CHI | LoopMaker: Automatic Creation of Music Loops from Pre-recorded Music. | Zhengshan Shi, Gautham J. Mysore |
| 2018 | ICASSP | Crowdsourced Pairwise-Comparison for Source Separation Evaluation. | Mark Cartwright, Bryan Pardo, Gautham J. Mysore |
| 2018 | ICASSP | Fftnet: A Real-Time Speaker-Dependent Neural Vocoder. | Zeyu Jin, Adam Finkelstein, Gautham J. Mysore, Jingwan Lu |
| 2018 | ICASSP | Blind Estimation of the Speech Transmission Index for Speech Quality Prediction. | Prem Seetharaman, Gautham J. Mysore, Paris Smaragdis, Bryan Pardo |
| 2018 | IUI | MedleyAssistant - A System for Personalized Music Medley Creation. | Zhengshan Shi, Gautham J. Mysore |
| 2017 | UIST | AutoDub: Automatic Redubbing for Voiceover Editing. | Shrikant Venkataramani, Paris Smaragdis, Gautham J. Mysore |
| 2016 | ICASSP | Fast and easy crowdsourced perceptual audio evaluation. | Mark Cartwright, Bryan Pardo, Gautham J. Mysore, Matthew D. Hoffman |
| 2016 | ICASSP | Equalization matching of speech recordings in real-world environments. | Franois G. Germain, Gautham J. Mysore, Takako Fujioka |
| 2016 | ICASSP | Cute: A concatenative method for voice conversion using exemplar-based unit selection. | Zeyu Jin, Adam Finkelstein, Stephen DiVerdi, Jingwan Lu, Gautham J. Mysore |
| 2016 | ICASSP | Structural segmentation with the Variable Markov Oracle and boundary adjustment. | Cheng-i Wang, Gautham J. Mysore |
| 2015 | CHI | Lamello: Passive Acoustic Sensing for Tangible Input Components. | Valkyrie Savage, Andrew Head, Bjrn Hartmann, Dan B. Goldman, Gautham J. Mysore, Wilmot Li |
| 2015 | ICASSP | Speaker and noise independent online single-channel speech enhancement. | Franois G. Germain, Gautham J. Mysore |
| 2015 | ICASSP | Efficient manifold preserving audio source separation using locality sensitive hashing. | Minje Kim, Paris Smaragdis, Gautham J. Mysore |
| 2015 | ICASSP | Speech dereverberation using a learned speech model. | Dawen Liang, Matthew D. Hoffman, Gautham J. Mysore |
| 2015 | UIST | Capture-Time Feedback for Recording Scripted Narration. | Steve Rubin, Floraine Berthouzoz, Gautham J. Mysore, Maneesh Agrawala |
| 2014 | CHI | ISSE: an interactive source separation editor. | Nicholas J. Bryan, Gautham J. Mysore, Ge Wang |
| 2014 | ICASSP | Exploiting long-term temporal dependencies in NMF using recurrent neural networks with application to source separation. | Nicolas Boulanger-Lewandowski, Gautham J. Mysore, Matthew D. Hoffman |
| 2014 | ICASSP | Speech decoloration based on the product-of-filters model. | Dawen Liang, Daniel P. W. Ellis, Matthew D. Hoffman, Gautham J. Mysore |
| 2013 | ICASSP | Interactive refinement of supervised and semi-supervised sound source separation estimates. | Nicholas J. Bryan, Gautham J. Mysore |
| 2013 | ICASSP | Universal speech models for speaker independent single channel source separation. | Dennis L. Sun, Gautham J. Mysore |
| 2013 | ICML | An Efficient Posterior Regularized Latent Variable Model for Interactive Sound Source Separation. | Nicholas J. Bryan, Gautham J. Mysore |
| 2013 | Interspeech | Speaker and noise independent voice activity detection. | Franois G. Germain, Dennis L. Sun, Gautham J. Mysore |
| 2013 | UIST | Content-based tools for editing audio stories. | Steve Rubin, Floraine Berthouzoz, Gautham J. Mysore, Wilmot Li, Maneesh Agrawala |
| 2012 | ICASSP | Clustering and synchronizing multi-camera video via landmark cross-correlation. | Nicholas J. Bryan, Paris Smaragdis, Gautham J. Mysore |
| 2012 | ICASSP | Noise-robust dynamic time warping using PLCA features. | Brian King, Paris Smaragdis, Gautham J. Mysore |
| 2012 | ICASSP | Following musical sources by example. | Paris Smaragdis, Gautham J. Mysore |
| 2012 | ICML | Variational Inference in Non-negative Factorial Hidden Markov Models for Efficient Audio Source Separation. | Gautham J. Mysore, Maneesh Sahani |
| 2012 | Interspeech | Speech Enhancement by Online Non-negative Spectrogram Decomposition in Non-stationary Noise Environments. | Zhiyao Duan, Gautham J. Mysore, Paris Smaragdis |
| 2012 | UIST | UnderScore: musical underlays for audio stories. | Steve Rubin, Floraine Berthouzoz, Gautham J. Mysore, Wilmot Li, Maneesh Agrawala |
| 2011 | ICASSP | A non-negative approach to semi-supervised separation of speech from noise with the use of temporal dynamics. | Gautham J. Mysore, Paris Smaragdis |
| 2010 | Interspeech | A super-resolution spectrogram using coupled PLCA. | Juhan Nam, Gautham J. Mysore, Joachim Ganseman, Kyogu Lee, Jonathan S. Abel |
| 2009 | ICASSP | Relative pitch estimation of multiple instruments. | Gautham J. Mysore, Paris Smaragdis |
| 2009 | IDA | Probabilistic Factorization of Non-negative Data with Entropic Co-occurrence Constraints. | Paris Smaragdis, Madhusudana V. S. Shashanka, Bhiksha Raj, Gautham J. Mysore |