| 2025 | ICASSP | Towards Sub-millisecond Latency Real-Time Speech Enhancement Models on Hearables. | Artem Dementyev, Chandan K. A. Reddy, Scott Wisdom, Navin Chatlani, John R. Hershey, Richard F. Lyon |
| 2024 | ICASSP | Unsupervised Multi-Channel Separation And Adaptation. | Cong Han, Kevin W. Wilson, Scott Wisdom, John R. Hershey |
| 2023 | ICASSP | Audioslots: A Slot-Centric Generative Model For Audio Separation. | Pradyumna Reddy, Scott Wisdom, Klaus Greff, John R. Hershey, Thomas Kipf |
| 2023 | Interspeech | TokenSplit: Using Discrete Speech Representations for Direct, Refined, and Transcript-Conditioned Speech Separation and Recognition. | Hakan Erdogan, Scott Wisdom, Xuankai Chang, Zaln Borsos, Marco Tagliasacchi, Neil Zeghidour, John R. Hershey |
| 2022 | ECCV | AudioScopeV2: Audio-Visual Attention Architectures for Calibrated Open-Domain On-Screen Sound Separation. | Efthymios Tzinis, Scott Wisdom, Tal Remez, John R. Hershey |
| 2022 | ICASSP | Improving Bird Classification with Unsupervised Sound Separation. | Tom Denton, Scott Wisdom, John R. Hershey |
| 2022 | ICASSP | Adapting Speech Separation to Real-World Meetings using Mixture Invariant Training. | Aswin Sivaraman, Scott Wisdom, Hakan Erdogan, John R. Hershey |
| 2022 | Interspeech | Text-Driven Separation of Arbitrary Sounds. | Kevin Kilgour, Beat Gfeller, Qingqing Huang, Aren Jansen, Scott Wisdom, Marco Tagliasacchi |
| 2022 | Interspeech | CycleGAN-based Unpaired Speech Dereverberation. | Hannah Muckenhirn, Aleksandr Safin, Hakan Erdogan, Felix de Chaumont Quitry, Marco Tagliasacchi, Scott Wisdom, John R. Hershey |
| 2022 | Interspeech | Distance-Based Sound Separation. | Katharine Patterson, Kevin W. Wilson, Scott Wisdom, John R. Hershey |
| 2022 | Interspeech | Listening with Googlears: Low-Latency Neural Multiframe Beamforming and Equalization for Hearing Aids. | Samuel J. Yang, Scott Wisdom, Chet Gnegy, Richard F. Lyon, Sagar Savla |
| 2021 | ICASSP | End-To-End Diarization for Variable Number of Speakers with Local-Global Networks and Discriminative Speaker Embeddings. | Soumi Maiti, Hakan Erdogan, Kevin W. Wilson, Scott Wisdom, Shinji Watanabe, John R. Hershey |
| 2021 | ICASSP | Sound Event Detection and Separation: A Benchmark on Desed Synthetic Soundscapes. | Nicolas Turpault, Romain Serizel, Scott Wisdom, Hakan Erdogan, John R. Hershey, Eduardo Fonseca, Prem Seetharaman, Justin Salamon |
| 2021 | ICASSP | What's all the Fuss about Free Universal Sound Separation Data? | Scott Wisdom, Hakan Erdogan, Daniel P. W. Ellis, Romain Serizel, Nicolas Turpault, Eduardo Fonseca, Justin Salamon, Prem Seetharaman, John R. Hershey |
| 2021 | ICLR | Into the Wild with AudioScope: Unsupervised Audio-Visual Separation of On-Screen Sounds. | Efthymios Tzinis, Scott Wisdom, Aren Jansen, Shawn Hershey, Tal Remez, Dan Ellis, John R. Hershey |
| 2020 | ICASSP | Performance Study of a Convolutional Time-Domain Audio Separation Network for Real-Time Speech Denoising. | Samuel Sonning, Christian Schldt, Hakan Erdogan, Scott Wisdom |
| 2020 | ICASSP | Improving Universal Sound Separation Using Sound Classification. | Efthymios Tzinis, Scott Wisdom, John R. Hershey, Aren Jansen, Daniel P. W. Ellis |
| 2019 | ICASSP | SDR - Half-baked or Well Done? | Jonathan Le Roux, Scott Wisdom, Hakan Erdogan, John R. Hershey |
| 2019 | ICASSP | Differentiable Consistency Constraints for Improved Deep Speech Enhancement. | Scott Wisdom, John R. Hershey, Kevin W. Wilson, Jeremy Thorpe, Michael Chinen, Brian Patton, Rif A. Saurous |
| 2017 | ICASSP | Building recurrent networks by unfolding iterative thresholding for sequential sparse recovery. | Scott Wisdom, Thomas Powers, James W. Pitton, Les E. Atlas |
| 2016 | ACSSC | Benefits of noncircular statistics for nonstationary signals. | Scott Wisdom, Les E. Atlas, James W. Pitton, Greg Okopal |
| 2016 | ICASSP | Deep unfolding for multichannel source separation. | Scott Wisdom, John R. Hershey, Jonathan Le Roux, Shinji Watanabe |
| 2015 | ICASSP | Voice activity detection using subband noncircularity. | Scott Wisdom, Greg Okopal, Les E. Atlas, James W. Pitton |
| 2014 | ACSSC | Estimating the noncircularity of latent components within complex-valued subband mixtures with applications to speech processing. | Greg Okopal, Scott Wisdom, Les Atlas |
| 2014 | ACSSC | Extending coherence for optimal detection of nonstationary harmonic signals. | Scott Wisdom, James W. Pitton, Les Atlas |
| 2014 | ICASSP | Extending coherence time for analysis of modulated random processes. | Scott Wisdom, Les Atlas, James Pittore |