Kevin W. Wilson
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
27
Venues
6
Active years
2002–2024
Best venue rank
A*
Where they publish
Papers
27 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2024 | ICASSP | Unsupervised Multi-Channel Separation And Adaptation. | Cong Han, Kevin W. Wilson, Scott Wisdom, John R. Hershey |
| 2022 | Interspeech | Distance-Based Sound Separation. | Katharine Patterson, Kevin W. Wilson, Scott Wisdom, John R. Hershey |
| 2021 | ICASSP | End-To-End Diarization for Variable Number of Speakers with Local-Global Networks and Discriminative Speaker Embeddings. | Soumi Maiti, Hakan Erdogan, Kevin W. Wilson, Scott Wisdom, Shinji Watanabe, John R. Hershey |
| 2020 | Interspeech | VoiceFilter-Lite: Streaming Targeted Voice Separation for On-Device Speech Recognition. | Quan Wang, Ignacio Lpez-Moreno, Mert Saglam, Kevin W. Wilson, Alan Chiao, Renjie Liu, Yanzhang He, Wei Li, Jason Pelecanos, Marily Nika, Alexander Gruenstein |
| 2019 | ICASSP | Differentiable Consistency Constraints for Improved Deep Speech Enhancement. | Scott Wisdom, John R. Hershey, Kevin W. Wilson, Jeremy Thorpe, Michael Chinen, Brian Patton, Rif A. Saurous |
| 2019 | Interspeech | VoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking. | Quan Wang, Hannah Muckenhirn, Kevin W. Wilson, Prashant Sridhar, Zelin Wu, John R. Hershey, Rif A. Saurous, Ron J. Weiss, Ye Jia, Ignacio Lpez-Moreno |
| 2018 | Interspeech | AVA-Speech: A Densely Labeled Dataset of Speech Activity in Movies. | Sourish Chaudhuri, Joseph Roth, Daniel P. W. Ellis, Andrew C. Gallagher, Liat Kaver, Radhika Marvin, Caroline Pantofaru, Nathan Reale, Loretta Guarino Reid, Kevin W. Wilson, Zhonghua Xi |
| 2017 | ICASSP | CNN architectures for large-scale audio classification. | Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis, Jort F. Gemmeke, Aren Jansen, R. Channing Moore, Manoj Plakal, Devin Platt, Rif A. Saurous, Bryan Seybold, Malcolm Slaney, Ron J. Weiss, Kevin W. Wilson |
| 2017 | Interspeech | Acoustic Modeling for Google Home. | Bo Li, Tara N. Sainath, Arun Narayanan, Joe Caroselli, Michiel Bacchiani, Ananya Misra, Izhak Shafran, Hasim Sak, Golan Pundak, Kean K. Chin, Khe Chai Sim, Ron J. Weiss, Kevin W. Wilson, Ehsan Variani, Chanwoo Kim, Olivier Siohan, Mitchel Weintraub, Erik McDermott, Richard Rose, Matt Shannon |
| 2016 | ICASSP | Factored spatial and spectral multichannel raw waveform CLDNNs. | Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Arun Narayanan, Michiel Bacchiani |
| 2016 | Interspeech | Neural Network Adaptive Beamforming for Robust Multichannel Speech Recognition. | Bo Li, Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Michiel Bacchiani |
| 2016 | Interspeech | Reducing the Computational Complexity of Multimicrophone Acoustic Models with Integrated Feature Extraction. | Tara N. Sainath, Arun Narayanan, Ron J. Weiss, Ehsan Variani, Kevin W. Wilson, Michiel Bacchiani, Izhak Shafran |
| 2015 | ASRU | Speaker location and microphone spacing invariant acoustic modeling from raw multichannel waveforms. | Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Arun Narayanan, Michiel Bacchiani, Andrew W. Senior |
| 2015 | ICASSP | Speech acoustic modeling from raw multichannel waveforms. | Yedid Hoshen, Ron J. Weiss, Kevin W. Wilson |
| 2015 | Interspeech | Learning the speech front-end with raw waveform CLDNNs. | Tara N. Sainath, Ron J. Weiss, Andrew W. Senior, Kevin W. Wilson, Oriol Vinyals |
| 2010 | ICASSP | Spectrogram dimensionality reductionwith independence constraints. | Kevin W. Wilson, Bhiksha Raj |
| 2010 | Interspeech | Ungrounded independent non-negative factor analysis. | Bhiksha Raj, Kevin W. Wilson, Alexander Krueger, Reinhold Haeb-Umbach |
| 2008 | ICASSP | Speech denoising using nonnegative matrix factorization with priors. | Kevin W. Wilson, Bhiksha Raj, Paris Smaragdis, Ajay Divakaran |
| 2008 | Interspeech | Regularized non-negative matrix factorization with temporal dependencies for speech denoising. | Kevin W. Wilson, Bhiksha Raj, Paris Smaragdis |
| 2005 | ICASSP | Improving audio source localization by learning the precedence effect. | Kevin W. Wilson, Trevor Darrell |
| 2005 | ICCV | Visual Speech Recognition with Loosely Synchronized Feature Streams. | Kate Saenko, Karen Livescu, Michael Siracusa, Kevin W. Wilson, James R. Glass, Trevor Darrell |
| 2004 | ICASSP | Multiple person and speaker activity tracking with a particle filter. | Neal Checka, Kevin W. Wilson, Michael Siracusa, Trevor Darrell |
| 2004 | ICMI | Real-time audio-visual tracking for meeting analysis. | David Demirdjian, Kevin W. Wilson, Michael Siracusa, Trevor Darrell |
| 2003 | CVPR | A Probabilistic Framework for Multi-modal Multi-Person Tracking. | Neal Checka, Kevin W. Wilson, Vibhav Rangarajan, Trevor Darrell |
| 2003 | ICMI | A multi-modal approach for determining speaker location and focus. | Michael Siracusa, Louis-Philippe Morency, Kevin W. Wilson, John W. Fisher III, Trevor Darrell |
| 2002 | ICASSP | Audio-video array source localization for intelligent environments. | Kevin W. Wilson, Trevor Darrell |
| 2002 | ICMI | Audiovisual Arrays for Untethered Spoken Interfaces. | Kevin W. Wilson, Vibhav Rangarajan, Neal Checka, Trevor Darrell |