Skip to content

Kevin W. Wilson

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

27

Venues

6

Active years

2002–2024

Best venue rank

A*

Where they publish

Papers

27 indexed papers, newest first.

YearVenueTitleAuthors
2024ICASSPUnsupervised Multi-Channel Separation And Adaptation.Cong Han, Kevin W. Wilson, Scott Wisdom, John R. Hershey
2022InterspeechDistance-Based Sound Separation.Katharine Patterson, Kevin W. Wilson, Scott Wisdom, John R. Hershey
2021ICASSPEnd-To-End Diarization for Variable Number of Speakers with Local-Global Networks and Discriminative Speaker Embeddings.Soumi Maiti, Hakan Erdogan, Kevin W. Wilson, Scott Wisdom, Shinji Watanabe, John R. Hershey
2020InterspeechVoiceFilter-Lite: Streaming Targeted Voice Separation for On-Device Speech Recognition.Quan Wang, Ignacio Lpez-Moreno, Mert Saglam, Kevin W. Wilson, Alan Chiao, Renjie Liu, Yanzhang He, Wei Li, Jason Pelecanos, Marily Nika, Alexander Gruenstein
2019ICASSPDifferentiable Consistency Constraints for Improved Deep Speech Enhancement.Scott Wisdom, John R. Hershey, Kevin W. Wilson, Jeremy Thorpe, Michael Chinen, Brian Patton, Rif A. Saurous
2019InterspeechVoiceFilter: Targeted Voice Separation by Speaker-Conditioned Spectrogram Masking.Quan Wang, Hannah Muckenhirn, Kevin W. Wilson, Prashant Sridhar, Zelin Wu, John R. Hershey, Rif A. Saurous, Ron J. Weiss, Ye Jia, Ignacio Lpez-Moreno
2018InterspeechAVA-Speech: A Densely Labeled Dataset of Speech Activity in Movies.Sourish Chaudhuri, Joseph Roth, Daniel P. W. Ellis, Andrew C. Gallagher, Liat Kaver, Radhika Marvin, Caroline Pantofaru, Nathan Reale, Loretta Guarino Reid, Kevin W. Wilson, Zhonghua Xi
2017ICASSPCNN architectures for large-scale audio classification.Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis, Jort F. Gemmeke, Aren Jansen, R. Channing Moore, Manoj Plakal, Devin Platt, Rif A. Saurous, Bryan Seybold, Malcolm Slaney, Ron J. Weiss, Kevin W. Wilson
2017InterspeechAcoustic Modeling for Google Home.Bo Li, Tara N. Sainath, Arun Narayanan, Joe Caroselli, Michiel Bacchiani, Ananya Misra, Izhak Shafran, Hasim Sak, Golan Pundak, Kean K. Chin, Khe Chai Sim, Ron J. Weiss, Kevin W. Wilson, Ehsan Variani, Chanwoo Kim, Olivier Siohan, Mitchel Weintraub, Erik McDermott, Richard Rose, Matt Shannon
2016ICASSPFactored spatial and spectral multichannel raw waveform CLDNNs.Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Arun Narayanan, Michiel Bacchiani
2016InterspeechNeural Network Adaptive Beamforming for Robust Multichannel Speech Recognition.Bo Li, Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Michiel Bacchiani
2016InterspeechReducing the Computational Complexity of Multimicrophone Acoustic Models with Integrated Feature Extraction.Tara N. Sainath, Arun Narayanan, Ron J. Weiss, Ehsan Variani, Kevin W. Wilson, Michiel Bacchiani, Izhak Shafran
2015ASRUSpeaker location and microphone spacing invariant acoustic modeling from raw multichannel waveforms.Tara N. Sainath, Ron J. Weiss, Kevin W. Wilson, Arun Narayanan, Michiel Bacchiani, Andrew W. Senior
2015ICASSPSpeech acoustic modeling from raw multichannel waveforms.Yedid Hoshen, Ron J. Weiss, Kevin W. Wilson
2015InterspeechLearning the speech front-end with raw waveform CLDNNs.Tara N. Sainath, Ron J. Weiss, Andrew W. Senior, Kevin W. Wilson, Oriol Vinyals
2010ICASSPSpectrogram dimensionality reductionwith independence constraints.Kevin W. Wilson, Bhiksha Raj
2010InterspeechUngrounded independent non-negative factor analysis.Bhiksha Raj, Kevin W. Wilson, Alexander Krueger, Reinhold Haeb-Umbach
2008ICASSPSpeech denoising using nonnegative matrix factorization with priors.Kevin W. Wilson, Bhiksha Raj, Paris Smaragdis, Ajay Divakaran
2008InterspeechRegularized non-negative matrix factorization with temporal dependencies for speech denoising.Kevin W. Wilson, Bhiksha Raj, Paris Smaragdis
2005ICASSPImproving audio source localization by learning the precedence effect.Kevin W. Wilson, Trevor Darrell
2005ICCVVisual Speech Recognition with Loosely Synchronized Feature Streams.Kate Saenko, Karen Livescu, Michael Siracusa, Kevin W. Wilson, James R. Glass, Trevor Darrell
2004ICASSPMultiple person and speaker activity tracking with a particle filter.Neal Checka, Kevin W. Wilson, Michael Siracusa, Trevor Darrell
2004ICMIReal-time audio-visual tracking for meeting analysis.David Demirdjian, Kevin W. Wilson, Michael Siracusa, Trevor Darrell
2003CVPRA Probabilistic Framework for Multi-modal Multi-Person Tracking.Neal Checka, Kevin W. Wilson, Vibhav Rangarajan, Trevor Darrell
2003ICMIA multi-modal approach for determining speaker location and focus.Michael Siracusa, Louis-Philippe Morency, Kevin W. Wilson, John W. Fisher III, Trevor Darrell
2002ICASSPAudio-video array source localization for intelligent environments.Kevin W. Wilson, Trevor Darrell
2002ICMIAudiovisual Arrays for Untethered Spoken Interfaces.Kevin W. Wilson, Vibhav Rangarajan, Neal Checka, Trevor Darrell