| 2025 | ASRU | Improving Resource-Efficient Speech Enhancement via Neural Differentiable DSP Vocoder Refinement. | Heitor R. Guimares, Ke Tan, Juan Azcarreta, Jesus Alvarez, Prabhav Agrawal, Ashutosh Pandey, Buye Xu |
| 2025 | ICASSP | Robust Frame-level Speaker Localization in Reverberant and Noisy Environments by Exploiting Phase Difference Losses. | Shanmukha Srinivas Battula, Hassan Taherian, Ashutosh Pandey, Daniel Wong, Buye Xu, DeLiang Wang |
| 2025 | ICASSP | Modulating State Space Model with SlowFast Framework for Compute-Efficient Ultra Low-Latency Speech Enhancement. | Longbiao Cheng, Ashutosh Pandey, Buye Xu, Tobi Delbruck, Vamsi Krishna Ithapu, Shih-Chii Liu |
| 2025 | ICASSP | Advancing Active Speaker Detection for Egocentric Videos. | Jaesung Huh, Juan Azcarreta Ortiz, Anurag Kumar, Ashutosh Pandey, Ali Aroudi, Daniel D. E. Wong, Francesco Nesta, Buye Xu, Jacob Donley |
| 2025 | ICASSP | Reexamining the Efficacy of MetricGAN for Speech Enhancement. | Haibin Wu, Ali Aroudi, Buye Xu, Ashutosh Pandey, Francesco Nesta, Anurag Kumar, Alexander Reich, Ke Tan |
| 2025 | Interspeech | A Novel Deep Learning Framework for Efficient Multichannel Acoustic Feedback Control. | Yuan-Kuei Wu, Juan Azcarreta Ortiz, Kashyap Patel, Buye Xu, Jung-Suk Lee, Sanha Lee, Ashutosh Pandey |
| 2025 | Interspeech | Online AV-CrossNet: a Causal and Efficient Audiovisual System for Speech Enhancement and Target Speaker Extraction. | Cheng Yu, Vahid Ahmadi Kalkhorani, Buye Xu, DeLiang Wang |
| 2024 | ICASSP | Decoupled Spatial and Temporal Processing for Resource Efficient Multichannel Speech Enhancement. | Ashutosh Pandey, Buye Xu |
| 2024 | ICASSP | On the Importance of Neural Wiener Filter for Resource Efficient Multichannel Speech Enhancement. | Tsun-An Hsieh, Jacob Donley, Daniel Wong, Buye Xu, Ashutosh Pandey |
| 2024 | ICASSP | Audiovisual Speaker Separation with Full- and Sub-Band Modeling in the Time-Frequency Domain. | Vahid Ahmadi Kalkhorani, Anurag Kumar, Ke Tan, Buye Xu, DeLiang Wang |
| 2024 | ICASSP | A Closer Look at Wav2vec2 Embeddings for On-Device Single-Channel Speech Enhancement. | Ravi Shankar, Ke Tan, Buye Xu, Anurag Kumar |
| 2024 | ICASSP | Leveraging Sound Localization to Improve Continuous Speaker Separation. | Hassan Taherian, Ashutosh Pandey, Daniel Wong, Buye Xu, DeLiang Wang |
| 2024 | Interspeech | Dynamic Gated Recurrent Neural Network for Compute-efficient Speech Enhancement. | Longbiao Cheng, Ashutosh Pandey, Buye Xu, Tobi Delbruck, Shih-Chii Liu |
| 2024 | Interspeech | All Neural Low-latency Directional Speech Extraction. | Ashutosh Pandey, Sanha Lee, Juan Azcarreta, Daniel Wong, Buye Xu |
| 2024 | Interspeech | Towards Explainable Monaural Speaker Separation with Auditory-based Training. | Hassan Taherian, Vahid Ahmadi Kalkhorani, Ashutosh Pandey, Daniel Wong, Buye Xu, DeLiang Wang |
| 2024 | Interspeech | FoVNet: Configurable Field-of-View Speech Enhancement with Low Computation and Distortion for Smart Glasses. | Zhongweiyang Xu, Ali Aroudi, Ke Tan, Ashutosh Pandey, Jung-Suk Lee, Buye Xu, Francesco Nesta |
| 2023 | ICASSP | Leveraging Heteroscedastic Uncertainty in Learning Complex Spectral Mapping for Single-Channel Speech Enhancement. | Kuan-Lin Chen, Daniel D. E. Wong, Ke Tan, Buye Xu, Anurag Kumar, Vamsi Krishna Ithapu |
| 2023 | ICASSP | Torchaudio-Squim: Reference-Less Speech Quality and Intelligibility Measures in Torchaudio. | Anurag Kumar, Ke Tan, Zhaoheng Ni, Pranay Manocha, Xiaohui Zhang, Ethan Henderson, Buye Xu |
| 2023 | ICASSP | LA-VOCE: LOW-SNR Audio-Visual Speech Enhancement Using Neural Vocoders. | Rodrigo Mira, Buye Xu, Jacob Donley, Anurag Kumar, Stavros Petridis, Vamsi Krishna Ithapu, Maja Pantic |
| 2023 | Interspeech | A Simple RNN Model for Lightweight, Low-compute and Low-latency Multichannel Speech Enhancement in the Time Domain. | Ashutosh Pandey, Ke Tan, Buye Xu |
| 2023 | Interspeech | Time-domain Transformer-based Audiovisual Speaker Separation. | Vahid Ahmadi Kalkhorani, Anurag Kumar, Ke Tan, Buye Xu, DeLiang Wang |
| 2023 | Interspeech | Multi-input Multi-output Complex Spectral Mapping for Speaker Separation. | Hassan Taherian, Ashutosh Pandey, Daniel Wong, Buye Xu, DeLiang Wang |
| 2023 | Interspeech | Rethinking Complex-Valued Deep Neural Networks for Monaural Speech Enhancement. | Haibin Wu, Ke Tan, Buye Xu, Anurag Kumar, Daniel Wong |
| 2022 | ICASSP | TPARN: Triple-Path Attentive Recurrent Network for Time-Domain Multichannel Speech Enhancement. | Ashutosh Pandey, Buye Xu, Anurag Kumar, Jacob Donley, Paul Calamia, DeLiang Wang |
| 2022 | ICASSP | Multichannel Speech Enhancement Without Beamforming. | Ashutosh Pandey, Buye Xu, Anurag Kumar, Jacob Donley, Paul Calamia, DeLiang Wang |
| 2022 | ICASSP | Continual Self-Training With Bootstrapped Remixing For Speech Enhancement. | Efthymios Tzinis, Yossi Adi, Vamsi K. Ithapu, Buye Xu, Anurag Kumar |
| 2022 | Interspeech | Time-domain Ad-hoc Array Speech Enhancement Using a Triple-path Network. | Ashutosh Pandey, Buye Xu, Anurag Kumar, Jacob Donley, Paul Calamia, DeLiang Wang |
| 2022 | Interspeech | SAQAM: Spatial Audio Quality Assessment Metric. | Pranay Manocha, Anurag Kumar, Buye Xu, Anjali Menon, Israel Dejene Gebru, Vamsi Krishna Ithapu, Paul Calamia |
| 2021 | ASRU | Incorporating Real-World Noisy Speech in Neural-Network-Based Speech Enhancement Systems. | Yangyang Xia, Buye Xu, Anurag Kumar |
| 2018 | ICASSP | Late Reverberation Suppression Using Recurrent Neural Networks with Long Short-Term Memory. | Yan Zhao, DeLiang Wang, Buye Xu, Tao Zhang |
| 2018 | ICASSP | Perceptually Guided Speech Enhancement Using Deep Neural Networks. | Yan Zhao, Buye Xu, Ritwik Giri, Tao Zhang |
| 2014 | ICASSP | Design of a high order binaural microphone array for hearing aids using a rigid spherical model. | Ivo Merks, Buye Xu, Tao Zhang |
| 2012 | ICASSP | Annoyance perception and modeling for hearing-impaired listeners. | Srikanth Vishnubhotla, Jinjun Xiao, Buye Xu, Martin F. McKinney, Tao Zhang |