Skip to content

Sameer Khurana

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

27

Venues

7

Active years

2016–2025

Best venue rank

A*

Where they publish

Papers

27 indexed papers, newest first.

YearVenueTitleAuthors
2025ICASSPInteractive Robot Action Replanning using Multimodal LLM Trained from Human Demonstration Videos.Chiori Hori, Motonari Kambara, Komei Sugiura, Kei Ota, Sameer Khurana, Siddarth Jain, Radu Corcodel, Devesh K. Jha, Diego Romeres, Jonathan Le Roux
2025ICASSPLeveraging Audio-Only Data for Text-Queried Target Sound Extraction.Kohei Saijo, Janek Ebbers, Franois G. Germain, Sameer Khurana, Gordon Wichern, Jonathan Le Roux
2025ICMLAligning Multimodal Representations through an Information Bottleneck.Antonio Almudvar, Jos Miguel Hernndez-Lobato, Sameer Khurana, Ricard Marxer, Alfonso Ortega
2025InterspeechHASRD: Hierarchical Acoustic and Semantic Representation Disentanglement.Amir Hussein, Sameer Khurana, Gordon Wichern, Franois G. Germain, Jonathan Le Roux
2025InterspeechFactorized RVQ-GAN For Disentangled Speech Tokenization.Sameer Khurana, Dominik Klement, Antoine Laurent, Dominik Bobos, Juraj Novosad, Peter Gazdik, Ellen Zhang, Zili Huang, Amir Hussein, Ricard Marxer, Yoshiki Masuyama, Ryo Aihara, Chiori Hori, Franois G. Germain, Gordon Wichern, Jonathan Le Roux
2025InterspeechInvestigating continuous autoregressive generative speech enhancement.Haici Yang, Gordon Wichern, Ryo Aihara, Yoshiki Masuyama, Sameer Khurana, Franois G. Germain, Jonathan Le Roux
2024ICASSPGeneration or Replication: Auscultating Audio Latent Diffusion Models.Dimitrios Bralios, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024ICASSPWI-FI based Indoor Monitoring Enhanced by Multimodal Fusion.Chiori Hori, Pu Wang, Mahbub Rahman, Cristian J. Vaca-Rubio, Sameer Khurana, Anoop Cherian, Jonathan Le Roux
2024ICASSPCross-Lingual Transfer Learning for Low-Resource Speech Translation.Sameer Khurana, Nauman Dawalatabad, Antoine Laurent, Luis Vicente, Pablo Gimeno, Victoria Mingote, James R. Glass
2024ICASSPNIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization.Yoshiki Masuyama, Gordon Wichern, Franois G. Germain, Zexu Pan, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2024ICASSPNeuroHeed+: Improving Neuro-Steered Speaker Extraction with Joint Auditory Attention Detection.Zexu Pan, Gordon Wichern, Franois G. Germain, Sameer Khurana, Jonathan Le Roux
2024InterspeechZeroST: Zero-Shot Speech Translation.Sameer Khurana, Chiori Hori, Antoine Laurent, Gordon Wichern, Jonathan Le Roux
2023ASRUScenario-Aware Audio-Visual TF-Gridnet for Target Speech Extraction.Zexu Pan, Gordon Wichern, Yoshiki Masuyama, Franois G. Germain, Sameer Khurana, Chiori Hori, Jonathan Le Roux
2023ICASSPOn Unsupervised Uncertainty-Driven Speech Pseudo-Label Filtering and Model Calibration.Nauman Dawalatabad, Sameer Khurana, Antoine Laurent, James R. Glass
2023InterspeechWhisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong General Audio Event Taggers.Yuan Gong, Sameer Khurana, Leonid Karlinsky, James R. Glass
2023InterspeechComparison of Multilingual Self-Supervised and Weakly-Supervised Speech Pre-Training for Adaptation to Unseen Languages.Andrew Rouditchenko, Sameer Khurana, Samuel Thomas, Rogrio Feris, Leonid Karlinsky, Hilde Kuehne, David Harwath, Brian Kingsbury, James R. Glass
2022EMNLPDetecting Dementia from Long Neuropsychological Interviews.Nauman Dawalatabad, Yuan Gong, Sameer Khurana, Rhoda Au, James R. Glass
2022ICASSPMagic Dust for Cross-Lingual Adaptation of Monolingual Wav2vec-2.0.Sameer Khurana, Antoine Laurent, James R. Glass
2021ICASSPUnsupervised Domain Adaptation for Speech Recognition via Uncertainty Driven Self-Training.Sameer Khurana, Niko Moritz, Takaaki Hori, Jonathan Le Roux
2020IJCNNRobust Training of Vector Quantized Bottleneck Models.Adrian Lancucki, Jan Chorowski, Guillaume Sanchez, Ricard Marxer, Nanxin Chen, Hans J. G. A. Dolfing, Sameer Khurana, Tanel Alume, Antoine Laurent
2020InterspeechA Convolutional Deep Markov Model for Unsupervised Speech Representation Learning.Sameer Khurana, Antoine Laurent, Wei-Ning Hsu, Jan Chorowski, Adrian Lancucki, Ricard Marxer, James R. Glass
2019ICASSPA Factorial Deep Markov Model for Unsupervised Disentangled Representation Learning from Speech.Sameer Khurana, Shafiq Rayhan Joty, Ahmed Ali, James R. Glass
2018ICASSPExploiting Convolutional Neural Networks for Phonotactic Based Dialect Identification.Maryam Najafian, Sameer Khurana, Suwon Shon, Ahmed Ali, James R. Glass
2017EACLQCRI Live Speech Translation System.Fahim Dalvi, Yifan Zhang, Sameer Khurana, Nadir Durrani, Hassan Sajjad, Ahmed Abdelali, Hamdy Mubarak, Ahmed M. Ali, Stephan Vogel
2017EACLThe SUMMA Platform Prototype.Renars Liepins, Ulrich Germann, Guntis Barzdins, Alexandra Birch, Steve Renals, Susanne Weber, Peggy van der Kreeft, Herv Bourlard, Joo Prieto, Ondrej Klejch, Peter Bell, Alexandros Lazaridis, Afonso Mendes, Sebastian Riedel, Mariana S. C. Almeida, Pedro Balage, Shay B. Cohen, Tomasz Dwojak, Philip N. Garner, Andreas Giefer, Marcin Junczys-Dowmunt, Hina Imran, David Nogueira, Ahmed M. Ali, Sebastio Miranda, Andrei Popescu-Belis, Lesly Miculicich Werlen, Nikos Papasarantopoulos, Abiola Obamuyide, Clive Jones, Fahim Dalvi, Andreas Vlachos, Yang Wang, Sibo Tong, Rico Sennrich, Nikolaos Pappas, Shashi Narayan, Marco Damonte, Nadir Durrani, Sameer Khurana, Ahmed Abdelali, Hassan Sajjad, Stephan Vogel, David Sheppey, Chris Hernon, Jeff Mitchell
2017InterspeechQMDIS: QCRI-MIT Advanced Dialect Identification System.Sameer Khurana, Maryam Najafian, Ahmed Ali, Tuka Al Hanai, Yonatan Belinkov, James R. Glass
2016InterspeechAutomatic Dialect Detection in Arabic Broadcast Speech.Ahmed Ali, Najim Dehak, Patrick Cardinal, Sameer Khurana, Sree Harsha Yella, James R. Glass, Peter Bell, Steve Renals