Daniel Garcia-Romero
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
51
Venues
6
Active years
2003–2025
Best venue rank
A*
Where they publish
Papers
51 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICASSP | Hyper-adapter for Parameter-Efficient Multilingual ASR Adaptation. | Zejiang Hou, Daniel Garcia-Romero, Kyu J. Han |
| 2025 | ICASSP | Zero-resource Speech Translation and Recognition with LLMs. | Karel Mundnich, Xing Niu, Prashant Mathur, Srikanth Ronanki, Brady Houston, Veera Raghavendra Elluru, Nilaksh Das, Zejiang Hou, Goeric Huybrechts, Anshu Bhatia, Daniel Garcia-Romero, Kyu J. Han, Katrin Kirchhoff |
| 2025 | ICASSP | Knowledge Distillation From Ensemble for Spoken Language Identification. | Raghuveer Peri, Seyed Omid Sadjadi, Daniel Garcia-Romero, Srikanth Vishnubhotla, Kyu J. Han |
| 2025 | ICASSP | Contextual ASR with Retrieval Augmented Large Language Model. | Cihan Xiao, Zejiang Hou, Daniel Garcia-Romero, Kyu J. Han |
| 2024 | ACL | SpeechGuard: Exploring the Adversarial Robustness of Multi-modal Large Language Models. | Raghuveer Peri, Sai Muralidhar Jayanthi, Srikanth Ronanki, Anshu Bhatia, Karel Mundnich, Saket Dingliwal, Nilaksh Das, Zejiang Hou, Goeric Huybrechts, Srikanth Vishnubhotla, Daniel Garcia-Romero, Sundararajan Srinivasan, Kyu J. Han, Katrin Kirchhoff |
| 2024 | Interspeech | Revisiting Convolution-free Transformer for Speech Recognition. | Zejiang Hou, Goeric Huybrechts, Anshu Bhatia, Daniel Garcia-Romero, Kyu J. Han, Katrin Kirchhoff |
| 2022 | Interspeech | Directed speech separation for automatic speech recognition of long form conversational speech. | Rohit Paturi, Sundararajan Srinivasan, Katrin Kirchhoff, Daniel Garcia-Romero |
| 2021 | ICASSP | Recent Developments on Espnet Toolkit Boosted By Conformer. | Pengcheng Guo, Florian Boyer, Xuankai Chang, Tomoki Hayashi, Yosuke Higuchi, Hirofumi Inaguma, Naoyuki Kamo, Chenda Li, Daniel Garcia-Romero, Jiatong Shi, Jing Shi, Shinji Watanabe, Kun Wei, Wangyou Zhang, Yuekai Zhang |
| 2020 | ICASSP | Jhu-HLTCOE System for the Voxsrc Speaker Recognition Challenge. | Daniel Garcia-Romero, Alan McCree, David Snyder, Gregory Sell |
| 2019 | ICASSP | Speaker Recognition for Multi-speaker Conversations Using X-vectors. | David Snyder, Daniel Garcia-Romero, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur |
| 2019 | ICDAR | Script Identification using Across- and Within-Image Distribution Estimation. | Gregory Sell, David Etter, Daniel Garcia-Romero, Alan McCree |
| 2019 | Interspeech | x-Vector DNN Refinement with Full-Length Recordings for Speaker Recognition. | Daniel Garcia-Romero, David Snyder, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Speaker Recognition Benchmark Using the CHiME-5 Corpus. | Daniel Garcia-Romero, David Snyder, Shinji Watanabe, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Speaker Diarization Using Leave-One-Out Gaussian PLDA Clustering of DNN Embeddings. | Alan McCree, Gregory Sell, Daniel Garcia-Romero |
| 2019 | Interspeech | State-of-the-Art Speaker Recognition for Telephone and Video Speech: The JHU-MIT Submission for NIST SRE18. | Jess Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, Franois Grondin, Rda Dehak, Leibny Paola Garca-Perera, Daniel Povey, Pedro A. Torres-Carrasquillo, Sanjeev Khudanpur, Najim Dehak |
| 2018 | ICASSP | Audio-Visual Person Recognition in Multimedia Data From the Iarpa Janus Program. | Gregory Sell, Kevin Duh, David Snyder, Dave Etter, Daniel Garcia-Romero |
| 2018 | ICASSP | X-Vectors: Robust DNN Embeddings for Speaker Recognition. | David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, Sanjeev Khudanpur |
| 2018 | Interspeech | Diarization is Hard: Some Experiences and Lessons Learned for the JHU Team in the Inaugural DIHARD Challenge. | Gregory Sell, David Snyder, Alan McCree, Daniel Garcia-Romero, Jess Villalba, Matthew Maciejewski, Vimal Manohar, Najim Dehak, Daniel Povey, Shinji Watanabe, Sanjeev Khudanpur |
| 2018 | Interspeech | Fast Variational Bayes for Heavy-tailed PLDA Applied to i-vectors and x-vectors. | Anna Silnova, Niko Brmmer, Daniel Garcia-Romero, David Snyder, Luks Burget |
| 2017 | ICASSP | Speaker diarization using deep neural network embeddings. | Daniel Garcia-Romero, David Snyder, Gregory Sell, Daniel Povey, Alan McCree |
| 2017 | Interspeech | Extended Variability Modeling and Unsupervised Adaptation for PLDA Speaker Recognition. | Alan McCree, Gregory Sell, Daniel Garcia-Romero |
| 2017 | Interspeech | Deep Neural Network Embeddings for Text-Independent Speaker Verification. | David Snyder, Daniel Garcia-Romero, Daniel Povey, Sanjeev Khudanpur |
| 2016 | Interspeech | Stacked Long-Term TDNN for Spoken Language Recognition. | Daniel Garcia-Romero, Alan McCree |
| 2016 | Interspeech | Priors for Speaker Counting and Diarization with AHC. | Gregory Sell, Alan McCree, Daniel Garcia-Romero |
| 2015 | ASRU | Time delay deep neural network-based universal background models for speaker recognition. | David Snyder, Daniel Garcia-Romero, Daniel Povey |
| 2015 | EMNLP | Topic Identification and Discovery on Text and Speech. | Chandler May, Francis Ferraro, Alan McCree, Jonathan Wintrode, Daniel Garcia-Romero, Benjamin Van Durme |
| 2015 | ICASSP | Diarization resegmentation in the factor analysis subspace. | Gregory Sell, Daniel Garcia-Romero |
| 2015 | ICASSP | Content-based recommender systems for spoken documents. | Jonathan Wintrode, Gregory Sell, Aren Jansen, Michelle Fox, Daniel Garcia-Romero, Alan McCree |
| 2015 | Interspeech | Analysis of the second phase of the 2013-2014 i-vector machine learning challenge. | Dsir Bans, George R. Doddington, Daniel Garcia-Romero, John J. Godfrey, Craig S. Greenberg, Jaime Hernandez-Cordero, John M. Howard, Alvin F. Martin, Lisa P. Mason, Alan McCree, Douglas A. Reynolds |
| 2015 | Interspeech | Insights into deep neural networks for speaker recognition. | Daniel Garcia-Romero, Alan McCree |
| 2015 | Interspeech | DNN senone MAP multinomial i-vectors for phonotactic language recognition. | Alan McCree, Daniel Garcia-Romero |
| 2015 | Interspeech | Speaker diarization with i-vectors from DNN senone posteriors. | Gregory Sell, Daniel Garcia-Romero, Alan McCree |
| 2014 | ICASSP | Generative modelling for unsupervised score calibration. | Niko Brmmer, Daniel Garcia-Romero |
| 2014 | ICASSP | Supervised domain adaptation for I-vector based speaker recognition. | Daniel Garcia-Romero, Alan McCree |
| 2014 | ICASSP | Unsupervised idiolect discovery for speaker recognition. | Aren Jansen, Daniel Garcia-Romero, Pascal Clark, Jaime Hernandez-Cordero |
| 2014 | Interspeech | Summary and initial results of the 2013-2014 speaker recognition i-vector machine learning challenge. | Dsir Bans, George R. Doddington, Daniel Garcia-Romero, John J. Godfrey, Craig S. Greenberg, Alvin F. Martin, Alan McCree, Mark A. Przybocki, Douglas A. Reynolds |
| 2013 | Interspeech | Subspace-constrained supervector PLDA for speaker verification. | Daniel Garcia-Romero, Alan McCree |
| 2012 | ICASSP | Multicondition training of Gaussian PLDA models in i-vector space for noise and reverberation robust speaker recognition. | Daniel Garcia-Romero, Xinhui Zhou, Carol Y. Espy-Wilson |
| 2012 | ICASSP | The UMD-JHU 2011 speaker recognition system. | Daniel Garcia-Romero, Xinhui Zhou, Dmitry N. Zotkin, Balaji Vasan Srinivasan, Yuancheng Luo, Sriram Ganapathy, Samuel Thomas, Sridhar Krishna Nemala, Garimella S. V. S. Sivaram, Majid Mirbagheri, Sri Harish Reddy Mallidi, Thomas Janu, Padmanabhan Rajan, Nima Mesgarani, Mounya Elhilali, Hynek Hermansky, Shihab A. Shamma, Ramani Duraiswami |
| 2012 | Interspeech | Automatic intelligibility assessment of pathologic speech in head and neck cancer based on auditory-inspired spectro-temporal modulations. | Xinhui Zhou, Daniel Garcia-Romero, Nima Mesgarani, Maureen L. Stone, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2011 | ASRU | Linear versus mel frequency cepstral coefficients for speaker recognition. | Xinhui Zhou, Daniel Garcia-Romero, Ramani Duraiswami, Carol Y. Espy-Wilson, Shihab A. Shamma |
| 2011 | Interspeech | Analysis of i-vector Length Normalization in Speaker Recognition Systems. | Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2011 | Interspeech | Kernel Partial Least Squares for Speaker Recognition. | Balaji Vasan Srinivasan, Daniel Garcia-Romero, Dmitry N. Zotkin, Ramani Duraiswami |
| 2011 | Interspeech | Automatic Speech Codec Identification with Applications to Tampering Detection of Speech Recordings. | Jingting Zhou, Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2010 | ICASSP | Automatic acquisition device identification from speech recordings. | Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2008 | ICASSP | Language detection in audio content analysis. | Vikramjit Mitra, Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2008 | Interspeech | Intersession variability in speaker recognition: a behind the scene analysis. | Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2008 | Interspeech | Language and genre detection in audio content analysis. | Vikramjit Mitra, Daniel Garcia-Romero, Carol Y. Espy-Wilson |
| 2004 | ICASSP | Exploiting general knowledge in user-dependent fusion strategies for multimodal biometric verification. | Julian Firrez-Aguilar, Daniel Garcia-Romero, Javier Ortega-Garcia, Joaqun Gonzlez-Rodrguez |
| 2003 | ICASSP | Support vector machine fusion of idiolectal and acoustic speaker information in Spanish conversational speech. | Daniel Garcia-Romero, Julian Firrez-Aguilar, Joaqun Gonzlez-Rodrguez, Javier Ortega-Garcia |
| 2003 | Interspeech | Robust likelihood ratio estimation in Bayesian forensic speaker recognition. | Joaquin Gonzalez-Rodriguez, Daniel Garcia-Romero, Marta Garcia-Gomar, Daniel Ramos, Javier Ortega-Garcia |