Leibny Paola Garca-Perera
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
28
Venues
8
Active years
2013–2026
Best venue rank
A*
Where they publish
Papers
28 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | AAAI | MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence. | Sonal Kumar, Simon Sedlcek, Vaibhavi Lokegaonkar, Fernando Lpez, Wenyi Yu, Nishit Anand, Hyeonggon Ryu, Lichang Chen, Maxim Plicka, Miroslav Hlavcek, William Fineas Ellingwood, Sathvik Udupa, Siyuan Hou, Allison Ferner, Sara Barahona, Cecilia Bolaos, Satish Rahi, Laura Herrera-Alarcn, Satvik Dixit, Rupali S. Patil, Soham Deshmukh, Lasha Koroshinadze, Yao Liu, Leibny Paola Garca-Perera, Eleni Zanou, Themos Stafylakis, Joon Son Chung, David Harwath, Chao Zhang, Dinesh Manocha, Alicia Lozano-Diez, Santosh Kesiraju, Sreyan Ghosh, Ramani Duraiswami |
| 2026 | EACL | CSPB: Conversational Speech Processing Benchmark for Self-supervised Speech Models. | Zili Huang, Matthew Maciejewski, Leibny Paola Garca-Perera, Shinji Watanabe, Sanjeev Khudanpur |
| 2025 | ASRU | GenVC: Self-Supervised Zero-Shot Voice Conversion. | Zexin Cai, Henry Li Xinyuan, Ashi Garg, Leibny Paola Garca-Perera, Kevin Duh, Sanjeev Khudanpur, Matthew Wiesner, Nicholas Andrews |
| 2025 | ASRU | WST: Weakly Supervised Transducer for Automatic Speech Recognition. | Dongji Gao, Chenda Liao, Changliang Liu, Matthew Wiesner, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur, Jian Wu |
| 2025 | ASRU | Rapidly Adapting to New Voice Spoofing: Few-Shot Detection of Synthesized Speech Under Distribution Shifts. | Ashi Garg, Zexin Cai, Henry Li Xinyuan, Leibny Paola Garca-Perera, Sanjeev Khudanpur, Matthew Wiesner, Nicholas Andrews |
| 2025 | ASRU | CASPER: A Large Scale Spontaneous Speech Dataset. | Cihan Xiao, Ruixing Liang, Xiangyu Zhang, Mehmet Emre Tiryaki, Veronica Bae, Lavanya Shankar, Rong Yang, Ethan Poon, Emmanuel Dupoux, Sanjeev Khudanpur, Leibny Paola Garca-Perera |
| 2025 | ASRU | Scalable Controllable Accented TTS. | Henry Li Xinyuan, Zexin Cai, Ashi Garg, Kevin Duh, Leibny Paola Garca-Perera, Sanjeev Khudanpur, Nicholas Andrews, Matthew Wiesner |
| 2025 | ICASSP | Constructing Datasets From Public Police Body Camera Footage. | Jamie Rosas-Smith, Martijn Bartelds, Ruizhe Huang, Leibny Paola Garca-Perera, Karen Livescu, Dan Jurafsky, Anjalie Field |
| 2025 | ICASSP | HLTCOE Submission to the VoicePrivacy Attacker Challenge. | Henry Li Xinyuan, Ashi Garg, Zexin Cai, Kevin Duh, Leibny Paola Garca-Perera, Sanjeev Khudanpur, Nicholas Andrews, Matthew Wiesner |
| 2024 | EMNLP | Speaking in Wavelet Domain: A Simple and Efficient Approach to Speed up Speech Diffusion Model. | Xiangyu Zhang, Daijiao Liu, Hexin Liu, Qiquan Zhang, Hanyu Meng, Leibny Paola Garca-Perera, Engsiong Chng, Lina Yao |
| 2024 | NAACL | Where are you from? Geolocating Speech and Applications to Language Identification. | Patrick Foley, Matthew Wiesner, Bismarck Bamfo Odoom, Leibny Paola Garca-Perera, Kenton Murray, Philipp Koehn |
| 2023 | ASRU | Learning From Flawed Data: Weakly Supervised Automatic Speech Recognition. | Dongji Gao, Hainan Xu, Desh Raj, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur |
| 2023 | ICASSP | Building Keyword Search System from End-To-End Asr Systems. | Ruizhe Huang, Matthew Wiesner, Leibny Paola Garca-Perera, Daniel Povey, Jan Trmal, Sanjeev Khudanpur |
| 2023 | ICDAR | Crosslingual Handwritten Text Generation Using GANs. | Chun Chieh Chang, Leibny Paola Garca-Perera, Sanjeev Khudanpur |
| 2022 | Interspeech | PHO-LID: A Unified Model Incorporating Acoustic-Phonetic and Phonotactic Information for Language Identification. | Hexin Liu, Leibny Paola Garca-Perera, Andy W. H. Khong, Suzy J. Styles, Sanjeev Khudanpur |
| 2022 | Interspeech | Updating Only Encoders Prevents Catastrophic Forgetting of End-to-End ASR Models. | Yuki Takashima, Shota Horiguchi, Shinji Watanabe, Leibny Paola Garca-Perera, Yohei Kawaguchi |
| 2021 | Interspeech | End-to-End Language Diarization for Bilingual Code-Switching Speech. | Hexin Liu, Leibny Paola Garca-Perera, Xinyi Zhang, Justin Dauwels, Andy W. H. Khong, Sanjeev Khudanpur, Suzy J. Styles |
| 2021 | Interspeech | Semi-Supervised Training with Pseudo-Labeling for End-To-End Neural Diarization. | Yuki Takashima, Yusuke Fujita, Shota Horiguchi, Shinji Watanabe, Leibny Paola Garca-Perera, Kenji Nagamatsu |
| 2021 | Interspeech | Training Hybrid Models on Noisy Transliterated Transcripts for Code-Switched Speech Recognition. | Matthew Wiesner, Mousmita Sarma, Ashish Arora, Desh Raj, Dongji Gao, Ruizhe Huang, Supreet Preet, Moris Johnson, Zikra Iqbal, Nagendra Goel, Jan Trmal, Leibny Paola Garca-Perera, Sanjeev Khudanpur |
| 2021 | Interspeech | Online Streaming End-to-End Neural Diarization Handling Overlapping Speech and Flexible Numbers of Speakers. | Yawen Xue, Shota Horiguchi, Yusuke Fujita, Yuki Takashima, Shinji Watanabe, Leibny Paola Garca-Perera, Kenji Nagamatsu |
| 2020 | ICASSP | Overlap-Aware Diarization: Resegmentation Using Neural End-to-End Overlapped Speech Detection. | Latan Bullock, Herv Bredin, Leibny Paola Garca-Perera |
| 2020 | Interspeech | End-to-End Domain-Adversarial Voice Activity Detection. | Marvin Lavechin, Marie-Philippe Gill, Ruben Bousbib, Herv Bredin, Leibny Paola Garca-Perera |
| 2019 | ICDAR | Optical Character Recognition with Chinese and Korean Character Decomposition. | Chun-Chieh Chang, Ashish Arora, Leibny Paola Garca-Perera, David Etter, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | State-of-the-Art Speaker Recognition for Telephone and Video Speech: The JHU-MIT Submission for NIST SRE18. | Jess Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, Franois Grondin, Rda Dehak, Leibny Paola Garca-Perera, Daniel Povey, Pedro A. Torres-Carrasquillo, Sanjeev Khudanpur, Najim Dehak |
| 2019 | Interspeech | Advances in Automatic Speech Recognition for Child Speech Using Factored Time Delay Neural Network. | Fei Wu, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Multi-PLDA Diarization on Children's Speech. | Jiamin Xie, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur |
| 2013 | ICASSP | Optimization of the DET curve in speaker verification under noisy conditions. | Leibny Paola Garca-Perera, Bhiksha Raj, Juan Arturo Nolazco-Flores |
| 2013 | Interspeech | Ensemble approach in speaker verification. | Leibny Paola Garca-Perera, Bhiksha Raj, Juan Arturo Nolazco-Flores |