Pedro J. Moreno
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
70
Venues
7
Active years
1991–2026
Best venue rank
A*
Where they publish
Papers
70 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ICDAR | Doc2Doc: Structure-Aware Generative Rendering for Bi-directional Document Translation. | Fahad Al-Otaibi, Daulet Toibazar, Renad A. Alnuaim, Ranya A. Alkahtani, Haneen A. Alhomoud, Asma A. Ibrahim, Yazeed Alharbi, Murtadha Al-Jubran, Pedro J. Moreno |
| 2024 | ICASSP | Extreme Encoder Output Frame Rate Reduction: Improving Computational Latencies of Large End-to-End Models. | Rohit Prabhavalkar, Zhong Meng, Weiran Wang, Adam Stooke, Xingyu Cai, Yanzhang He, Arun Narayanan, Dongseong Hwang, Tara N. Sainath, Pedro J. Moreno |
| 2023 | ICASSP | Modular Conformer Training for Flexible End-to-End ASR. | Kartik Audhkhasi, Brian Farris, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2023 | ICASSP | Large-Scale Language Model Rescoring on Long-Form Data. | Tongzhou Chen, Cyril Allauzen, Yinghui Huang, Daniel S. Park, David Rybach, W. Ronny Huang, Rodrigo Cabrera, Kartik Audhkhasi, Bhuvana Ramabhadran, Pedro J. Moreno, Michael Riley |
| 2022 | ICASSP | Tts4pretrain 2.0: Advancing the use of Text and Speech in ASR Pretraining with Consistency and Contrastive Losses. | Zhehuai Chen, Yu Zhang, Andrew Rosenberg, Bhuvana Ramabhadran, Pedro J. Moreno, Gary Wang |
| 2022 | ICASSP | Multilingual Second-Pass Rescoring for Automatic Speech Recognition Systems. | Neeraj Gaur, Tongzhou Chen, Ehsan Variani, Parisa Haghani, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2022 | Interspeech | Analysis of Self-Attention Head Diversity for Conformer-based Automatic Speech Recognition. | Kartik Audhkhasi, Yinghui Huang, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2022 | Interspeech | A Scalable Model Specialization Framework for Training and Inference using Submodels and its Application to Speech Model Personalization. | Fadi Biadsy, Youzheng Chen, Xia Zhang, Oleg Rybakov, Andrew Rosenberg, Pedro J. Moreno |
| 2022 | Interspeech | MAESTRO: Matched Speech Text Representations through Modality Matching. | Zhehuai Chen, Yu Zhang, Andrew Rosenberg, Bhuvana Ramabhadran, Pedro J. Moreno, Ankur Bapna, Heiga Zen |
| 2022 | Interspeech | Non-Parallel Voice Conversion for ASR Augmentation. | Gary Wang, Andrew Rosenberg, Bhuvana Ramabhadran, Fadi Biadsy, Jesse Emond, Yinghui Huang, Pedro J. Moreno |
| 2021 | ASRU | Injecting Text in Self-Supervised Speech Pretraining. | Zhehuai Chen, Yu Zhang, Andrew Rosenberg, Bhuvana Ramabhadran, Gary Wang, Pedro J. Moreno |
| 2021 | ICASSP | Extending Parrotron: An End-to-End, Speech Conversion and Speech Recognition Model for Atypical Speech. | Rohan Doshi, Youzheng Chen, Liyang Jiang, Xia Zhang, Fadi Biadsy, Bhuvana Ramabhadran, Fang Chu, Andrew Rosenberg, Pedro J. Moreno |
| 2021 | ICASSP | Mixture of Informed Experts for Multilingual Speech Recognition. | Neeraj Gaur, Brian Farris, Parisa Haghani, Isabel Leal, Pedro J. Moreno, Manasa Prasad, Bhuvana Ramabhadran, Yun Zhu |
| 2021 | Interspeech | Mixture Model Attention: Flexible Streaming and Non-Streaming Automatic Speech Recognition. | Kartik Audhkhasi, Tongzhou Chen, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2021 | Interspeech | Conformer Parrotron: A Faster and Stronger End-to-End Speech Conversion and Recognition Model for Atypical Speech. | Zhehuai Chen, Bhuvana Ramabhadran, Fadi Biadsy, Xia Zhang, Youzheng Chen, Liyang Jiang, Fang Chu, Rohan Doshi, Pedro J. Moreno |
| 2021 | Interspeech | Semi-Supervision in ASR: Sequential MixMatch and Factorized TTS-Based Augmentation. | Zhehuai Chen, Andrew Rosenberg, Yu Zhang, Heiga Zen, Mohammadreza Ghodsi, Yinghui Huang, Jesse Emond, Gary Wang, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2021 | Interspeech | Self-Adaptive Distillation for Multilingual Speech Recognition: Leveraging Student Independence. | Isabel Leal, Neeraj Gaur, Parisa Haghani, Brian Farris, Pedro J. Moreno, Manasa Prasad, Bhuvana Ramabhadran, Yun Zhu |
| 2020 | ICASSP | Neural Oracle Search on N-BEST Hypotheses. | Ehsan Variani, Tongzhou Chen, James Apfel, Bhuvana Ramabhadran, Seungji Lee, Pedro J. Moreno |
| 2020 | ICASSP | Improving Speech Recognition Using Consistent Predictions on Synthesized Speech. | Gary Wang, Andrew Rosenberg, Zhehuai Chen, Yu Zhang, Bhuvana Ramabhadran, Yonghui Wu, Pedro J. Moreno |
| 2020 | Interspeech | Improving Speech Recognition Using GAN-Based Speech Synthesis and Contrastive Unspoken Text Selection. | Zhehuai Chen, Andrew Rosenberg, Yu Zhang, Gary Wang, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2020 | Interspeech | SCADA: Stochastic, Consistent and Adversarial Data Augmentation to Improve ASR. | Gary Wang, Andrew Rosenberg, Zhehuai Chen, Yu Zhang, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2020 | Interspeech | Multilingual Speech Recognition with Self-Attention Structured Parameterization. | Yun Zhu, Parisa Haghani, Anshuman Tripathi, Bhuvana Ramabhadran, Brian Farris, Hainan Xu, Han Lu, Hasim Sak, Isabel Leal, Neeraj Gaur, Pedro J. Moreno, Qian Zhang |
| 2019 | ASRU | Speech Recognition with Augmented Synthesized Speech. | Andrew Rosenberg, Yu Zhang, Bhuvana Ramabhadran, Ye Jia, Pedro J. Moreno, Yonghui Wu, Zelin Wu |
| 2019 | ASRU | Leveraging Language ID in Multilingual End-to-End Speech Recognition. | Austin Waters, Neeraj Gaur, Parisa Haghani, Pedro J. Moreno, Zhongdi Qu |
| 2019 | Interspeech | Parrotron: An End-to-End Speech-to-Speech Conversion Model and its Applications to Hearing-Impaired Speech and Speech Separation. | Fadi Biadsy, Ron J. Weiss, Pedro J. Moreno, Dimitri Kanvesky, Ye Jia |
| 2018 | ICASSP | Modeling Non-Linguistic Contextual Signals in LSTM Language Models Via Domain Adaptation. | Min Ma, Shankar Kumar, Fadi Biadsy, Michael Nirschl, Tomas Vykruta, Pedro J. Moreno |
| 2018 | ICASSP | Hybrid Lstm-Fsmn Networks for Acoustic Modeling. | Asa Oines, Eugene Weinstein, Pedro J. Moreno |
| 2018 | ICASSP | Multilingual Speech Recognition with a Single End-to-End Model. | Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss, Bo Li, Pedro J. Moreno, Eugene Weinstein, Kanishka Rao |
| 2018 | Interspeech | Semantic Lattice Processing in Contextual Automatic Speech Recognition for Google Assistant. | Leonid Velikovich, Ian Williams, Justin Scheiner, Petar S. Aleksic, Pedro J. Moreno, Michael Riley |
| 2017 | ASRU | Syllable-based acoustic modeling with CTC-SMBR-LSTM. | Zhongdi Qu, Parisa Haghani, Eugene Weinstein, Pedro J. Moreno |
| 2016 | ICASSP | Selection and combination of hypotheses for dialectal speech recognition. | Victor Soto, Olivier Siohan, Mohamed Elfeky, Pedro J. Moreno |
| 2015 | ICASSP | Improved recognition of contact names in voice commands. | Petar S. Aleksic, Cyril Allauzen, David Elson, Aleksandar Kracun, Diego Melendo Casado, Pedro J. Moreno |
| 2015 | Interspeech | Bringing contextual information to google speech recognition. | Petar S. Aleksic, Mohammadreza Ghodsi, Assaf Hurwitz Michaely, Cyril Allauzen, Keith B. Hall, Brian Roark, David Rybach, Pedro J. Moreno |
| 2014 | ICASSP | Automatic language identification using deep neural networks. | Ignacio Lpez-Moreno, Javier Gonzalez-Dominguez, Oldrich Plchot, David Martinez, Joaquin Gonzalez-Rodriguez, Pedro J. Moreno |
| 2014 | Interspeech | Backoff inspired features for maximum entropy language models. | Fadi Biadsy, Keith B. Hall, Pedro J. Moreno, Brian Roark |
| 2014 | Interspeech | Automatic language identification using long short-term memory recurrent neural networks. | Javier Gonzalez-Dominguez, Ignacio Lpez-Moreno, Hasim Sak, Joaquin Gonzalez-Rodriguez, Pedro J. Moreno |
| 2014 | Interspeech | A big data approach to acoustic model training corpus selection. | Olga Kapralova, John Alex, Eugene Weinstein, Pedro J. Moreno, Olivier Siohan |
| 2014 | Interspeech | Asynchronous stochastic optimization for sequence training of deep neural networks: towards big data. | Erik McDermott, Georg Heigold, Pedro J. Moreno, Andrew W. Senior, Michiel Bacchiani |
| 2012 | ICASSP | Google's cross-dialect Arabic voice search. | Fadi Biadsy, Pedro J. Moreno, Martin Jansche |
| 2011 | Interspeech | Deploying Google Search by Voice in Cantonese. | Yun-Hsuan Sung, Martin Jansche, Pedro J. Moreno |
| 2010 | Interspeech | Voice search for development. | Etienne Barnard, Johan Schalkwyk, Charl Johannes van Heerden, Pedro J. Moreno |
| 2010 | Interspeech | Building transcribed speech corpora quickly and cheaply for many languages. | Thad Hughes, Kaisuke Nakajima, Linne Ha, Atul Vasu, Pedro J. Moreno, Mike LeBeau |
| 2010 | Interspeech | Search by voice in Mandarin Chinese. | Jiulong Shan, Genqing Wu, Zhihong Hu, Xiliu Tang, Martin Jansche, Pedro J. Moreno |
| 2009 | ICASSP | An audio indexing system for election video material. | Christopher Alberti, Michiel Bacchiani, Ari Bezman, Ciprian Chelba, Anastassia Drofa, Hank Liao, Pedro J. Moreno, Ted Power, Arnaud Sahuguet, Maria Shugrina, Olivier Siohan |
| 2009 | ICASSP | A factor automaton approach for the forced alignment of long speech recordings. | Pedro J. Moreno, Christopher Alberti |
| 2009 | ICASSP | Audiovisual celebrity recognition in unconstrained web videos. | Mehmet Emre Sargin, Hrishikesh B. Aradhye, Pedro J. Moreno, Ming Zhao |
| 2009 | Interspeech | A new quality measure for topic segmentation of text and speech. | Mehryar Mohri, Pedro J. Moreno, Eugene Weinstein |
| 2007 | ICASSP | Music Identification with Weighted Finite-State Transducers. | Eugene Weinstein, Pedro J. Moreno |
| 2004 | ECCV | The Kullback-Leibler Kernel as a Framework for Discriminant and Localized Representations for Visual Recognition. | Nuno Vasconcelos, Purdy Ho, Pedro J. Moreno |
| 2004 | Interspeech | SVM kernel adaptation in speaker classification and verification. | Purdy Ho, Pedro J. Moreno |
| 2003 | Interspeech | A new SVM approach to speaker identification and verification using probabilistic distance kernels. | Pedro J. Moreno, Purdy Ho |
| 2001 | Interspeech | A boosting approach for confidence scoring. | Pedro J. Moreno, Beth Logan, Bhiksha Raj |
| 2001 | SIGIR | Topic Segmentation with an Aspect Hidden Markov Model. | David M. Blei, Pedro J. Moreno |
| 2000 | ICASSP | Using the Fisher kernel method for Web audio classification. | Pedro J. Moreno, Ryan Rifkin |
| 2000 | Interspeech | An experimental study of an audio indexing system for the web. | Beth Logan, Pedro J. Moreno, Jean-Manuel Van Thong, Edward W. D. Whittaker |
| 1999 | ICASSP | On the use of support vector machines for phonetic classification. | Philip Clarkson, Pedro J. Moreno |
| 1998 | ICASSP | Factorial HMMs for acoustic modeling. | Beth Logan, Pedro J. Moreno |
| 1998 | Interspeech | A recursive algorithm for the forced alignment of very long audio segments. | Pedro J. Moreno, Christopher F. Joerg, Jean-Manuel Van Thong, Oren Glickman |
| 1997 | Interspeech | Delta vector taylor series environment compensation for speaker recognition. | Brian S. Eberman, Pedro J. Moreno |
| 1997 | Interspeech | A new algorithm for robust speech recognition: the delta vector taylor series approach. | Pedro J. Moreno, Brian S. Eberman |
| 1996 | ICASSP | A vector Taylor series approach for environment-independent speech recognition. | Pedro J. Moreno, Bhiksha Raj, Richard M. Stern |
| 1996 | Interspeech | Cepstral compensation by polynomial approximation for environment-independent speech recognition. | Bhiksha Raj, Evandro Bacci Gouva, Pedro J. Moreno, Richard M. Stern |
| 1995 | ICASSP | Multivariate-Gaussian-based cepstral normalization for robust speech recognition. | Pedro J. Moreno, Bhiksha Raj, Evandro B. Gouva, Richard M. Stern |
| 1995 | Interspeech | A unified approach for robust speech recognition. | Pedro J. Moreno, Bhiksha Raj, Richard M. Stern |
| 1994 | ICASSP | Environment normalization for robust speech recognition using direct cepstral comparison. | Fu-Hua Liu, Richard M. Stern, Alejandro Acero, Pedro J. Moreno |
| 1994 | ICASSP | Sources of degradation of speech recognition in the telephone network. | Pedro J. Moreno, Richard M. Stern |
| 1994 | Interspeech | Signal processing for robust speech recognition. | Richard M. Stern, Fu-Hua Liu, Pedro J. Moreno, Alejandro Acero |
| 1994 | NAACL | Signal Processing for Robust Speech Recognition. | Fu-Hua Liu, Pedro J. Moreno, Richard M. Stern, Alejandro Acero |
| 1992 | ICASSP | Efficient grammar processing for a spoken language translation system. | David B. Roe, Fernando C. N. Pereira, Richard Sproat, Michael D. Riley, Pedro J. Moreno, Alejandro Macarrn |
| 1991 | Interspeech | Toward a spoken language translator for restricted-domain context-free languages. | David B. Roe, Fernando Pereira, Richard Sproat, Michael D. Riley, Pedro J. Moreno, Alejandro Macarrn |