Ehsan Variani
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
27
Venues
4
Active years
2011–2023
Best venue rank
A
Where they publish
Papers
27 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | ICASSP | JEIT: Joint End-to-End Model and Internal Language Model Training for Speech Recognition. | Zhong Meng, Weiran Wang, Rohit Prabhavalkar, Tara N. Sainath, Tongzhou Chen, Ehsan Variani, Yu Zhang, Bo Li, Andrew Rosenberg, Bhuvana Ramabhadran |
| 2023 | ICASSP | Alignment Entropy Regularization. | Ehsan Variani, Ke Wu, David Rybach, Cyril Allauzen, Michael Riley |
| 2023 | ICASSP | Last: Scalable Lattice-Based Speech Modelling in Jax. | Ke Wu, Ehsan Variani, Tom Bagby, Michael Riley |
| 2022 | ICASSP | Multilingual Second-Pass Rescoring for Automatic Speech Recognition Systems. | Neeraj Gaur, Tongzhou Chen, Ehsan Variani, Parisa Haghani, Bhuvana Ramabhadran, Pedro J. Moreno |
| 2022 | Interspeech | UserLibri: A Dataset for ASR Personalization Using Only Text. | Theresa Breiner, Swaroop Ramaswamy, Ehsan Variani, Shefali Garg, Rajiv Mathews, Khe Chai Sim, Kilol Gupta, Mingqing Chen, Lara McConnaughey |
| 2022 | Interspeech | On Adaptive Weight Interpolation of the Hybrid Autoregressive Transducer. | Ehsan Variani, Michael Riley, David Rybach, Cyril Allauzen, Tongzhou Chen, Bhuvana Ramabhadran |
| 2022 | Interspeech | Improving Rare Word Recognition with LM-aware MWER Training. | Weiran Wang, Tongzhou Chen, Tara N. Sainath, Ehsan Variani, Rohit Prabhavalkar, W. Ronny Huang, Bhuvana Ramabhadran, Neeraj Gaur, Sepand Mavandadi, Cal Peyser, Trevor Strohman, Yanzhang He, David Rybach |
| 2021 | ICASSP | Cascaded Encoders for Unifying Streaming and Non-Streaming ASR. | Arun Narayanan, Tara N. Sainath, Ruoming Pang, Jiahui Yu, Chung-Cheng Chiu, Rohit Prabhavalkar, Ehsan Variani, Trevor Strohman |
| 2021 | Interspeech | A Hybrid Seq-2-Seq ASR Design for On-Device and Server Applications. | Cyril Allauzen, Ehsan Variani, Michael Riley, David Rybach, Hao Zhang |
| 2021 | Interspeech | An Efficient Streaming Non-Recurrent On-Device End-to-End Model with Improvements to Rare-Word Modeling. | Tara N. Sainath, Yanzhang He, Arun Narayanan, Rami Botros, Ruoming Pang, David Rybach, Cyril Allauzen, Ehsan Variani, James Qin, Quoc-Nam Le-The, Shuo-Yiin Chang, Bo Li, Anmol Gulati, Jiahui Yu, Chung-Cheng Chiu, Diamantino Caseiro, Wei Li, Qiao Liang, Pat Rondon |
| 2020 | ICASSP | Neural Oracle Search on N-BEST Hypotheses. | Ehsan Variani, Tongzhou Chen, James Apfel, Bhuvana Ramabhadran, Seungji Lee, Pedro J. Moreno |
| 2020 | ICASSP | Hybrid Autoregressive Transducer (HAT). | Ehsan Variani, David Rybach, Cyril Allauzen, Michael Riley |
| 2019 | ASRU | A Density Ratio Approach to Language Model Fusion in End-to-End Automatic Speech Recognition. | Erik McDermott, Hasim Sak, Ehsan Variani |
| 2019 | ICASSP | West: Word Encoded Sequence Transducers. | Ehsan Variani, Ananda Theertha Suresh, Mitchel Weintraub |
| 2018 | ICASSP | Sampled Connectionist Temporal Classification. | Ehsan Variani, Tom Bagby, Kamel Lahouel, Erik McDermott, Michiel Bacchiani |
| 2018 | Interspeech | Efficient Implementation of the Room Simulator for Training Deep Neural Network Acoustic Models. | Chanwoo Kim, Ehsan Variani, Arun Narayanan, Michiel Bacchiani |
| 2017 | Interspeech | Acoustic Modeling for Google Home. | Bo Li, Tara N. Sainath, Arun Narayanan, Joe Caroselli, Michiel Bacchiani, Ananya Misra, Izhak Shafran, Hasim Sak, Golan Pundak, Kean K. Chin, Khe Chai Sim, Ron J. Weiss, Kevin W. Wilson, Ehsan Variani, Chanwoo Kim, Olivier Siohan, Mitchel Weintraub, Erik McDermott, Richard Rose, Matt Shannon |
| 2017 | Interspeech | End-to-End Training of Acoustic Models for Large Vocabulary Continuous Speech Recognition with TensorFlow. | Ehsan Variani, Tom Bagby, Erik McDermott, Michiel Bacchiani |
| 2016 | Interspeech | Reducing the Computational Complexity of Multimicrophone Acoustic Models with Integrated Feature Extraction. | Tara N. Sainath, Arun Narayanan, Ron J. Weiss, Ehsan Variani, Kevin W. Wilson, Michiel Bacchiani, Izhak Shafran |
| 2016 | Interspeech | Complex Linear Projection (CLP): A Discriminative Approach to Joint Feature Extraction and Acoustic Modeling. | Ehsan Variani, Tara N. Sainath, Izhak Shafran, Michiel Bacchiani |
| 2015 | ICASSP | A Gaussian Mixture Model layer jointly optimized with discriminative features within a Deep Neural Network architecture. | Ehsan Variani, Erik McDermott, Georg Heigold |
| 2015 | ISIT | NON-adaptive policies for 20 questions target localization. | Ehsan Variani, Kamel Lahouel, Avner Bar-Hen, Bruno Jedynak |
| 2014 | ICASSP | Deep neural networks for small footprint text-dependent speaker verification. | Ehsan Variani, Xin Lei, Erik McDermott, Ignacio Lpez-Moreno, Javier Gonzalez-Dominguez |
| 2013 | ICASSP | Mean temporal distance: Predicting ASR error from temporal properties of speech signal. | Hynek Hermansky, Ehsan Variani, Vijayaditya Peddinti |
| 2013 | Interspeech | Multi-stream recognition of noisy speech with performance monitoring. | Ehsan Variani, Feipeng Li, Hynek Hermansky |
| 2012 | Interspeech | Estimating Classifier Performance in Unknown Noise. | Ehsan Variani, Hynek Hermansky |
| 2011 | Interspeech | VTLN in the MFCC Domain: Band-Limited versus Local Interpolation. | Ehsan Variani, Thomas Schaaf |