Daniel Povey
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
156
Venues
9
Active years
1999–2026
Best venue rank
A*
Where they publish
Papers
156 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching. | Han Zhu, Wei Kang, Liyong Guo, Zengwei Yao, Fangjun Kuang, Weiji Zhuang, Zhaoqing Li, Zhifeng Han, Dong Zhang, Xin Zhang, Xingchen Song, Lingxuan Ye, Long Lin, Daniel Povey |
| 2025 | ASRU | WST: Weakly Supervised Transducer for Automatic Speech Recognition. | Dongji Gao, Chenda Liao, Changliang Liu, Matthew Wiesner, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur, Jian Wu |
| 2025 | ASRU | ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching. | Zhu Han, Wei Kang, Zengwei Yao, Liyong Guo, Fangjun Kuang, Zhaoqing Li, Weiji Zhuang, Long Lin, Daniel Povey |
| 2025 | ICLR | CR-CTC: Consistency regularization on CTC for improved speech recognition. | Zengwei Yao, Wei Kang, Xiaoyu Yang, Fangjun Kuang, Liyong Guo, Han Zhu, Zengrui Jin, Zhaoqing Li, Long Lin, Daniel Povey |
| 2024 | ECAI | SUBLLM: A Novel Efficient Architecture with Token Sequence Subsampling for LLM. | Quandong Wang, Yuxuan Yuan, Xiaoyu Yang, Ruike Zhang, Kang Zhao, Wei Liu, Jian Luan, Daniel Povey, Bin Wang |
| 2024 | ICASSP | Less Peaky and More Accurate CTC Forced Alignment by Label Priors. | Ruizhe Huang, Xiaohui Zhang, Zhaoheng Ni, Li Sun, Moto Hira, Jeff Hwang, Vimal Manohar, Vineel Pratap, Matthew Wiesner, Shinji Watanabe, Daniel Povey, Sanjeev Khudanpur |
| 2024 | ICASSP | Libriheavy: A 50, 000 Hours ASR Corpus with Punctuation Casing and Context. | Wei Kang, Xiaoyu Yang, Zengwei Yao, Fangjun Kuang, Yifan Yang, Liyong Guo, Long Lin, Daniel Povey |
| 2024 | ICASSP | PromptASR for Contextualized ASR with Controllable Style. | Xiaoyu Yang, Wei Kang, Zengwei Yao, Yifan Yang, Liyong Guo, Fangjun Kuang, Long Lin, Daniel Povey |
| 2024 | ICASSP | Towards Universal Speech Discrete Tokens: A Case Study for ASR and TTS. | Yifan Yang, Feiyu Shen, Chenpeng Du, Ziyang Ma, Kai Yu, Daniel Povey, Xie Chen |
| 2024 | ICLR | Zipformer: A faster and better encoder for automatic speech recognition. | Zengwei Yao, Liyong Guo, Xiaoyu Yang, Wei Kang, Fangjun Kuang, Yifan Yang, Zengrui Jin, Long Lin, Daniel Povey |
| 2024 | Interspeech | Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation. | Ruizhe Huang, Mahsa Yarmohammadi, Sanjeev Khudanpur, Daniel Povey |
| 2024 | Interspeech | Enhancing Neural Transducer for Multilingual ASR with Synchronized Language Diarization. | Amir Hussein, Desh Raj, Matthew Wiesner, Daniel Povey, Paola Garca, Sanjeev Khudanpur |
| 2024 | Interspeech | LibriheavyMix: A 20, 000-Hour Dataset for Single-Channel Reverberant Multi-Talker Speech Separation, ASR and Speaker Diarization. | Zengrui Jin, Yifan Yang, Mohan Shi, Wei Kang, Xiaoyu Yang, Zengwei Yao, Fangjun Kuang, Liyong Guo, Lingwei Meng, Long Lin, Yong Xu, Shi-Xiong Zhang, Daniel Povey |
| 2024 | Interspeech | Multi-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment. | Yiwen Shao, Shi-Xiong Zhang, Yong Xu, Meng Yu, Dong Yu, Daniel Povey, Sanjeev Khudanpur |
| 2023 | ASRU | Learning From Flawed Data: Weakly Supervised Automatic Speech Recognition. | Dongji Gao, Hainan Xu, Desh Raj, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur |
| 2023 | ICASSP | Predicting Multi-Codebook Vector Quantization Indexes for Knowledge Distillation. | Liyong Guo, Xiaoyu Yang, Quandong Wang, Yuxiang Kong, Zengwei Yao, Fan Cui, Fangjun Kuang, Wei Kang, Long Lin, Mingshuang Luo, Piotr Zelasko, Daniel Povey |
| 2023 | ICASSP | Building Keyword Search System from End-To-End Asr Systems. | Ruizhe Huang, Matthew Wiesner, Leibny Paola Garca-Perera, Daniel Povey, Jan Trmal, Sanjeev Khudanpur |
| 2023 | ICASSP | Fast and Parallel Decoding for Transducer. | Wei Kang, Liyong Guo, Fangjun Kuang, Long Lin, Mingshuang Luo, Zengwei Yao, Xiaoyu Yang, Piotr Zelasko, Daniel Povey |
| 2023 | ICASSP | Delay-Penalized Transducer for Low-Latency Streaming ASR. | Wei Kang, Zengwei Yao, Fangjun Kuang, Liyong Guo, Xiaoyu Yang, Long Lin, Piotr Zelasko, Daniel Povey |
| 2023 | Interspeech | Bypass Temporal Classification: Weakly Supervised Automatic Speech Recognition with Imperfect Transcripts. | Dongji Gao, Matthew Wiesner, Hainan Xu, Leibny Paola Garca, Daniel Povey, Sanjeev Khudanpur |
| 2023 | Interspeech | GPU-accelerated Guided Source Separation for Meeting Transcription. | Desh Raj, Daniel Povey, Sanjeev Khudanpur |
| 2023 | Interspeech | Blank-regularized CTC for Frame Skipping in Neural Transducer. | Yifan Yang, Xiaoyu Yang, Liyong Guo, Zengwei Yao, Wei Kang, Fangjun Kuang, Long Lin, Xie Chen, Daniel Povey |
| 2023 | Interspeech | Delay-penalized CTC Implemented Based on Finite State Transducer. | Zengwei Yao, Wei Kang, Fangjun Kuang, Liyong Guo, Xiaoyu Yang, Yifan Yang, Long Lin, Daniel Povey |
| 2022 | Interspeech | Pruned RNN-T for fast, memory-efficient ASR training. | Fangjun Kuang, Liyong Guo, Wei Kang, Long Lin, Mingshuang Luo, Zengwei Yao, Daniel Povey |
| 2021 | ICASSP | An Asynchronous WFST-Based Decoder for Automatic Speech Recognition. | Hang Lv, Zhehuai Chen, Hainan Xu, Daniel Povey, Lei Xie, Sanjeev Khudanpur |
| 2021 | ICASSP | A Parallelizable Lattice Rescoring Strategy with Neural Language Models. | Ke Li, Daniel Povey, Sanjeev Khudanpur |
| 2021 | ICASSP | Wake Word Detection with Streaming Transformers. | Yiming Wang, Hang Lv, Daniel Povey, Lei Xie, Sanjeev Khudanpur |
| 2021 | Interspeech | GigaSpeech: An Evolving, Multi-Domain ASR Corpus with 10, 000 Hours of Transcribed Audio. | Guoguo Chen, Shuzhou Chai, Guan-Bo Wang, Jiayu Du, Wei-Qiang Zhang, Chao Weng, Dan Su, Daniel Povey, Jan Trmal, Junbo Zhang, Mingjie Jin, Sanjeev Khudanpur, Shinji Watanabe, Shuaijiang Zhao, Wei Zou, Xiangang Li, Xuchen Yao, Yongqing Wang, Zhao You, Zhiyong Yan |
| 2021 | Interspeech | speechocean762: An Open-Source Non-Native English Speech Corpus for Pronunciation Assessment. | Junbo Zhang, Zhiwen Zhang, Yongqing Wang, Zhiyong Yan, Qiong Song, Yukai Huang, Ke Li, Daniel Povey, Yujun Wang |
| 2020 | ICASSP | Gpu-Accelerated Viterbi Exact Lattice Decoder for Batched Online and Offline Speech Recognition. | Hugo Braun, Justin Luitjens, Ryan Leary, Tim Kaldewey, Daniel Povey |
| 2020 | ICASSP | Speaker Diarization with Region Proposal Network. | Zili Huang, Shinji Watanabe, Yusuke Fujita, Paola Garca, Yiwen Shao, Daniel Povey, Sanjeev Khudanpur |
| 2020 | ICASSP | An Empirical Study of Transformer-Based Neural Language Model Adaptation. | Ke Li, Zhe Liu, Tianxing He, Hongzhao Huang, Fuchun Peng, Daniel Povey, Sanjeev Khudanpur |
| 2020 | ICASSP | OOV Recovery with Efficient 2nd Pass Decoding and Open-vocabulary Word-level RNNLM Rescoring for Hybrid ASR. | Xiaohui Zhang, Daniel Povey, Sanjeev Khudanpur |
| 2020 | Interspeech | An Alternative to MFCCs for ASR. | Pegah Ghahramani, Hossein Hadian, Daniel Povey, Hynek Hermansky, Sanjeev Khudanpur |
| 2020 | Interspeech | Efficient MDI Adaptation for n-Gram Language Models. | Ruizhe Huang, Ke Li, Ashish Arora, Daniel Povey, Sanjeev Khudanpur |
| 2020 | Interspeech | Neural Language Modeling with Implicit Cache Pointers. | Ke Li, Daniel Povey, Sanjeev Khudanpur |
| 2020 | Interspeech | Lattice-Free Maximum Mutual Information Training of Multilingual Speech Recognition Systems. | Srikanth R. Madikeri, Banriskhem K. Khonglah, Sibo Tong, Petr Motlcek, Herv Bourlard, Daniel Povey |
| 2020 | Interspeech | PyChain: A Fully Parallelized PyTorch Implementation of LF-MMI for End-to-End ASR. | Yiwen Shao, Yiming Wang, Daniel Povey, Sanjeev Khudanpur |
| 2020 | Interspeech | Wake Word Detection with Alignment-Free Lattice-Free MMI. | Yiming Wang, Hang Lv, Daniel Povey, Lei Xie, Sanjeev Khudanpur |
| 2019 | ASRU | Incremental Lattice Determinization for WFST Decoders. | Zhehuai Chen, Mahsa Yarmohammadi, Hainan Xu, Hang Lv, Lei Xie, Daniel Povey, Sanjeev Khudanpur |
| 2019 | ASRU | Probing the Information Encoded in X-Vectors. | Desh Raj, David Snyder, Daniel Povey, Sanjeev Khudanpur |
| 2019 | ICASSP | Speaker Recognition for Multi-speaker Conversations Using X-vectors. | David Snyder, Daniel Garcia-Romero, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur |
| 2019 | ICDAR | Using ASR Methods for OCR. | Ashish Arora, Paola Garca, Shinji Watanabe, Vimal Manohar, Yiwen Shao, Sanjeev Khudanpur, Chun-Chieh Chang, Babak Rekabdar, Bagher BabaAli, Daniel Povey, David Etter, Desh Raj, Hossein Hadian, Jan Trmal |
| 2019 | ICDAR | Optical Character Recognition with Chinese and Korean Character Decomposition. | Chun-Chieh Chang, Ashish Arora, Leibny Paola Garca-Perera, David Etter, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | x-Vector DNN Refinement with Full-Length Recordings for Speaker Recognition. | Daniel Garcia-Romero, David Snyder, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Speaker Recognition Benchmark Using the CHiME-5 Corpus. | Daniel Garcia-Romero, David Snyder, Shinji Watanabe, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Improving Emotion Identification Using Phone Posteriors in Raw Speech Waveform Based DNN. | Mousmita Sarma, Pegah Ghahremani, Daniel Povey, Nagendra Kumar Goel, Kandarpa Kumar Sarma, Najim Dehak |
| 2019 | Interspeech | The JHU Speaker Recognition System for the VOiCES 2019 Challenge. | David Snyder, Jess Villalba, Nanxin Chen, Daniel Povey, Gregory Sell, Najim Dehak, Sanjeev Khudanpur |
| 2019 | Interspeech | State-of-the-Art Speaker Recognition for Telephone and Video Speech: The JHU-MIT Submission for NIST SRE18. | Jess Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, Franois Grondin, Rda Dehak, Leibny Paola Garca-Perera, Daniel Povey, Pedro A. Torres-Carrasquillo, Sanjeev Khudanpur, Najim Dehak |
| 2019 | Interspeech | The JHU ASR System for VOiCES from a Distance Challenge 2019. | Yiming Wang, David Snyder, Hainan Xu, Vimal Manohar, Phani Sankar Nidadavolu, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Advances in Automatic Speech Recognition for Child Speech Using Factored Time Delay Neural Network. | Fei Wu, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur |
| 2019 | Interspeech | Multi-PLDA Diarization on Children's Speech. | Jiamin Xie, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur |
| 2018 | ICASSP | Semi-Supervised Training of Acoustic Models Using Lattice-Free MMI. | Vimal Manohar, Hossein Hadian, Daniel Povey, Sanjeev Khudanpur |
| 2018 | ICASSP | A Time-Restricted Self-Attention Layer for ASR. | Daniel Povey, Hossein Hadian, Pegah Ghahremani, Ke Li, Sanjeev Khudanpur |
| 2018 | ICASSP | X-Vectors: Robust DNN Embeddings for Speaker Recognition. | David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, Sanjeev Khudanpur |
| 2018 | ICASSP | A Pruned Rnnlm Lattice-Rescoring Algorithm for Automatic Speech Recognition. | Hainan Xu, Tongfei Chen, Dongji Gao, Yiming Wang, Ke Li, Nagendra Goel, Yishay Carmiel, Daniel Povey, Sanjeev Khudanpur |
| 2018 | ICASSP | Neural Network Language Modeling with Letter-Based Features and Importance Sampling. | Hainan Xu, Ke Li, Yiming Wang, Jian Wang, Shiyin Kang, Xie Chen, Daniel Povey, Sanjeev Khudanpur |
| 2018 | Interspeech | Output-Gate Projected Gated Recurrent Unit for Speech Recognition. | Gaofeng Cheng, Daniel Povey, Lu Huang, Ji Xu, Sanjeev Khudanpur, Yonghong Yan |
| 2018 | Interspeech | A GPU-based WFST Decoder with Exact Lattice Generation. | Zhehuai Chen, Justin Luitjens, Hainan Xu, Yiming Wang, Daniel Povey, Sanjeev Khudanpur |
| 2018 | Interspeech | Acoustic Modeling from Frequency Domain Representations of Speech. | Pegah Ghahremani, Hossein Hadian, Hang Lv, Daniel Povey, Sanjeev Khudanpur |
| 2018 | Interspeech | End-to-end Deep Neural Network Age Estimation. | Pegah Ghahremani, Phani Sankar Nidadavolu, Nanxin Chen, Jess Villalba, Daniel Povey, Sanjeev Khudanpur, Najim Dehak |
| 2018 | Interspeech | End-to-end Speech Recognition Using Lattice-free MMI. | Hossein Hadian, Hossein Sameti, Daniel Povey, Sanjeev Khudanpur |
| 2018 | Interspeech | Recurrent Neural Network Language Model Adaptation for Conversational Speech Recognition. | Ke Li, Hainan Xu, Yiming Wang, Daniel Povey, Sanjeev Khudanpur |
| 2018 | Interspeech | Semi-Orthogonal Low-Rank Matrix Factorization for Deep Neural Networks. | Daniel Povey, Gaofeng Cheng, Yiming Wang, Ke Li, Hainan Xu, Mahsa Yarmohammadi, Sanjeev Khudanpur |
| 2018 | Interspeech | Emotion Identification from Raw Speech Signals Using DNNs. | Mousmita Sarma, Pegah Ghahremani, Daniel Povey, Nagendra Kumar Goel, Kandarpa Kumar Sarma, Najim Dehak |
| 2018 | Interspeech | Diarization is Hard: Some Experiences and Lessons Learned for the JHU Team in the Inaugural DIHARD Challenge. | Gregory Sell, David Snyder, Alan McCree, Daniel Garcia-Romero, Jess Villalba, Matthew Maciejewski, Vimal Manohar, Najim Dehak, Daniel Povey, Shinji Watanabe, Sanjeev Khudanpur |
| 2018 | Interspeech | Self-Attentive Speaker Embeddings for Text-Independent Speaker Verification. | Yingke Zhu, Tom Ko, David Snyder, Brian Mak, Daniel Povey |
| 2017 | ASRU | Investigation of transfer learning for ASR using LF-MMI trained neural networks. | Pegah Ghahremani, Vimal Manohar, Hossein Hadian, Daniel Povey, Sanjeev Khudanpur |
| 2017 | ASRU | JHU Kaldi system for Arabic MGB-3 ASR challenge using diarization, audio-transcript alignment and transfer learning. | Vimal Manohar, Daniel Povey, Sanjeev Khudanpur |
| 2017 | ICASSP | Speaker diarization using deep neural network embeddings. | Daniel Garcia-Romero, David Snyder, Gregory Sell, Daniel Povey, Alan McCree |
| 2017 | ICASSP | A study on data augmentation of reverberant speech for robust speech recognition. | Tom Ko, Vijayaditya Peddinti, Daniel Povey, Michael L. Seltzer, Sanjeev Khudanpur |
| 2017 | Interspeech | An Exploration of Dropout with LSTMs. | Gaofeng Cheng, Vijayaditya Peddinti, Daniel Povey, Vimal Manohar, Sanjeev Khudanpur, Yonghong Yan |
| 2017 | Interspeech | Phone Duration Modeling for LVCSR Using Neural Networks. | Hossein Hadian, Daniel Povey, Hossein Sameti, Sanjeev Khudanpur |
| 2017 | Interspeech | Deep Neural Network Embeddings for Text-Independent Speaker Verification. | David Snyder, Daniel Garcia-Romero, Daniel Povey, Sanjeev Khudanpur |
| 2017 | Interspeech | The Kaldi OpenKWS System: Improving Low Resource Keyword Search. | Jan Trmal, Matthew Wiesner, Vijayaditya Peddinti, Xiaohui Zhang, Pegah Ghahremani, Yiming Wang, Vimal Manohar, Hainan Xu, Daniel Povey, Sanjeev Khudanpur |
| 2017 | Interspeech | Backstitch: Counteracting Finite-Sample Bias via Negative Steps. | Yiming Wang, Vijayaditya Peddinti, Hainan Xu, Xiaohui Zhang, Daniel Povey, Sanjeev Khudanpur |
| 2017 | Interspeech | Acoustic Data-Driven Lexicon Learning Based on a Greedy Pronunciation Selection Framework. | Xiaohui Zhang, Vimal Manohar, Daniel Povey, Sanjeev Khudanpur |
| 2016 | ICASSP | Acoustic data-driven pronunciation lexicon generation for logographic languages. | Guoguo Chen, Daniel Povey, Sanjeev Khudanpur |
| 2016 | Interspeech | Acoustic Modelling from the Signal Domain Using CNNs. | Pegah Ghahremani, Vimal Manohar, Daniel Povey, Sanjeev Khudanpur |
| 2016 | Interspeech | Far-Field ASR Without Parallel Data. | Vijayaditya Peddinti, Vimal Manohar, Yiming Wang, Daniel Povey, Sanjeev Khudanpur |
| 2016 | Interspeech | Purely Sequence-Trained Neural Networks for ASR Based on Lattice-Free MMI. | Daniel Povey, Vijayaditya Peddinti, Daniel Galvez, Pegah Ghahremani, Vimal Manohar, Xingyu Na, Yiming Wang, Sanjeev Khudanpur |
| 2015 | ASRU | JHU ASpIRE system: Robust LVCSR with TDNNS, iVector adaptation and RNN-LMS. | Vijayaditya Peddinti, Guoguo Chen, Vimal Manohar, Tom Ko, Daniel Povey, Sanjeev Khudanpur |
| 2015 | ASRU | Time delay deep neural network-based universal background models for speaker recognition. | David Snyder, Daniel Garcia-Romero, Daniel Povey |
| 2015 | EMNLP | A Coarse-Grained Model for Optimal Coupling of ASR and SMT Systems for Speech Translation. | Gaurav Kumar, Graeme W. Blackwood, Jan Trmal, Daniel Povey, Sanjeev Khudanpur |
| 2015 | ICASSP | Librispeech: An ASR corpus based on public domain audio books. | Vassil Panayotov, Guoguo Chen, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | Pronunciation and silence probability modeling for ASR. | Guoguo Chen, Hainan Xu, Minhua Wu, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | Audio augmentation for speech recognition. | Tom Ko, Vijayaditya Peddinti, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | Semi-supervised maximum mutual information training of deep neural network acoustic models. | Vimal Manohar, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | Reverberation robust acoustic modeling using i-vectors with time delay neural networks. | Vijayaditya Peddinti, Guoguo Chen, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | A time delay neural network architecture for efficient modeling of long temporal contexts. | Vijayaditya Peddinti, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | Modeling phonetic context with non-random forests for speech recognition. | Hainan Xu, Guoguo Chen, Daniel Povey, Sanjeev Khudanpur |
| 2015 | Interspeech | A diversity-penalizing ensemble training method for deep learning. | Xiaohui Zhang, Daniel Povey, Sanjeev Khudanpur |
| 2014 | ICASSP | A pitch extraction algorithm tuned for automatic speech recognition. | Pegah Ghahremani, Bagher BabaAli, Daniel Povey, Korbinian Riedhammer, Jan Trmal, Sanjeev Khudanpur |
| 2014 | ICASSP | Some insights from translating conversational telephone speech. | Gaurav Kumar, Matt Post, Daniel Povey, Sanjeev Khudanpur |
| 2014 | ICASSP | Multilingual deep neural network based acoustic modeling for rapid language adaptation. | Ngoc Thang Vu, David Imseng, Daniel Povey, Petr Motlcek, Tanja Schultz, Herv Bourlard |
| 2014 | ICASSP | Improving deep neural network acoustic models using generalized maxout networks. | Xiaohui Zhang, Jan Trmal, Daniel Povey, Sanjeev Khudanpur |
| 2014 | Interspeech | Combination of FST and CN search in spoken term detection. | Justin T. Chiu, Yun Wang, Jan Trmal, Daniel Povey, Guoguo Chen, Alexander I. Rudnicky |
| 2014 | Interspeech | Removing redundancy from lattices. | David Nolden, Hagen Soltau, Daniel Povey, Pegah Ghahremani, Lidia Mangu, Hermann Ney |
| 2013 | ASRU | Using proxies for OOV keywords in the keyword search task. | Guoguo Chen, Oguz Yilmaz, Jan Trmal, Daniel Povey, Sanjeev Khudanpur |
| 2013 | ICASSP | Quantifying the value of pronunciation lexicons for keyword search in lowresource languages. | Guoguo Chen, Sanjeev Khudanpur, Daniel Povey, Jan Trmal, David Yarowsky, Oguz Yilmaz |
| 2013 | ICASSP | Combining forward and backward search in decoding. | Mirko Hannemann, Daniel Povey, Geoffrey Zweig |
| 2013 | ICASSP | Feature and score level combination of subspace Gaussinas in LVCSR task. | Petr Motlcek, Daniel Povey, Martin Karafit |
| 2013 | Interspeech | Improved feature processing for deep neural networks. | Shakti P. Rath, Daniel Povey, Karel Vesel, Jan Cernock |
| 2013 | Interspeech | Sequence-discriminative training of deep neural networks. | Karel Vesel, Arnab Ghoshal, Luks Burget, Daniel Povey |
| 2012 | ICASSP | Generating exact lattices in the WFST framework. | Daniel Povey, Mirko Hannemann, Gilles Boulianne, Luks Burget, Arnab Ghoshal, Milos Janda, Martin Karafit, Stefan Kombrink, Petr Motlcek, Yanmin Qian, Korbinian Riedhammer, Karel Vesel, Ngoc Thang Vu |
| 2012 | ICASSP | Revisiting semi-continuous hidden Markov models. | Korbinian Riedhammer, Tobias Bocklet, Arnab Ghoshal, Daniel Povey |
| 2012 | ICASSP | Revisiting Recurrent Neural Networks for robust ASR. | Oriol Vinyals, Suman V. Ravuri, Daniel Povey |
| 2012 | ICASSP | Modeling gender dependency in the Subspace GMM framework. | Ngoc Thang Vu, Tanja Schultz, Daniel Povey |
| 2012 | Interspeech | Discriminative Training Using Non-uniform Criteria for Keyword Spotting on Spontaneous Speech. | Chao Weng, Biing-Hwang Juang, Daniel Povey |
| 2011 | ASRU | Strategies for training large scale neural network language models. | Toms Mikolov, Anoop Deoras, Daniel Povey, Luks Burget, Jan Cernock |
| 2011 | ASRU | Speaker adaptation with an Exponential Transform. | Daniel Povey, Geoffrey Zweig, Alex Acero |
| 2011 | ASRU | Strategies for using MLP based features with limited target-language training data. | Yanmin Qian, Ji Xu, Daniel Povey, Jia Liu |
| 2011 | ICASSP | A symmetrization of the Subspace Gaussian Mixture Model. | Daniel Povey, Martin Karafit, Arnab Ghoshal, Petr Schwarz |
| 2011 | ICASSP | A basis method for robust estimation of constrained MLLR. | Daniel Povey, Kaisheng Yao |
| 2011 | Interspeech | State-Level Data Borrowing for Low-Resource Speech Recognition Based on Subspace GMMs. | Yanmin Qian, Daniel Povey, Jia Liu |
| 2010 | ICASSP | Multilingual acoustic modeling for speech recognition based on subspace Gaussian Mixture Models. | Luks Burget, Petr Schwarz, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Samuel Thomas |
| 2010 | ICASSP | Speaking rate adaptation using continuous frame rate normalization. | Stephen M. Chu, Daniel Povey |
| 2010 | ICASSP | The 2009 IBM GALE Mandarin broadcast transcription system. | Stephen M. Chu, Daniel Povey, Hong-Kwang Kuo, Lidia Mangu, Shilei Zhang, Qin Shi, Yong Qin |
| 2010 | ICASSP | A novel estimation of feature-space MLLR for full-covariance models. | Arnab Ghoshal, Daniel Povey, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Ondrej Glembek, Nagendra Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas |
| 2010 | ICASSP | Approaches to automatic lexicon learning with limited training examples. | Nagendra Goel, Samuel Thomas, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Petr Schwarz |
| 2010 | ICASSP | Subspace Gaussian Mixture Models for speech recognition. | Daniel Povey, Luks Burget, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas |
| 2010 | ICASSP | The IBM 2008 GALE Arabic speech transcription system. | George Saon, Hagen Soltau, Upendra V. Chaudhari, Stephen M. Chu, Brian Kingsbury, Hong-Kwang Kuo, Lidia Mangu, Daniel Povey |
| 2010 | ICASSP | An improved consensus-like method for Minimum Bayes Risk decoding and lattice combination. | Haihua Xu, Daniel Povey, Lidia Mangu, Jie Zhu |
| 2009 | ICASSP | Large margin semi-tied covariance transforms for discriminative training. | George Saon, Daniel Povey, Hagen Soltau |
| 2009 | Interspeech | Minimum hypothesis phone error as a decoding method for speech recognition. | Haihua Xu, Daniel Povey, Jie Zhu, Guanyong Wu |
| 2008 | ICASSP | Universal background model based speech recognition. | Daniel Povey, Selina M. Chu, Balakrishnan Varadarajan |
| 2008 | ICASSP | Boosted MMI for model and feature-space discriminative training. | Daniel Povey, Dimitri Kanevsky, Brian Kingsbury, Bhuvana Ramabhadran, George Saon, Karthik Visweswariah |
| 2008 | ICASSP | Quick fmllr for speaker adaptation in speech recognition. | Balakrishnan Varadarajan, Daniel Povey, Selina M. Chu |
| 2008 | Interspeech | Monte Carlo model-space noise adaptation for speech recognition. | Daniel Povey, Brian Kingsbury |
| 2008 | Interspeech | XMLLR for improved speaker adaptation in speech recognition. | Daniel Povey, Hong-Kwang Jeff Kuo |
| 2008 | Interspeech | Fast speaker adaptive training for speech recognition. | Daniel Povey, Hong-Kwang Jeff Kuo, Hagen Soltau |
| 2008 | Interspeech | Penalty function maximization for large margin HMM training. | George Saon, Daniel Povey |
| 2007 | ICASSP | Evaluation of Proposed Modifications to MPE for Large Scale Discriminative Training. | Daniel Povey, Brian Kingsbury |
| 2007 | ICASSP | The Impact of ASR on Speech-to-Speech Translation Performance. | Ruhi Sarikaya, Bowen Zhou, Daniel Povey, Mohamed Afify, Yuqing Gao |
| 2007 | ICASSP | The IBM 2006 Gale Arabic ASR System. | Hagen Soltau, George Saon, Brian Kingsbury, Hong-Kwang Jeff Kuo, Lidia Mangu, Daniel Povey, Geoffrey Zweig |
| 2006 | ICASSP | Morpheme-Based Language Modeling for Arabic Lvcsr. | Ghinwa F. Choueiter, Daniel Povey, Stanley F. Chen, Geoffrey Zweig |
| 2006 | ICASSP | Secondary Classification for GMM Based Speaker Recognition. | Jason W. Pelecanos, Daniel Povey, Ganesh N. Ramaswamy |
| 2006 | ICASSP | Automated Quality Monitoring in the Call Center with ASR and Maximum Entropy. | Geoffrey Zweig, Olivier Siohan, George Saon, Bhuvana Ramabhadran, Daniel Povey, Lidia Mangu, Brian Kingsbury |
| 2006 | Interspeech | SPAM and full covariance for speech recognition. | Daniel Povey |
| 2006 | Interspeech | Feature and model space speaker adaptation with full covariance Gaussians. | Daniel Povey, George Saon |
| 2006 | NAACL | Automated Quality Monitoring for Call Centers using Speech and NLP Technologies. | Geoffrey Zweig, Olivier Siohan, George Saon, Bhuvana Ramabhadran, Daniel Povey, Lidia Mangu, Brian Kingsbury |
| 2005 | ICASSP | fMPE: Discriminatively Trained Features for Speech Recognition. | Daniel Povey, Brian Kingsbury, Lidia Mangu, George Saon, Hagen Soltau, Geoffrey Zweig |
| 2005 | ICASSP | The IBM 2004 Conversational Telephony System for Rich Transcription. | Hagen Soltau, Brian Kingsbury, Lidia Mangu, Daniel Povey, George Saon, Geoffrey Zweig |
| 2005 | Interspeech | Discriminatively trained features using fMPE for multi-stream audio-visual speech recognition. | Jing Huang, Daniel Povey |
| 2005 | Interspeech | Improvements to fMPE for discriminative training of features. | Daniel Povey |
| 2005 | Interspeech | Anatomy of an extremely fast LVCSR decoder. | George Saon, Daniel Povey, Geoffrey Zweig |
| 2004 | ICASSP | Phone duration modeling for LVCSR. | Daniel Povey |
| 2004 | ICASSP | Feature space Gaussianization. | George Saon, Satya Dharanipragada, Daniel Povey |
| 2003 | ICASSP | Porting: SwitchBoard to the VoiceMail task. | Mark J. F. Gales, Yuan Dong, Daniel Povey, Philip C. Woodland |
| 2003 | ICASSP | Discriminative map for acoustic model adaptation. | Daniel Povey, Philip C. Woodland, Mark J. F. Gales |
| 2003 | ICDAR | Discriminative Training for HMM-Based Offline Handwritten Character Recognition. | Roongroj Nopsuwanchai, Daniel Povey |
| 2003 | Interspeech | MMI-MAP and MPE-MAP for acoustic model adaptation. | Daniel Povey, Mark J. F. Gales, Do Yeong Kim, Philip C. Woodland |
| 2002 | ICASSP | Minimum Phone Error and I-smoothing for improved discriminative training. | Daniel Povey, Philip C. Woodland |
| 2001 | ICASSP | New features in the CU-HTK system for transcription of conversational telephone speech. | Thomas Hain, Philip C. Woodland, Gunnar Evermann, Daniel Povey |
| 2001 | ICASSP | Improved discriminative training techniques for large vocabulary continuous speech recognition. | Daniel Povey, Philip C. Woodland |
| 1999 | ICASSP | Frame discrimination training for HMMs for large vocabulary speech recognition. | Daniel Povey, Philip C. Woodland |