Skip to content

Daniel Povey

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

156

Venues

9

Active years

1999–2026

Best venue rank

A*

Where they publish

Papers

156 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching.Han Zhu, Wei Kang, Liyong Guo, Zengwei Yao, Fangjun Kuang, Weiji Zhuang, Zhaoqing Li, Zhifeng Han, Dong Zhang, Xin Zhang, Xingchen Song, Lingxuan Ye, Long Lin, Daniel Povey
2025ASRUWST: Weakly Supervised Transducer for Automatic Speech Recognition.Dongji Gao, Chenda Liao, Changliang Liu, Matthew Wiesner, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur, Jian Wu
2025ASRUZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching.Zhu Han, Wei Kang, Zengwei Yao, Liyong Guo, Fangjun Kuang, Zhaoqing Li, Weiji Zhuang, Long Lin, Daniel Povey
2025ICLRCR-CTC: Consistency regularization on CTC for improved speech recognition.Zengwei Yao, Wei Kang, Xiaoyu Yang, Fangjun Kuang, Liyong Guo, Han Zhu, Zengrui Jin, Zhaoqing Li, Long Lin, Daniel Povey
2024ECAISUBLLM: A Novel Efficient Architecture with Token Sequence Subsampling for LLM.Quandong Wang, Yuxuan Yuan, Xiaoyu Yang, Ruike Zhang, Kang Zhao, Wei Liu, Jian Luan, Daniel Povey, Bin Wang
2024ICASSPLess Peaky and More Accurate CTC Forced Alignment by Label Priors.Ruizhe Huang, Xiaohui Zhang, Zhaoheng Ni, Li Sun, Moto Hira, Jeff Hwang, Vimal Manohar, Vineel Pratap, Matthew Wiesner, Shinji Watanabe, Daniel Povey, Sanjeev Khudanpur
2024ICASSPLibriheavy: A 50, 000 Hours ASR Corpus with Punctuation Casing and Context.Wei Kang, Xiaoyu Yang, Zengwei Yao, Fangjun Kuang, Yifan Yang, Liyong Guo, Long Lin, Daniel Povey
2024ICASSPPromptASR for Contextualized ASR with Controllable Style.Xiaoyu Yang, Wei Kang, Zengwei Yao, Yifan Yang, Liyong Guo, Fangjun Kuang, Long Lin, Daniel Povey
2024ICASSPTowards Universal Speech Discrete Tokens: A Case Study for ASR and TTS.Yifan Yang, Feiyu Shen, Chenpeng Du, Ziyang Ma, Kai Yu, Daniel Povey, Xie Chen
2024ICLRZipformer: A faster and better encoder for automatic speech recognition.Zengwei Yao, Liyong Guo, Xiaoyu Yang, Wei Kang, Fangjun Kuang, Yifan Yang, Zengrui Jin, Long Lin, Daniel Povey
2024InterspeechImproving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation.Ruizhe Huang, Mahsa Yarmohammadi, Sanjeev Khudanpur, Daniel Povey
2024InterspeechEnhancing Neural Transducer for Multilingual ASR with Synchronized Language Diarization.Amir Hussein, Desh Raj, Matthew Wiesner, Daniel Povey, Paola Garca, Sanjeev Khudanpur
2024InterspeechLibriheavyMix: A 20, 000-Hour Dataset for Single-Channel Reverberant Multi-Talker Speech Separation, ASR and Speaker Diarization.Zengrui Jin, Yifan Yang, Mohan Shi, Wei Kang, Xiaoyu Yang, Zengwei Yao, Fangjun Kuang, Liyong Guo, Lingwei Meng, Long Lin, Yong Xu, Shi-Xiong Zhang, Daniel Povey
2024InterspeechMulti-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment.Yiwen Shao, Shi-Xiong Zhang, Yong Xu, Meng Yu, Dong Yu, Daniel Povey, Sanjeev Khudanpur
2023ASRULearning From Flawed Data: Weakly Supervised Automatic Speech Recognition.Dongji Gao, Hainan Xu, Desh Raj, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur
2023ICASSPPredicting Multi-Codebook Vector Quantization Indexes for Knowledge Distillation.Liyong Guo, Xiaoyu Yang, Quandong Wang, Yuxiang Kong, Zengwei Yao, Fan Cui, Fangjun Kuang, Wei Kang, Long Lin, Mingshuang Luo, Piotr Zelasko, Daniel Povey
2023ICASSPBuilding Keyword Search System from End-To-End Asr Systems.Ruizhe Huang, Matthew Wiesner, Leibny Paola Garca-Perera, Daniel Povey, Jan Trmal, Sanjeev Khudanpur
2023ICASSPFast and Parallel Decoding for Transducer.Wei Kang, Liyong Guo, Fangjun Kuang, Long Lin, Mingshuang Luo, Zengwei Yao, Xiaoyu Yang, Piotr Zelasko, Daniel Povey
2023ICASSPDelay-Penalized Transducer for Low-Latency Streaming ASR.Wei Kang, Zengwei Yao, Fangjun Kuang, Liyong Guo, Xiaoyu Yang, Long Lin, Piotr Zelasko, Daniel Povey
2023InterspeechBypass Temporal Classification: Weakly Supervised Automatic Speech Recognition with Imperfect Transcripts.Dongji Gao, Matthew Wiesner, Hainan Xu, Leibny Paola Garca, Daniel Povey, Sanjeev Khudanpur
2023InterspeechGPU-accelerated Guided Source Separation for Meeting Transcription.Desh Raj, Daniel Povey, Sanjeev Khudanpur
2023InterspeechBlank-regularized CTC for Frame Skipping in Neural Transducer.Yifan Yang, Xiaoyu Yang, Liyong Guo, Zengwei Yao, Wei Kang, Fangjun Kuang, Long Lin, Xie Chen, Daniel Povey
2023InterspeechDelay-penalized CTC Implemented Based on Finite State Transducer.Zengwei Yao, Wei Kang, Fangjun Kuang, Liyong Guo, Xiaoyu Yang, Yifan Yang, Long Lin, Daniel Povey
2022InterspeechPruned RNN-T for fast, memory-efficient ASR training.Fangjun Kuang, Liyong Guo, Wei Kang, Long Lin, Mingshuang Luo, Zengwei Yao, Daniel Povey
2021ICASSPAn Asynchronous WFST-Based Decoder for Automatic Speech Recognition.Hang Lv, Zhehuai Chen, Hainan Xu, Daniel Povey, Lei Xie, Sanjeev Khudanpur
2021ICASSPA Parallelizable Lattice Rescoring Strategy with Neural Language Models.Ke Li, Daniel Povey, Sanjeev Khudanpur
2021ICASSPWake Word Detection with Streaming Transformers.Yiming Wang, Hang Lv, Daniel Povey, Lei Xie, Sanjeev Khudanpur
2021InterspeechGigaSpeech: An Evolving, Multi-Domain ASR Corpus with 10, 000 Hours of Transcribed Audio.Guoguo Chen, Shuzhou Chai, Guan-Bo Wang, Jiayu Du, Wei-Qiang Zhang, Chao Weng, Dan Su, Daniel Povey, Jan Trmal, Junbo Zhang, Mingjie Jin, Sanjeev Khudanpur, Shinji Watanabe, Shuaijiang Zhao, Wei Zou, Xiangang Li, Xuchen Yao, Yongqing Wang, Zhao You, Zhiyong Yan
2021Interspeechspeechocean762: An Open-Source Non-Native English Speech Corpus for Pronunciation Assessment.Junbo Zhang, Zhiwen Zhang, Yongqing Wang, Zhiyong Yan, Qiong Song, Yukai Huang, Ke Li, Daniel Povey, Yujun Wang
2020ICASSPGpu-Accelerated Viterbi Exact Lattice Decoder for Batched Online and Offline Speech Recognition.Hugo Braun, Justin Luitjens, Ryan Leary, Tim Kaldewey, Daniel Povey
2020ICASSPSpeaker Diarization with Region Proposal Network.Zili Huang, Shinji Watanabe, Yusuke Fujita, Paola Garca, Yiwen Shao, Daniel Povey, Sanjeev Khudanpur
2020ICASSPAn Empirical Study of Transformer-Based Neural Language Model Adaptation.Ke Li, Zhe Liu, Tianxing He, Hongzhao Huang, Fuchun Peng, Daniel Povey, Sanjeev Khudanpur
2020ICASSPOOV Recovery with Efficient 2nd Pass Decoding and Open-vocabulary Word-level RNNLM Rescoring for Hybrid ASR.Xiaohui Zhang, Daniel Povey, Sanjeev Khudanpur
2020InterspeechAn Alternative to MFCCs for ASR.Pegah Ghahramani, Hossein Hadian, Daniel Povey, Hynek Hermansky, Sanjeev Khudanpur
2020InterspeechEfficient MDI Adaptation for n-Gram Language Models.Ruizhe Huang, Ke Li, Ashish Arora, Daniel Povey, Sanjeev Khudanpur
2020InterspeechNeural Language Modeling with Implicit Cache Pointers.Ke Li, Daniel Povey, Sanjeev Khudanpur
2020InterspeechLattice-Free Maximum Mutual Information Training of Multilingual Speech Recognition Systems.Srikanth R. Madikeri, Banriskhem K. Khonglah, Sibo Tong, Petr Motlcek, Herv Bourlard, Daniel Povey
2020InterspeechPyChain: A Fully Parallelized PyTorch Implementation of LF-MMI for End-to-End ASR.Yiwen Shao, Yiming Wang, Daniel Povey, Sanjeev Khudanpur
2020InterspeechWake Word Detection with Alignment-Free Lattice-Free MMI.Yiming Wang, Hang Lv, Daniel Povey, Lei Xie, Sanjeev Khudanpur
2019ASRUIncremental Lattice Determinization for WFST Decoders.Zhehuai Chen, Mahsa Yarmohammadi, Hainan Xu, Hang Lv, Lei Xie, Daniel Povey, Sanjeev Khudanpur
2019ASRUProbing the Information Encoded in X-Vectors.Desh Raj, David Snyder, Daniel Povey, Sanjeev Khudanpur
2019ICASSPSpeaker Recognition for Multi-speaker Conversations Using X-vectors.David Snyder, Daniel Garcia-Romero, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur
2019ICDARUsing ASR Methods for OCR.Ashish Arora, Paola Garca, Shinji Watanabe, Vimal Manohar, Yiwen Shao, Sanjeev Khudanpur, Chun-Chieh Chang, Babak Rekabdar, Bagher BabaAli, Daniel Povey, David Etter, Desh Raj, Hossein Hadian, Jan Trmal
2019ICDAROptical Character Recognition with Chinese and Korean Character Decomposition.Chun-Chieh Chang, Ashish Arora, Leibny Paola Garca-Perera, David Etter, Daniel Povey, Sanjeev Khudanpur
2019Interspeechx-Vector DNN Refinement with Full-Length Recordings for Speaker Recognition.Daniel Garcia-Romero, David Snyder, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur
2019InterspeechSpeaker Recognition Benchmark Using the CHiME-5 Corpus.Daniel Garcia-Romero, David Snyder, Shinji Watanabe, Gregory Sell, Alan McCree, Daniel Povey, Sanjeev Khudanpur
2019InterspeechImproving Emotion Identification Using Phone Posteriors in Raw Speech Waveform Based DNN.Mousmita Sarma, Pegah Ghahremani, Daniel Povey, Nagendra Kumar Goel, Kandarpa Kumar Sarma, Najim Dehak
2019InterspeechThe JHU Speaker Recognition System for the VOiCES 2019 Challenge.David Snyder, Jess Villalba, Nanxin Chen, Daniel Povey, Gregory Sell, Najim Dehak, Sanjeev Khudanpur
2019InterspeechState-of-the-Art Speaker Recognition for Telephone and Video Speech: The JHU-MIT Submission for NIST SRE18.Jess Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, Franois Grondin, Rda Dehak, Leibny Paola Garca-Perera, Daniel Povey, Pedro A. Torres-Carrasquillo, Sanjeev Khudanpur, Najim Dehak
2019InterspeechThe JHU ASR System for VOiCES from a Distance Challenge 2019.Yiming Wang, David Snyder, Hainan Xu, Vimal Manohar, Phani Sankar Nidadavolu, Daniel Povey, Sanjeev Khudanpur
2019InterspeechAdvances in Automatic Speech Recognition for Child Speech Using Factored Time Delay Neural Network.Fei Wu, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur
2019InterspeechMulti-PLDA Diarization on Children's Speech.Jiamin Xie, Leibny Paola Garca-Perera, Daniel Povey, Sanjeev Khudanpur
2018ICASSPSemi-Supervised Training of Acoustic Models Using Lattice-Free MMI.Vimal Manohar, Hossein Hadian, Daniel Povey, Sanjeev Khudanpur
2018ICASSPA Time-Restricted Self-Attention Layer for ASR.Daniel Povey, Hossein Hadian, Pegah Ghahremani, Ke Li, Sanjeev Khudanpur
2018ICASSPX-Vectors: Robust DNN Embeddings for Speaker Recognition.David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, Sanjeev Khudanpur
2018ICASSPA Pruned Rnnlm Lattice-Rescoring Algorithm for Automatic Speech Recognition.Hainan Xu, Tongfei Chen, Dongji Gao, Yiming Wang, Ke Li, Nagendra Goel, Yishay Carmiel, Daniel Povey, Sanjeev Khudanpur
2018ICASSPNeural Network Language Modeling with Letter-Based Features and Importance Sampling.Hainan Xu, Ke Li, Yiming Wang, Jian Wang, Shiyin Kang, Xie Chen, Daniel Povey, Sanjeev Khudanpur
2018InterspeechOutput-Gate Projected Gated Recurrent Unit for Speech Recognition.Gaofeng Cheng, Daniel Povey, Lu Huang, Ji Xu, Sanjeev Khudanpur, Yonghong Yan
2018InterspeechA GPU-based WFST Decoder with Exact Lattice Generation.Zhehuai Chen, Justin Luitjens, Hainan Xu, Yiming Wang, Daniel Povey, Sanjeev Khudanpur
2018InterspeechAcoustic Modeling from Frequency Domain Representations of Speech.Pegah Ghahremani, Hossein Hadian, Hang Lv, Daniel Povey, Sanjeev Khudanpur
2018InterspeechEnd-to-end Deep Neural Network Age Estimation.Pegah Ghahremani, Phani Sankar Nidadavolu, Nanxin Chen, Jess Villalba, Daniel Povey, Sanjeev Khudanpur, Najim Dehak
2018InterspeechEnd-to-end Speech Recognition Using Lattice-free MMI.Hossein Hadian, Hossein Sameti, Daniel Povey, Sanjeev Khudanpur
2018InterspeechRecurrent Neural Network Language Model Adaptation for Conversational Speech Recognition.Ke Li, Hainan Xu, Yiming Wang, Daniel Povey, Sanjeev Khudanpur
2018InterspeechSemi-Orthogonal Low-Rank Matrix Factorization for Deep Neural Networks.Daniel Povey, Gaofeng Cheng, Yiming Wang, Ke Li, Hainan Xu, Mahsa Yarmohammadi, Sanjeev Khudanpur
2018InterspeechEmotion Identification from Raw Speech Signals Using DNNs.Mousmita Sarma, Pegah Ghahremani, Daniel Povey, Nagendra Kumar Goel, Kandarpa Kumar Sarma, Najim Dehak
2018InterspeechDiarization is Hard: Some Experiences and Lessons Learned for the JHU Team in the Inaugural DIHARD Challenge.Gregory Sell, David Snyder, Alan McCree, Daniel Garcia-Romero, Jess Villalba, Matthew Maciejewski, Vimal Manohar, Najim Dehak, Daniel Povey, Shinji Watanabe, Sanjeev Khudanpur
2018InterspeechSelf-Attentive Speaker Embeddings for Text-Independent Speaker Verification.Yingke Zhu, Tom Ko, David Snyder, Brian Mak, Daniel Povey
2017ASRUInvestigation of transfer learning for ASR using LF-MMI trained neural networks.Pegah Ghahremani, Vimal Manohar, Hossein Hadian, Daniel Povey, Sanjeev Khudanpur
2017ASRUJHU Kaldi system for Arabic MGB-3 ASR challenge using diarization, audio-transcript alignment and transfer learning.Vimal Manohar, Daniel Povey, Sanjeev Khudanpur
2017ICASSPSpeaker diarization using deep neural network embeddings.Daniel Garcia-Romero, David Snyder, Gregory Sell, Daniel Povey, Alan McCree
2017ICASSPA study on data augmentation of reverberant speech for robust speech recognition.Tom Ko, Vijayaditya Peddinti, Daniel Povey, Michael L. Seltzer, Sanjeev Khudanpur
2017InterspeechAn Exploration of Dropout with LSTMs.Gaofeng Cheng, Vijayaditya Peddinti, Daniel Povey, Vimal Manohar, Sanjeev Khudanpur, Yonghong Yan
2017InterspeechPhone Duration Modeling for LVCSR Using Neural Networks.Hossein Hadian, Daniel Povey, Hossein Sameti, Sanjeev Khudanpur
2017InterspeechDeep Neural Network Embeddings for Text-Independent Speaker Verification.David Snyder, Daniel Garcia-Romero, Daniel Povey, Sanjeev Khudanpur
2017InterspeechThe Kaldi OpenKWS System: Improving Low Resource Keyword Search.Jan Trmal, Matthew Wiesner, Vijayaditya Peddinti, Xiaohui Zhang, Pegah Ghahremani, Yiming Wang, Vimal Manohar, Hainan Xu, Daniel Povey, Sanjeev Khudanpur
2017InterspeechBackstitch: Counteracting Finite-Sample Bias via Negative Steps.Yiming Wang, Vijayaditya Peddinti, Hainan Xu, Xiaohui Zhang, Daniel Povey, Sanjeev Khudanpur
2017InterspeechAcoustic Data-Driven Lexicon Learning Based on a Greedy Pronunciation Selection Framework.Xiaohui Zhang, Vimal Manohar, Daniel Povey, Sanjeev Khudanpur
2016ICASSPAcoustic data-driven pronunciation lexicon generation for logographic languages.Guoguo Chen, Daniel Povey, Sanjeev Khudanpur
2016InterspeechAcoustic Modelling from the Signal Domain Using CNNs.Pegah Ghahremani, Vimal Manohar, Daniel Povey, Sanjeev Khudanpur
2016InterspeechFar-Field ASR Without Parallel Data.Vijayaditya Peddinti, Vimal Manohar, Yiming Wang, Daniel Povey, Sanjeev Khudanpur
2016InterspeechPurely Sequence-Trained Neural Networks for ASR Based on Lattice-Free MMI.Daniel Povey, Vijayaditya Peddinti, Daniel Galvez, Pegah Ghahremani, Vimal Manohar, Xingyu Na, Yiming Wang, Sanjeev Khudanpur
2015ASRUJHU ASpIRE system: Robust LVCSR with TDNNS, iVector adaptation and RNN-LMS.Vijayaditya Peddinti, Guoguo Chen, Vimal Manohar, Tom Ko, Daniel Povey, Sanjeev Khudanpur
2015ASRUTime delay deep neural network-based universal background models for speaker recognition.David Snyder, Daniel Garcia-Romero, Daniel Povey
2015EMNLPA Coarse-Grained Model for Optimal Coupling of ASR and SMT Systems for Speech Translation.Gaurav Kumar, Graeme W. Blackwood, Jan Trmal, Daniel Povey, Sanjeev Khudanpur
2015ICASSPLibrispeech: An ASR corpus based on public domain audio books.Vassil Panayotov, Guoguo Chen, Daniel Povey, Sanjeev Khudanpur
2015InterspeechPronunciation and silence probability modeling for ASR.Guoguo Chen, Hainan Xu, Minhua Wu, Daniel Povey, Sanjeev Khudanpur
2015InterspeechAudio augmentation for speech recognition.Tom Ko, Vijayaditya Peddinti, Daniel Povey, Sanjeev Khudanpur
2015InterspeechSemi-supervised maximum mutual information training of deep neural network acoustic models.Vimal Manohar, Daniel Povey, Sanjeev Khudanpur
2015InterspeechReverberation robust acoustic modeling using i-vectors with time delay neural networks.Vijayaditya Peddinti, Guoguo Chen, Daniel Povey, Sanjeev Khudanpur
2015InterspeechA time delay neural network architecture for efficient modeling of long temporal contexts.Vijayaditya Peddinti, Daniel Povey, Sanjeev Khudanpur
2015InterspeechModeling phonetic context with non-random forests for speech recognition.Hainan Xu, Guoguo Chen, Daniel Povey, Sanjeev Khudanpur
2015InterspeechA diversity-penalizing ensemble training method for deep learning.Xiaohui Zhang, Daniel Povey, Sanjeev Khudanpur
2014ICASSPA pitch extraction algorithm tuned for automatic speech recognition.Pegah Ghahremani, Bagher BabaAli, Daniel Povey, Korbinian Riedhammer, Jan Trmal, Sanjeev Khudanpur
2014ICASSPSome insights from translating conversational telephone speech.Gaurav Kumar, Matt Post, Daniel Povey, Sanjeev Khudanpur
2014ICASSPMultilingual deep neural network based acoustic modeling for rapid language adaptation.Ngoc Thang Vu, David Imseng, Daniel Povey, Petr Motlcek, Tanja Schultz, Herv Bourlard
2014ICASSPImproving deep neural network acoustic models using generalized maxout networks.Xiaohui Zhang, Jan Trmal, Daniel Povey, Sanjeev Khudanpur
2014InterspeechCombination of FST and CN search in spoken term detection.Justin T. Chiu, Yun Wang, Jan Trmal, Daniel Povey, Guoguo Chen, Alexander I. Rudnicky
2014InterspeechRemoving redundancy from lattices.David Nolden, Hagen Soltau, Daniel Povey, Pegah Ghahremani, Lidia Mangu, Hermann Ney
2013ASRUUsing proxies for OOV keywords in the keyword search task.Guoguo Chen, Oguz Yilmaz, Jan Trmal, Daniel Povey, Sanjeev Khudanpur
2013ICASSPQuantifying the value of pronunciation lexicons for keyword search in lowresource languages.Guoguo Chen, Sanjeev Khudanpur, Daniel Povey, Jan Trmal, David Yarowsky, Oguz Yilmaz
2013ICASSPCombining forward and backward search in decoding.Mirko Hannemann, Daniel Povey, Geoffrey Zweig
2013ICASSPFeature and score level combination of subspace Gaussinas in LVCSR task.Petr Motlcek, Daniel Povey, Martin Karafit
2013InterspeechImproved feature processing for deep neural networks.Shakti P. Rath, Daniel Povey, Karel Vesel, Jan Cernock
2013InterspeechSequence-discriminative training of deep neural networks.Karel Vesel, Arnab Ghoshal, Luks Burget, Daniel Povey
2012ICASSPGenerating exact lattices in the WFST framework.Daniel Povey, Mirko Hannemann, Gilles Boulianne, Luks Burget, Arnab Ghoshal, Milos Janda, Martin Karafit, Stefan Kombrink, Petr Motlcek, Yanmin Qian, Korbinian Riedhammer, Karel Vesel, Ngoc Thang Vu
2012ICASSPRevisiting semi-continuous hidden Markov models.Korbinian Riedhammer, Tobias Bocklet, Arnab Ghoshal, Daniel Povey
2012ICASSPRevisiting Recurrent Neural Networks for robust ASR.Oriol Vinyals, Suman V. Ravuri, Daniel Povey
2012ICASSPModeling gender dependency in the Subspace GMM framework.Ngoc Thang Vu, Tanja Schultz, Daniel Povey
2012InterspeechDiscriminative Training Using Non-uniform Criteria for Keyword Spotting on Spontaneous Speech.Chao Weng, Biing-Hwang Juang, Daniel Povey
2011ASRUStrategies for training large scale neural network language models.Toms Mikolov, Anoop Deoras, Daniel Povey, Luks Burget, Jan Cernock
2011ASRUSpeaker adaptation with an Exponential Transform.Daniel Povey, Geoffrey Zweig, Alex Acero
2011ASRUStrategies for using MLP based features with limited target-language training data.Yanmin Qian, Ji Xu, Daniel Povey, Jia Liu
2011ICASSPA symmetrization of the Subspace Gaussian Mixture Model.Daniel Povey, Martin Karafit, Arnab Ghoshal, Petr Schwarz
2011ICASSPA basis method for robust estimation of constrained MLLR.Daniel Povey, Kaisheng Yao
2011InterspeechState-Level Data Borrowing for Low-Resource Speech Recognition Based on Subspace GMMs.Yanmin Qian, Daniel Povey, Jia Liu
2010ICASSPMultilingual acoustic modeling for speech recognition based on subspace Gaussian Mixture Models.Luks Burget, Petr Schwarz, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Samuel Thomas
2010ICASSPSpeaking rate adaptation using continuous frame rate normalization.Stephen M. Chu, Daniel Povey
2010ICASSPThe 2009 IBM GALE Mandarin broadcast transcription system.Stephen M. Chu, Daniel Povey, Hong-Kwang Kuo, Lidia Mangu, Shilei Zhang, Qin Shi, Yong Qin
2010ICASSPA novel estimation of feature-space MLLR for full-covariance models.Arnab Ghoshal, Daniel Povey, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Ondrej Glembek, Nagendra Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas
2010ICASSPApproaches to automatic lexicon learning with limited training examples.Nagendra Goel, Samuel Thomas, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Petr Schwarz
2010ICASSPSubspace Gaussian Mixture Models for speech recognition.Daniel Povey, Luks Burget, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas
2010ICASSPThe IBM 2008 GALE Arabic speech transcription system.George Saon, Hagen Soltau, Upendra V. Chaudhari, Stephen M. Chu, Brian Kingsbury, Hong-Kwang Kuo, Lidia Mangu, Daniel Povey
2010ICASSPAn improved consensus-like method for Minimum Bayes Risk decoding and lattice combination.Haihua Xu, Daniel Povey, Lidia Mangu, Jie Zhu
2009ICASSPLarge margin semi-tied covariance transforms for discriminative training.George Saon, Daniel Povey, Hagen Soltau
2009InterspeechMinimum hypothesis phone error as a decoding method for speech recognition.Haihua Xu, Daniel Povey, Jie Zhu, Guanyong Wu
2008ICASSPUniversal background model based speech recognition.Daniel Povey, Selina M. Chu, Balakrishnan Varadarajan
2008ICASSPBoosted MMI for model and feature-space discriminative training.Daniel Povey, Dimitri Kanevsky, Brian Kingsbury, Bhuvana Ramabhadran, George Saon, Karthik Visweswariah
2008ICASSPQuick fmllr for speaker adaptation in speech recognition.Balakrishnan Varadarajan, Daniel Povey, Selina M. Chu
2008InterspeechMonte Carlo model-space noise adaptation for speech recognition.Daniel Povey, Brian Kingsbury
2008InterspeechXMLLR for improved speaker adaptation in speech recognition.Daniel Povey, Hong-Kwang Jeff Kuo
2008InterspeechFast speaker adaptive training for speech recognition.Daniel Povey, Hong-Kwang Jeff Kuo, Hagen Soltau
2008InterspeechPenalty function maximization for large margin HMM training.George Saon, Daniel Povey
2007ICASSPEvaluation of Proposed Modifications to MPE for Large Scale Discriminative Training.Daniel Povey, Brian Kingsbury
2007ICASSPThe Impact of ASR on Speech-to-Speech Translation Performance.Ruhi Sarikaya, Bowen Zhou, Daniel Povey, Mohamed Afify, Yuqing Gao
2007ICASSPThe IBM 2006 Gale Arabic ASR System.Hagen Soltau, George Saon, Brian Kingsbury, Hong-Kwang Jeff Kuo, Lidia Mangu, Daniel Povey, Geoffrey Zweig
2006ICASSPMorpheme-Based Language Modeling for Arabic Lvcsr.Ghinwa F. Choueiter, Daniel Povey, Stanley F. Chen, Geoffrey Zweig
2006ICASSPSecondary Classification for GMM Based Speaker Recognition.Jason W. Pelecanos, Daniel Povey, Ganesh N. Ramaswamy
2006ICASSPAutomated Quality Monitoring in the Call Center with ASR and Maximum Entropy.Geoffrey Zweig, Olivier Siohan, George Saon, Bhuvana Ramabhadran, Daniel Povey, Lidia Mangu, Brian Kingsbury
2006InterspeechSPAM and full covariance for speech recognition.Daniel Povey
2006InterspeechFeature and model space speaker adaptation with full covariance Gaussians.Daniel Povey, George Saon
2006NAACLAutomated Quality Monitoring for Call Centers using Speech and NLP Technologies.Geoffrey Zweig, Olivier Siohan, George Saon, Bhuvana Ramabhadran, Daniel Povey, Lidia Mangu, Brian Kingsbury
2005ICASSPfMPE: Discriminatively Trained Features for Speech Recognition.Daniel Povey, Brian Kingsbury, Lidia Mangu, George Saon, Hagen Soltau, Geoffrey Zweig
2005ICASSPThe IBM 2004 Conversational Telephony System for Rich Transcription.Hagen Soltau, Brian Kingsbury, Lidia Mangu, Daniel Povey, George Saon, Geoffrey Zweig
2005InterspeechDiscriminatively trained features using fMPE for multi-stream audio-visual speech recognition.Jing Huang, Daniel Povey
2005InterspeechImprovements to fMPE for discriminative training of features.Daniel Povey
2005InterspeechAnatomy of an extremely fast LVCSR decoder.George Saon, Daniel Povey, Geoffrey Zweig
2004ICASSPPhone duration modeling for LVCSR.Daniel Povey
2004ICASSPFeature space Gaussianization.George Saon, Satya Dharanipragada, Daniel Povey
2003ICASSPPorting: SwitchBoard to the VoiceMail task.Mark J. F. Gales, Yuan Dong, Daniel Povey, Philip C. Woodland
2003ICASSPDiscriminative map for acoustic model adaptation.Daniel Povey, Philip C. Woodland, Mark J. F. Gales
2003ICDARDiscriminative Training for HMM-Based Offline Handwritten Character Recognition.Roongroj Nopsuwanchai, Daniel Povey
2003InterspeechMMI-MAP and MPE-MAP for acoustic model adaptation.Daniel Povey, Mark J. F. Gales, Do Yeong Kim, Philip C. Woodland
2002ICASSPMinimum Phone Error and I-smoothing for improved discriminative training.Daniel Povey, Philip C. Woodland
2001ICASSPNew features in the CU-HTK system for transcription of conversational telephone speech.Thomas Hain, Philip C. Woodland, Gunnar Evermann, Daniel Povey
2001ICASSPImproved discriminative training techniques for large vocabulary continuous speech recognition.Daniel Povey, Philip C. Woodland
1999ICASSPFrame discrimination training for HMMs for large vocabulary speech recognition.Daniel Povey, Philip C. Woodland