Skip to content

Michael Picheny

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

93

Venues

8

Active years

1983–2023

Best venue rank

A*

Where they publish

Papers

93 indexed papers, newest first.

YearVenueTitleAuthors
2023ICASSPA Comparison of Semi-Supervised Learning Techniques for Streaming ASR at Scale.Cal Peyser, Michael Picheny, Kyunghyun Cho, Rohit Prabhavalkar, W. Ronny Huang, Tara N. Sainath
2023InterspeechImproving Joint Speech-Text Representations Without Alignment.Cal Peyser, Zhong Meng, Rohit Prabhavalkar, Andrew Rosenberg, Tara N. Sainath, Michael Picheny, Kyunghyun Cho, Ke Hu
2023InterspeechThe MALACH Corpus: Results with End-to-End Architectures and Pretraining.Michael Picheny, Qin Yang, Daiheng Zhang, Lining Zhang
2022ICASSPTowards Measuring Fairness in Speech Recognition: Casual Conversations Dataset Transcriptions.Chunxi Liu, Michael Picheny, Leda Sari, Pooja Chitkara, Alex Xiao, Xiaohui Zhang, Mark Chou, Andres Alvarado, Caner Hazirbas, Yatharth Saraf
2022InterspeechTowards Disentangled Speech Representations.Cal Peyser, W. Ronny Huang, Andrew Rosenberg, Tara N. Sainath, Michael Picheny, Kyunghyun Cho
2021ICCVMultimodal Clustering Networks for Self-supervised Learning from Unlabeled Videos.Brian Chen, Andrew Rouditchenko, Kevin Duarte, Hilde Kuehne, Samuel Thomas, Angie W. Boggust, Rameswar Panda, Brian Kingsbury, Rogrio Feris, David Harwath, James R. Glass, Michael Picheny, Shih-Fu Chang
2021InterspeechSpeak or Chat with Me: End-to-End Spoken Language Understanding System with Flexible Inputs.Sujeong Cha, Wangrui Hou, Hyun Jung, My Phung, Michael Picheny, Hong-Kwang Jeff Kuo, Samuel Thomas, Edmilson da Silva Morais
2021InterspeechCascaded Multilingual Audio-Visual Learning from Videos.Andrew Rouditchenko, Angie W. Boggust, David Harwath, Samuel Thomas, Hilde Kuehne, Brian Chen, Rameswar Panda, Rogrio Feris, Brian Kingsbury, Michael Picheny, James R. Glass
2021InterspeechAVLnet: Learning Audio-Visual Language Representations from Instructional Videos.Andrew Rouditchenko, Angie W. Boggust, David Harwath, Brian Chen, Dhiraj Joshi, Samuel Thomas, Kartik Audhkhasi, Hilde Kuehne, Rameswar Panda, Rogrio Schmidt Feris, Brian Kingsbury, Michael Picheny, Antonio Torralba, James R. Glass
2020ICASSPLeveraging Unpaired Text Data for Training End-To-End Speech-to-Intent Systems.Yinghui Huang, Hong-Kwang Kuo, Samuel Thomas, Zvi Kons, Kartik Audhkhasi, Brian Kingsbury, Ron Hoory, Michael Picheny
2020ICASSPImproving Efficiency in Large-Scale Decentralized Distributed Training.Wei Zhang, Xiaodong Cui, Abdullah Kayi, Mingrui Liu, Ulrich Finkler, Brian Kingsbury, George Saon, Youssef Mroueh, Alper Buyuktosunoglu, Payel Das, David S. Kung, Michael Picheny
2019ASRUSemi-Supervised Training and Data Augmentation for Adaptation of Automatic Broadcast News Captioning Systems.Yinghui Huang, Samuel Thomas, Masayuki Suzuki, Zoltn Tske, Larry Sansone, Michael Picheny
2019ASRUSimplified LSTMS for Speech Recognition.George Saon, Zoltn Tske, Kartik Audhkhasi, Brian Kingsbury, Michael Picheny, Samuel Thomas
2019CVPRGrounding Spoken Words in Unlabeled Video.Angie W. Boggust, Kartik Audhkhasi, Dhiraj Joshi, David Harwath, Samuel Thomas, Rogrio Schmidt Feris, Danny Gutfreund, Yang Zhang, Antonio Torralba, Michael Picheny, James R. Glass
2019ICASSPPre-training of Speaker Embeddings for Low-latency Speaker Change Detection in Broadcast News.Leda Sari, Samuel Thomas, Mark Hasegawa-Johnson, Michael Picheny
2019ICASSPAcoustically Grounded Word Embeddings for Improved Acoustics-to-word Speech Recognition.Shane Settle, Kartik Audhkhasi, Karen Livescu, Michael Picheny
2019ICASSPEnglish Broadcast News Speech Recognition by Humans and Machines.Samuel Thomas, Masayuki Suzuki, Yinghui Huang, Gakuto Kurata, Zoltn Tske, George Saon, Brian Kingsbury, Michael Picheny, Tom Dibert, Alice Kaiser-Schatzlein, Bern Samko
2019ICASSPDistributed Deep Learning Strategies for Automatic Speech Recognition.Wei Zhang, Xiaodong Cui, Ulrich Finkler, Brian Kingsbury, George Saon, David S. Kung, Michael Picheny
2019InterspeechIdentifying Mood Episodes Using Dialogue Features from Clinical Interviews.Zakaria Aldeneh, Mimansa Jaiswal, Michael Picheny, Melvin G. McInnis, Emily Mower Provost
2019InterspeechForget a Bit to Learn Better: Soft Forgetting for CTC-Based Automatic Speech Recognition.Kartik Audhkhasi, George Saon, Zoltn Tske, Brian Kingsbury, Michael Picheny
2019InterspeechAcoustic Model Optimization Based on Evolutionary Stochastic Gradient Descent with Anchors for Automatic Speech Recognition.Xiaodong Cui, Michael Picheny
2019InterspeechLarge-Scale Mixed-Bandwidth Deep Neural Network Acoustic Modeling for Automatic Speech Recognition.Khoi-Nguyen C. Mac, Xiaodong Cui, Wei Zhang, Michael Picheny
2019InterspeechChallenging the Boundaries of Speech Recognition: The MALACH Corpus.Michael Picheny, Zoltn Tske, Brian Kingsbury, Kartik Audhkhasi, Xiaodong Cui, George Saon
2019InterspeechDetection and Recovery of OOVs for Improved English Broadcast News Captioning.Samuel Thomas, Kartik Audhkhasi, Zoltn Tske, Yinghui Huang, Michael Picheny
2019InterspeechA Highly Efficient Distributed Deep Learning System for Automatic Speech Recognition.Wei Zhang, Xiaodong Cui, Ulrich Finkler, George Saon, Abdullah Kayi, Alper Buyuktosunoglu, Brian Kingsbury, David S. Kung, Michael Picheny
2018ICASSPBuilding Competitive Direct Acoustics-to-Word Models for English Conversational Speech Recognition.Kartik Audhkhasi, Brian Kingsbury, Bhuvana Ramabhadran, George Saon, Michael Picheny
2017ICASSPTraining variance and performance evaluation of neural networks in speech.Ewout van den Berg, Bhuvana Ramabhadran, Michael Picheny
2017ICASSPEnd-to-end speech recognition and keyword search on low-resource languages.Andrew Rosenberg, Kartik Audhkhasi, Abhinav Sethy, Bhuvana Ramabhadran, Michael Picheny
2017InterspeechDirect Acoustics-to-Word Models for English Conversational Speech Recognition.Kartik Audhkhasi, Bhuvana Ramabhadran, George Saon, Michael Picheny, David Nahamoo
2017InterspeechEnglish Conversational Telephone Speech Recognition by Humans and Machines.George Saon, Gakuto Kurata, Tom Sercu, Kartik Audhkhasi, Samuel Thomas, Dimitrios Dimitriadis, Xiaodong Cui, Bhuvana Ramabhadran, Michael Picheny, Lynn-Li Lim, Bergul Roomi, Phil Hall
2016ICASSPOn the importance of event detection for ASR.David Haws, Dimitrios Dimitriadis, George Saon, Samuel Thomas, Michael Picheny
2016ICASSPA comparison between deep neural nets and kernel acoustic models for speech recognition.Zhiyun Lu, Dong Guo, Alireza Bagheri Garakani, Kuan Liu, Avner May, Aurlien Bellet, Linxi Fan, Michael Collins, Brian Kingsbury, Michael Picheny, Fei Sha
2015ASRUMultilingual representations for low resource speech recognition and keyword search.Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran, Abhinav Sethy, Kartik Audhkhasi, Xiaodong Cui, Ellen Kislal, Lidia Mangu, Markus Nubaum-Thom, Michael Picheny, Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney, Mark J. F. Gales, Kate M. Knill, Anton Ragni, Haipeng Wang, Philip C. Woodland
2015ICASSPOrder-free spoken term detection.Lidia Mangu, George Saon, Michael Picheny, Brian Kingsbury
2015InterspeechThe IBM 2015 English conversational telephone speech recognition system.George Saon, Hong-Kwang Jeff Kuo, Steven J. Rennie, Michael Picheny
2014ICASSPEfficient spoken term detection using confusion networks.Lidia Mangu, Brian Kingsbury, Hagen Soltau, Hong-Kwang Kuo, Michael Picheny
2014InterspeechParallel deep neural network training for LVCSR tasks using blue gene/Q.Tara N. Sainath, I-Hsin Chung, Bhuvana Ramabhadran, Michael Picheny, John A. Gunnels, Brian Kingsbury, George Saon, Vernon Austel, Upendra V. Chaudhari
2014InterspeechUnfolded recurrent neural networks for speech recognition.George Saon, Hagen Soltau, Ahmad Emami, Michael Picheny
2014SCParallel Deep Neural Network Training for Big Data on Blue Gene/Q.I-Hsin Chung, Tara N. Sainath, Bhuvana Ramabhadran, Michael Picheny, John A. Gunnels, Vernon Austel, Upendra V. Chaudhari, Brian Kingsbury
2013ASRUSpeaker adaptation of neural network acoustic models using i-vectors.George Saon, Hagen Soltau, David Nahamoo, Michael Picheny
2013ICASSPDeveloping speech recognition systems for corpus indexing under the IARPA Babel program.Jia Cui, Xiaodong Cui, Bhuvana Ramabhadran, Janice Kim, Brian Kingsbury, Jonathan Mamou, Lidia Mangu, Michael Picheny, Tara N. Sainath, Abhinav Sethy
2013ICASSPA high-performance Cantonese keyword search system.Brian Kingsbury, Jia Cui, Xiaodong Cui, Mark J. F. Gales, Kate M. Knill, Jonathan Mamou, Lidia Mangu, David Nolden, Michael Picheny, Bhuvana Ramabhadran, Ralf Schlter, Abhinav Sethy, Philip C. Woodland
2013ICASSPSystem combination and score normalization for spoken term detection.Jonathan Mamou, Jia Cui, Xiaodong Cui, Mark J. F. Gales, Brian Kingsbury, Kate M. Knill, Lidia Mangu, David Nolden, Michael Picheny, Bhuvana Ramabhadran, Ralf Schlter, Abhinav Sethy, Philip C. Woodland
2010CHIEffects of automated transcription quality on non-native speakers' comprehension in real-time computer-mediated communication.Yingxin Pan, Danning Jiang, Lin Yao, Michael Picheny, Yong Qin
2009ASRUArticulatory feature detection with Support Vector Machines for integration into ASR and phone recognition.Upendra V. Chaudhari, Michael Picheny
2009ASRUImproved vocabulary independent search with approximate match based on Conditional Random Fields.Upendra V. Chaudhari, Michael Picheny
2009ASRUAn exploration of large vocabulary tools for small vocabulary phonetic recognition.Tara N. Sainath, Bhuvana Ramabhadran, Michael Picheny
2009CHIEffects of real-time transcription on non-native speaker's comprehension in computer-mediated communications.Yingxin Pan, Danning Jiang, Michael Picheny, Yong Qin
2007ASRUImprovements in phone based audio search via constrained match with high order confusion estimates.Upendra V. Chaudhari, Michael Picheny
2007ASRULattice-based Viterbi decoding techniques for speech translation.George Saon, Michael Picheny
2007ICASSPVoice-Melody Transcription Under a Speech Recognition Framework.Danning Jiang, Michael Picheny, Yong Qin
2006ICASSPTowards Pooled-Speaker Concatenative Text-to-Speech.Ellen Eide, Michael Picheny
2005InterspeechToward multiple-language TTS: experiments in English and Mandarin.Raul Fernandez, Wei Zhang, Ellen Eide, Raimo Bakis, Wael Hamza, Yi Liu, Michael Picheny, John F. Pitrelli, Yong Qing, Zhiwei Shuang, Li Qin Shen
2004InterspeechThe IBM expressive speech synthesis system.Wael Hamza, Ellen Eide, Raimo Bakis, Michael Picheny, John F. Pitrelli
2004NAACLA Comparison of Rule-Based and Statistical Methods for Semantic Language Modeling and Confidence Measurement.Ruhi Sarikaya, Yuqing Gao, Michael Picheny
2003ICASSPRecent improvements to the IBM trainable speech synthesis system.Ellen Eide, Andrew Aaron, Raimo Bakis, Paul S. Cohen, Robert E. Donovan, Wael Hamza, T. Mathes, Michael Picheny, M. Polkosky, M. Smith, Mahesh Viswanathan
2003ICASSPUse of statistical N-gram models in natural language generation for machine translation.Fu-Hua Liu, Liang Gu, Yuqing Gao, Michael Picheny
2003ICASSPTowards automatic transcription of large spoken archives - English ASR for the MALACH project.Bhuvana Ramabhadran, Jing Huang, Michael Picheny
2003ICASSPWord level confidence measurement using semantic features.Ruhi Sarikaya, Yuqing Gao, Michael Picheny
2003InterspeechAutomated transcription and topic segmentation of large spoken archives.Martin Franz, Bhuvana Ramabhadran, Todd Ward, Michael Picheny
2003InterspeechImproving statistical natural concept generation in interlingua-based speech-to-speech translation.Liang Gu, Yuqing Gao, Michael Picheny
2003InterspeechToward domain-independent conversational speech recognition.Brian Kingsbury, Lidia Mangu, George Saon, Geoffrey Zweig, Scott Axelrod, Vaibhava Goel, Karthik Visweswariah, Michael Picheny
2003InterspeechNoise robustness in speech to speech translation.Fu-Hua Liu, Yuqing Gao, Liang Gu, Michael Picheny
2002ICASSPTurn-Based Language Modeling for spoken dialog systems.Ruhi Sarikaya, Yuqing Gao, Hakan Erdogan, Michael Picheny
2002InterspeechSemantic structured language models.Hakan Erdogan, Ruhi Sarikaya, Yuqing Gao, Michael Picheny
2002InterspeechStatistical natural language generation for speech-to-speech machine translation systems.Bowen Zhou, Yuqing Gao, Jeffrey S. Sorensen, Zijian Diao, Michael Picheny
2001ICASSPSpeech recognition for DARPA Communicator.Andrew Aaron, Scott Saobing Chen, Paul S. Cohen, Satya Dharanipragada, Ellen Eide, Martin Franz, Jean-Michel LeRoux, X. Luo, Benot Maison, Lidia Mangu, T. Mathes, Miroslav Novak, Peder A. Olsen, Michael Picheny, Harry Printz, Bhuvana Ramabhadran, Andrej Sakrajda, George Saon, Borivoj Tydlitt, Karthik Visweswariah, D. Yuk
2001ICASSPRapid adaptation using penalized-likelihood methods.Hakan Erdogan, Yuqing Gao, Michael Picheny
2001ICASSPInnovative approaches for large vocabulary name recognition.Yuqing Gao, Bhuvana Ramabhadran, C. Julian Chen, Hakan Erdogan, Michael Picheny
2001InterspeechRecent advances in speech recognition system for IBM DARPA communicator.Yuqing Gao, Hakan Erdogan, Yongxin Li, Vaibhava Goel, Michael Picheny
2000InterspeechMaximal rank likelihood as an optimization function for speech recognition.Yuqing Gao, Yongxin Li, Michael Picheny
2000InterspeechSpeed improvement of the tree-based time asynchronous search.Miroslav Novak, Michael Picheny
2000InterspeechDynamic selection of feature spaces for robust speech recognition.Bhuvana Ramabhadran, Yuqing Gao, Michael Picheny
2000InterspeechImpact of bucketing on performance of linearly interpolated language models.Karthik Visweswariah, Harry Printz, Michael Picheny
1999ICASSPHMM training based on quality measurement.Yuqing Gao, Ea-Ee Jan, Mukund Padmanabhan, Michael Picheny
1999InterspeechSpeed improvement of the time-asynchronous acoustic fast match.Miroslav Novak, Michael Picheny
1999InterspeechEnhanced likelihood computation using regression.Peter V. de Souza, Bhuvana Ramabhadran, Yuqing Gao, Michael Picheny
1998ICASSPImprovements in children's speech recognition performance.Subrata K. Das, Don Nix, Michael Picheny
1998InterspeechTelephone band LVCSR for hearing-impaired users.Ea-Ee Jan, Raimo Bakis, Fu-Hua Liu, Michael Picheny
1998InterspeechA new confidence measure based on rank-ordering subphone scores.Qiguang Lin, Subrata K. Das, David M. Lubensky, Michael Picheny
1998InterspeechOn variable sampling frequencies in speech recognition.Fu-Hua Liu, Michael Picheny
1997InterspeechSpeaker adaptation based on pre-clustering training speakers.Yuqing Gao, Mukund Padmanabhan, Michael Picheny
1997InterspeechKey-phrase spotting using an integrated language model of n-grams and finite-state grammar.Qiguang Lin, David M. Lubensky, Michael Picheny, P. Srinivasa Rao
1996ICASSPSpeech recognition on Mandarin Call Home: a large-vocabulary, conversational, and telephone speech corpus.Fu-Hua Liu, Michael Picheny, Patibandla Srinivasa, Michael D. Monkowski, C. Julian Chen
1994ICASSPAdaptation techniques for ambience and microphone compensation in the IBM Tangora speech recognition system.Subrata K. Das, Arthur Ndas, David Nahamoo, Michael Picheny
1993ICASSPInfluence of background noise and microphone on the performance of the IBM Tangora speech recognition system.Subrata K. Das, Raimo Bakis, Arthur Ndas, David Nahamoo, Michael Picheny
1993InterspeechWord lookahead scheme for cross-word right context models in a stack decoder.Lalit R. Bahl, Peter V. de Souza, P. S. Gopalakrishnan, David Nahamoo, Michael Picheny
1991ICASSPAn iterative 'flip-flop' approximation of the most informative split in the construction of decision trees.Arthur Ndas, David Nahamoo, Michael Picheny, J. Powell
1991NAACLContext Dependent Modeling of Phones in Continuous Speech Using Decision Trees.Lalit R. Bahl, Peter V. de Souza, P. S. Gopalakrishnan, David Nahamoo, Michael Picheny
1990NAACLAutomatic Phonetic Baseform Determination.Lalit R. Bahl, Subrata K. Das, Peter DeSouza, M. Epstein, Robert L. Mercer, Bernard Mrialdo, David Nahamoo, Michael Picheny, J. Powell
1987ICASSPExperiments with the Tangora 20, 000 word speech recognizer.Amir Averbuch, Lalit R. Bahl, Raimo Bakis, Peter F. Brown, Gregg Daggett, S. Das, Ken Davies, Steven V. De Gennaro, Peter V. de Souza, E. Epstein, D. Fraleigh, Frederick Jelinek, B. Lewis, Robert L. Mercer, J. Moorhead, Arthur Ndas, David Nahamoo, Michael Picheny, G. Shichman, P. Spinelli, Dirk Van Compernolle, H. Wilkens
1986ICASSPAn IBM PC based large-vocabulary isolated-utterance speech recognizer.Amir Averbuch, Lalit R. Bahl, Raimo Bakis, Peter F. Brown, A. G. Cole, Gregg Daggett, Subrata K. Das, Ken Davies, S. DeGennaro, Peter V. de Souza, E. Epstein, D. Fraleigh, Frederick Jelinek, Slava M. Katz, B. Lewis, Robert L. Mercer, Arthur Ndas, David Nahamoo, Michael Picheny, G. Shichman, P. Spinelli
1983ICASSPRecognition of isolated-word sentences from a 5000-word vocabulary office correspondence task.Lalit R. Bahl, A. G. Cole, Frederick Jelinek, Robert L. Mercer, Arthur Ndas, David Nahamoo, Michael Picheny