Michael Picheny
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
93
Venues
8
Active years
1983–2023
Best venue rank
A*
Where they publish
Papers
93 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | ICASSP | A Comparison of Semi-Supervised Learning Techniques for Streaming ASR at Scale. | Cal Peyser, Michael Picheny, Kyunghyun Cho, Rohit Prabhavalkar, W. Ronny Huang, Tara N. Sainath |
| 2023 | Interspeech | Improving Joint Speech-Text Representations Without Alignment. | Cal Peyser, Zhong Meng, Rohit Prabhavalkar, Andrew Rosenberg, Tara N. Sainath, Michael Picheny, Kyunghyun Cho, Ke Hu |
| 2023 | Interspeech | The MALACH Corpus: Results with End-to-End Architectures and Pretraining. | Michael Picheny, Qin Yang, Daiheng Zhang, Lining Zhang |
| 2022 | ICASSP | Towards Measuring Fairness in Speech Recognition: Casual Conversations Dataset Transcriptions. | Chunxi Liu, Michael Picheny, Leda Sari, Pooja Chitkara, Alex Xiao, Xiaohui Zhang, Mark Chou, Andres Alvarado, Caner Hazirbas, Yatharth Saraf |
| 2022 | Interspeech | Towards Disentangled Speech Representations. | Cal Peyser, W. Ronny Huang, Andrew Rosenberg, Tara N. Sainath, Michael Picheny, Kyunghyun Cho |
| 2021 | ICCV | Multimodal Clustering Networks for Self-supervised Learning from Unlabeled Videos. | Brian Chen, Andrew Rouditchenko, Kevin Duarte, Hilde Kuehne, Samuel Thomas, Angie W. Boggust, Rameswar Panda, Brian Kingsbury, Rogrio Feris, David Harwath, James R. Glass, Michael Picheny, Shih-Fu Chang |
| 2021 | Interspeech | Speak or Chat with Me: End-to-End Spoken Language Understanding System with Flexible Inputs. | Sujeong Cha, Wangrui Hou, Hyun Jung, My Phung, Michael Picheny, Hong-Kwang Jeff Kuo, Samuel Thomas, Edmilson da Silva Morais |
| 2021 | Interspeech | Cascaded Multilingual Audio-Visual Learning from Videos. | Andrew Rouditchenko, Angie W. Boggust, David Harwath, Samuel Thomas, Hilde Kuehne, Brian Chen, Rameswar Panda, Rogrio Feris, Brian Kingsbury, Michael Picheny, James R. Glass |
| 2021 | Interspeech | AVLnet: Learning Audio-Visual Language Representations from Instructional Videos. | Andrew Rouditchenko, Angie W. Boggust, David Harwath, Brian Chen, Dhiraj Joshi, Samuel Thomas, Kartik Audhkhasi, Hilde Kuehne, Rameswar Panda, Rogrio Schmidt Feris, Brian Kingsbury, Michael Picheny, Antonio Torralba, James R. Glass |
| 2020 | ICASSP | Leveraging Unpaired Text Data for Training End-To-End Speech-to-Intent Systems. | Yinghui Huang, Hong-Kwang Kuo, Samuel Thomas, Zvi Kons, Kartik Audhkhasi, Brian Kingsbury, Ron Hoory, Michael Picheny |
| 2020 | ICASSP | Improving Efficiency in Large-Scale Decentralized Distributed Training. | Wei Zhang, Xiaodong Cui, Abdullah Kayi, Mingrui Liu, Ulrich Finkler, Brian Kingsbury, George Saon, Youssef Mroueh, Alper Buyuktosunoglu, Payel Das, David S. Kung, Michael Picheny |
| 2019 | ASRU | Semi-Supervised Training and Data Augmentation for Adaptation of Automatic Broadcast News Captioning Systems. | Yinghui Huang, Samuel Thomas, Masayuki Suzuki, Zoltn Tske, Larry Sansone, Michael Picheny |
| 2019 | ASRU | Simplified LSTMS for Speech Recognition. | George Saon, Zoltn Tske, Kartik Audhkhasi, Brian Kingsbury, Michael Picheny, Samuel Thomas |
| 2019 | CVPR | Grounding Spoken Words in Unlabeled Video. | Angie W. Boggust, Kartik Audhkhasi, Dhiraj Joshi, David Harwath, Samuel Thomas, Rogrio Schmidt Feris, Danny Gutfreund, Yang Zhang, Antonio Torralba, Michael Picheny, James R. Glass |
| 2019 | ICASSP | Pre-training of Speaker Embeddings for Low-latency Speaker Change Detection in Broadcast News. | Leda Sari, Samuel Thomas, Mark Hasegawa-Johnson, Michael Picheny |
| 2019 | ICASSP | Acoustically Grounded Word Embeddings for Improved Acoustics-to-word Speech Recognition. | Shane Settle, Kartik Audhkhasi, Karen Livescu, Michael Picheny |
| 2019 | ICASSP | English Broadcast News Speech Recognition by Humans and Machines. | Samuel Thomas, Masayuki Suzuki, Yinghui Huang, Gakuto Kurata, Zoltn Tske, George Saon, Brian Kingsbury, Michael Picheny, Tom Dibert, Alice Kaiser-Schatzlein, Bern Samko |
| 2019 | ICASSP | Distributed Deep Learning Strategies for Automatic Speech Recognition. | Wei Zhang, Xiaodong Cui, Ulrich Finkler, Brian Kingsbury, George Saon, David S. Kung, Michael Picheny |
| 2019 | Interspeech | Identifying Mood Episodes Using Dialogue Features from Clinical Interviews. | Zakaria Aldeneh, Mimansa Jaiswal, Michael Picheny, Melvin G. McInnis, Emily Mower Provost |
| 2019 | Interspeech | Forget a Bit to Learn Better: Soft Forgetting for CTC-Based Automatic Speech Recognition. | Kartik Audhkhasi, George Saon, Zoltn Tske, Brian Kingsbury, Michael Picheny |
| 2019 | Interspeech | Acoustic Model Optimization Based on Evolutionary Stochastic Gradient Descent with Anchors for Automatic Speech Recognition. | Xiaodong Cui, Michael Picheny |
| 2019 | Interspeech | Large-Scale Mixed-Bandwidth Deep Neural Network Acoustic Modeling for Automatic Speech Recognition. | Khoi-Nguyen C. Mac, Xiaodong Cui, Wei Zhang, Michael Picheny |
| 2019 | Interspeech | Challenging the Boundaries of Speech Recognition: The MALACH Corpus. | Michael Picheny, Zoltn Tske, Brian Kingsbury, Kartik Audhkhasi, Xiaodong Cui, George Saon |
| 2019 | Interspeech | Detection and Recovery of OOVs for Improved English Broadcast News Captioning. | Samuel Thomas, Kartik Audhkhasi, Zoltn Tske, Yinghui Huang, Michael Picheny |
| 2019 | Interspeech | A Highly Efficient Distributed Deep Learning System for Automatic Speech Recognition. | Wei Zhang, Xiaodong Cui, Ulrich Finkler, George Saon, Abdullah Kayi, Alper Buyuktosunoglu, Brian Kingsbury, David S. Kung, Michael Picheny |
| 2018 | ICASSP | Building Competitive Direct Acoustics-to-Word Models for English Conversational Speech Recognition. | Kartik Audhkhasi, Brian Kingsbury, Bhuvana Ramabhadran, George Saon, Michael Picheny |
| 2017 | ICASSP | Training variance and performance evaluation of neural networks in speech. | Ewout van den Berg, Bhuvana Ramabhadran, Michael Picheny |
| 2017 | ICASSP | End-to-end speech recognition and keyword search on low-resource languages. | Andrew Rosenberg, Kartik Audhkhasi, Abhinav Sethy, Bhuvana Ramabhadran, Michael Picheny |
| 2017 | Interspeech | Direct Acoustics-to-Word Models for English Conversational Speech Recognition. | Kartik Audhkhasi, Bhuvana Ramabhadran, George Saon, Michael Picheny, David Nahamoo |
| 2017 | Interspeech | English Conversational Telephone Speech Recognition by Humans and Machines. | George Saon, Gakuto Kurata, Tom Sercu, Kartik Audhkhasi, Samuel Thomas, Dimitrios Dimitriadis, Xiaodong Cui, Bhuvana Ramabhadran, Michael Picheny, Lynn-Li Lim, Bergul Roomi, Phil Hall |
| 2016 | ICASSP | On the importance of event detection for ASR. | David Haws, Dimitrios Dimitriadis, George Saon, Samuel Thomas, Michael Picheny |
| 2016 | ICASSP | A comparison between deep neural nets and kernel acoustic models for speech recognition. | Zhiyun Lu, Dong Guo, Alireza Bagheri Garakani, Kuan Liu, Avner May, Aurlien Bellet, Linxi Fan, Michael Collins, Brian Kingsbury, Michael Picheny, Fei Sha |
| 2015 | ASRU | Multilingual representations for low resource speech recognition and keyword search. | Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran, Abhinav Sethy, Kartik Audhkhasi, Xiaodong Cui, Ellen Kislal, Lidia Mangu, Markus Nubaum-Thom, Michael Picheny, Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney, Mark J. F. Gales, Kate M. Knill, Anton Ragni, Haipeng Wang, Philip C. Woodland |
| 2015 | ICASSP | Order-free spoken term detection. | Lidia Mangu, George Saon, Michael Picheny, Brian Kingsbury |
| 2015 | Interspeech | The IBM 2015 English conversational telephone speech recognition system. | George Saon, Hong-Kwang Jeff Kuo, Steven J. Rennie, Michael Picheny |
| 2014 | ICASSP | Efficient spoken term detection using confusion networks. | Lidia Mangu, Brian Kingsbury, Hagen Soltau, Hong-Kwang Kuo, Michael Picheny |
| 2014 | Interspeech | Parallel deep neural network training for LVCSR tasks using blue gene/Q. | Tara N. Sainath, I-Hsin Chung, Bhuvana Ramabhadran, Michael Picheny, John A. Gunnels, Brian Kingsbury, George Saon, Vernon Austel, Upendra V. Chaudhari |
| 2014 | Interspeech | Unfolded recurrent neural networks for speech recognition. | George Saon, Hagen Soltau, Ahmad Emami, Michael Picheny |
| 2014 | SC | Parallel Deep Neural Network Training for Big Data on Blue Gene/Q. | I-Hsin Chung, Tara N. Sainath, Bhuvana Ramabhadran, Michael Picheny, John A. Gunnels, Vernon Austel, Upendra V. Chaudhari, Brian Kingsbury |
| 2013 | ASRU | Speaker adaptation of neural network acoustic models using i-vectors. | George Saon, Hagen Soltau, David Nahamoo, Michael Picheny |
| 2013 | ICASSP | Developing speech recognition systems for corpus indexing under the IARPA Babel program. | Jia Cui, Xiaodong Cui, Bhuvana Ramabhadran, Janice Kim, Brian Kingsbury, Jonathan Mamou, Lidia Mangu, Michael Picheny, Tara N. Sainath, Abhinav Sethy |
| 2013 | ICASSP | A high-performance Cantonese keyword search system. | Brian Kingsbury, Jia Cui, Xiaodong Cui, Mark J. F. Gales, Kate M. Knill, Jonathan Mamou, Lidia Mangu, David Nolden, Michael Picheny, Bhuvana Ramabhadran, Ralf Schlter, Abhinav Sethy, Philip C. Woodland |
| 2013 | ICASSP | System combination and score normalization for spoken term detection. | Jonathan Mamou, Jia Cui, Xiaodong Cui, Mark J. F. Gales, Brian Kingsbury, Kate M. Knill, Lidia Mangu, David Nolden, Michael Picheny, Bhuvana Ramabhadran, Ralf Schlter, Abhinav Sethy, Philip C. Woodland |
| 2010 | CHI | Effects of automated transcription quality on non-native speakers' comprehension in real-time computer-mediated communication. | Yingxin Pan, Danning Jiang, Lin Yao, Michael Picheny, Yong Qin |
| 2009 | ASRU | Articulatory feature detection with Support Vector Machines for integration into ASR and phone recognition. | Upendra V. Chaudhari, Michael Picheny |
| 2009 | ASRU | Improved vocabulary independent search with approximate match based on Conditional Random Fields. | Upendra V. Chaudhari, Michael Picheny |
| 2009 | ASRU | An exploration of large vocabulary tools for small vocabulary phonetic recognition. | Tara N. Sainath, Bhuvana Ramabhadran, Michael Picheny |
| 2009 | CHI | Effects of real-time transcription on non-native speaker's comprehension in computer-mediated communications. | Yingxin Pan, Danning Jiang, Michael Picheny, Yong Qin |
| 2007 | ASRU | Improvements in phone based audio search via constrained match with high order confusion estimates. | Upendra V. Chaudhari, Michael Picheny |
| 2007 | ASRU | Lattice-based Viterbi decoding techniques for speech translation. | George Saon, Michael Picheny |
| 2007 | ICASSP | Voice-Melody Transcription Under a Speech Recognition Framework. | Danning Jiang, Michael Picheny, Yong Qin |
| 2006 | ICASSP | Towards Pooled-Speaker Concatenative Text-to-Speech. | Ellen Eide, Michael Picheny |
| 2005 | Interspeech | Toward multiple-language TTS: experiments in English and Mandarin. | Raul Fernandez, Wei Zhang, Ellen Eide, Raimo Bakis, Wael Hamza, Yi Liu, Michael Picheny, John F. Pitrelli, Yong Qing, Zhiwei Shuang, Li Qin Shen |
| 2004 | Interspeech | The IBM expressive speech synthesis system. | Wael Hamza, Ellen Eide, Raimo Bakis, Michael Picheny, John F. Pitrelli |
| 2004 | NAACL | A Comparison of Rule-Based and Statistical Methods for Semantic Language Modeling and Confidence Measurement. | Ruhi Sarikaya, Yuqing Gao, Michael Picheny |
| 2003 | ICASSP | Recent improvements to the IBM trainable speech synthesis system. | Ellen Eide, Andrew Aaron, Raimo Bakis, Paul S. Cohen, Robert E. Donovan, Wael Hamza, T. Mathes, Michael Picheny, M. Polkosky, M. Smith, Mahesh Viswanathan |
| 2003 | ICASSP | Use of statistical N-gram models in natural language generation for machine translation. | Fu-Hua Liu, Liang Gu, Yuqing Gao, Michael Picheny |
| 2003 | ICASSP | Towards automatic transcription of large spoken archives - English ASR for the MALACH project. | Bhuvana Ramabhadran, Jing Huang, Michael Picheny |
| 2003 | ICASSP | Word level confidence measurement using semantic features. | Ruhi Sarikaya, Yuqing Gao, Michael Picheny |
| 2003 | Interspeech | Automated transcription and topic segmentation of large spoken archives. | Martin Franz, Bhuvana Ramabhadran, Todd Ward, Michael Picheny |
| 2003 | Interspeech | Improving statistical natural concept generation in interlingua-based speech-to-speech translation. | Liang Gu, Yuqing Gao, Michael Picheny |
| 2003 | Interspeech | Toward domain-independent conversational speech recognition. | Brian Kingsbury, Lidia Mangu, George Saon, Geoffrey Zweig, Scott Axelrod, Vaibhava Goel, Karthik Visweswariah, Michael Picheny |
| 2003 | Interspeech | Noise robustness in speech to speech translation. | Fu-Hua Liu, Yuqing Gao, Liang Gu, Michael Picheny |
| 2002 | ICASSP | Turn-Based Language Modeling for spoken dialog systems. | Ruhi Sarikaya, Yuqing Gao, Hakan Erdogan, Michael Picheny |
| 2002 | Interspeech | Semantic structured language models. | Hakan Erdogan, Ruhi Sarikaya, Yuqing Gao, Michael Picheny |
| 2002 | Interspeech | Statistical natural language generation for speech-to-speech machine translation systems. | Bowen Zhou, Yuqing Gao, Jeffrey S. Sorensen, Zijian Diao, Michael Picheny |
| 2001 | ICASSP | Speech recognition for DARPA Communicator. | Andrew Aaron, Scott Saobing Chen, Paul S. Cohen, Satya Dharanipragada, Ellen Eide, Martin Franz, Jean-Michel LeRoux, X. Luo, Benot Maison, Lidia Mangu, T. Mathes, Miroslav Novak, Peder A. Olsen, Michael Picheny, Harry Printz, Bhuvana Ramabhadran, Andrej Sakrajda, George Saon, Borivoj Tydlitt, Karthik Visweswariah, D. Yuk |
| 2001 | ICASSP | Rapid adaptation using penalized-likelihood methods. | Hakan Erdogan, Yuqing Gao, Michael Picheny |
| 2001 | ICASSP | Innovative approaches for large vocabulary name recognition. | Yuqing Gao, Bhuvana Ramabhadran, C. Julian Chen, Hakan Erdogan, Michael Picheny |
| 2001 | Interspeech | Recent advances in speech recognition system for IBM DARPA communicator. | Yuqing Gao, Hakan Erdogan, Yongxin Li, Vaibhava Goel, Michael Picheny |
| 2000 | Interspeech | Maximal rank likelihood as an optimization function for speech recognition. | Yuqing Gao, Yongxin Li, Michael Picheny |
| 2000 | Interspeech | Speed improvement of the tree-based time asynchronous search. | Miroslav Novak, Michael Picheny |
| 2000 | Interspeech | Dynamic selection of feature spaces for robust speech recognition. | Bhuvana Ramabhadran, Yuqing Gao, Michael Picheny |
| 2000 | Interspeech | Impact of bucketing on performance of linearly interpolated language models. | Karthik Visweswariah, Harry Printz, Michael Picheny |
| 1999 | ICASSP | HMM training based on quality measurement. | Yuqing Gao, Ea-Ee Jan, Mukund Padmanabhan, Michael Picheny |
| 1999 | Interspeech | Speed improvement of the time-asynchronous acoustic fast match. | Miroslav Novak, Michael Picheny |
| 1999 | Interspeech | Enhanced likelihood computation using regression. | Peter V. de Souza, Bhuvana Ramabhadran, Yuqing Gao, Michael Picheny |
| 1998 | ICASSP | Improvements in children's speech recognition performance. | Subrata K. Das, Don Nix, Michael Picheny |
| 1998 | Interspeech | Telephone band LVCSR for hearing-impaired users. | Ea-Ee Jan, Raimo Bakis, Fu-Hua Liu, Michael Picheny |
| 1998 | Interspeech | A new confidence measure based on rank-ordering subphone scores. | Qiguang Lin, Subrata K. Das, David M. Lubensky, Michael Picheny |
| 1998 | Interspeech | On variable sampling frequencies in speech recognition. | Fu-Hua Liu, Michael Picheny |
| 1997 | Interspeech | Speaker adaptation based on pre-clustering training speakers. | Yuqing Gao, Mukund Padmanabhan, Michael Picheny |
| 1997 | Interspeech | Key-phrase spotting using an integrated language model of n-grams and finite-state grammar. | Qiguang Lin, David M. Lubensky, Michael Picheny, P. Srinivasa Rao |
| 1996 | ICASSP | Speech recognition on Mandarin Call Home: a large-vocabulary, conversational, and telephone speech corpus. | Fu-Hua Liu, Michael Picheny, Patibandla Srinivasa, Michael D. Monkowski, C. Julian Chen |
| 1994 | ICASSP | Adaptation techniques for ambience and microphone compensation in the IBM Tangora speech recognition system. | Subrata K. Das, Arthur Ndas, David Nahamoo, Michael Picheny |
| 1993 | ICASSP | Influence of background noise and microphone on the performance of the IBM Tangora speech recognition system. | Subrata K. Das, Raimo Bakis, Arthur Ndas, David Nahamoo, Michael Picheny |
| 1993 | Interspeech | Word lookahead scheme for cross-word right context models in a stack decoder. | Lalit R. Bahl, Peter V. de Souza, P. S. Gopalakrishnan, David Nahamoo, Michael Picheny |
| 1991 | ICASSP | An iterative 'flip-flop' approximation of the most informative split in the construction of decision trees. | Arthur Ndas, David Nahamoo, Michael Picheny, J. Powell |
| 1991 | NAACL | Context Dependent Modeling of Phones in Continuous Speech Using Decision Trees. | Lalit R. Bahl, Peter V. de Souza, P. S. Gopalakrishnan, David Nahamoo, Michael Picheny |
| 1990 | NAACL | Automatic Phonetic Baseform Determination. | Lalit R. Bahl, Subrata K. Das, Peter DeSouza, M. Epstein, Robert L. Mercer, Bernard Mrialdo, David Nahamoo, Michael Picheny, J. Powell |
| 1987 | ICASSP | Experiments with the Tangora 20, 000 word speech recognizer. | Amir Averbuch, Lalit R. Bahl, Raimo Bakis, Peter F. Brown, Gregg Daggett, S. Das, Ken Davies, Steven V. De Gennaro, Peter V. de Souza, E. Epstein, D. Fraleigh, Frederick Jelinek, B. Lewis, Robert L. Mercer, J. Moorhead, Arthur Ndas, David Nahamoo, Michael Picheny, G. Shichman, P. Spinelli, Dirk Van Compernolle, H. Wilkens |
| 1986 | ICASSP | An IBM PC based large-vocabulary isolated-utterance speech recognizer. | Amir Averbuch, Lalit R. Bahl, Raimo Bakis, Peter F. Brown, A. G. Cole, Gregg Daggett, Subrata K. Das, Ken Davies, S. DeGennaro, Peter V. de Souza, E. Epstein, D. Fraleigh, Frederick Jelinek, Slava M. Katz, B. Lewis, Robert L. Mercer, Arthur Ndas, David Nahamoo, Michael Picheny, G. Shichman, P. Spinelli |
| 1983 | ICASSP | Recognition of isolated-word sentences from a 5000-word vocabulary office correspondence task. | Lalit R. Bahl, A. G. Cole, Frederick Jelinek, Robert L. Mercer, Arthur Ndas, David Nahamoo, Michael Picheny |