Samuel Thomas
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
105
Venues
13
Active years
2007–2026
Best venue rank
A*
Where they publish
Papers
105 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ISPASS | Characterizing and Optimizing Cache Placement for Secure Memory Metadata. | Samuel Thomas, Blake Cragen, R. Iris Bahar, Tamara Silbergleit Lehman |
| 2025 | ASRU | Omni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM? | Andrew Rouditchenko, Saurabhchand Bhati, Edson Araujo, Samuel Thomas, Hilde Kuehne, Rogrio Feris, James R. Glass |
| 2025 | ASRU | Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities. | George Saon, Avihu Dekel, Alexander Brooks, Tohru Nagano, Abraham Daniels, Aharon Satt, Ashish R. Mittal, Brian Kingsbury, David Haws, Edmilson da Silva Morais, Gakuto Kurata, Hagai Aronowitz, Ibrahim Ibrahim, Hong-Kwang Kuo, Kate Soule, Luis A. Lastras, Masayuki Suzuki, Ron Hoory, Samuel Thomas, Sashi Novitasari, Takashi Fukuda, Vishal Sunder, Xiaodong Cui, Zvi Kons |
| 2025 | CVPR | CAV-MAE Sync: Improving Contrastive Audio-Visual Mask Autoencoders via Fine-Grained Alignment. | Edson Araujo, Andrew Rouditchenko, Yuan Gong, Saurabhchand Bhati, Samuel Thomas, Brian Kingsbury, Leonid Karlinsky, Rogrio Feris, James R. Glass, Hilde Kuehne |
| 2025 | ICASSP | LLM based Text Generation for Improved Low-resource Speech Recognition Models. | Tohru Nagano, Gakuto Kurata, Samuel Thomas, Hong-Kwang Jeff Kuo, Daniel Bolaos, Hyun Jung, George Saon |
| 2025 | ICASSP | A Non-autoregressive Model for Joint STT and TTS. | Vishal Sunder, Brian Kingsbury, George Saon, Samuel Thomas, Slava Shechtman, Hagai Aronowitz, Eric Fosler-Lussier, Luis A. Lastras |
| 2024 | ASPLOS | Automatic Generation of Vectorizing Compilers for Customizable Digital Signal Processors. | Samuel Thomas, James Bornholt |
| 2024 | ASPLOS | A Midsummer Night's Tree: Efficient and High Performance Secure SCM. | Samuel Thomas, Kidus Workneh, Jac McCarty, Joseph Izraelevitz, Tamara Lehman, R. Iris Bahar |
| 2024 | CVPR | What, When, and Where? Self-Supervised Spatio- Temporal Grounding in Untrimmed Multi-Action Videos from Narrated Instructions. | Brian Chen, Nina Shvetsova, Andrew Rouditchenko, Daniel Kondermann, Samuel Thomas, Shih-Fu Chang, Rogrio Feris, James R. Glass, Hilde Kuehne |
| 2024 | Interspeech | Whisper-Flamingo: Integrating Visual Features into Whisper for Audio-Visual Speech Recognition and Translation. | Andrew Rouditchenko, Yuan Gong, Samuel Thomas, Leonid Karlinsky, Hilde Kuehne, Rogrio Feris, James Glass |
| 2024 | WAOA | Lower Bounds for Approximate (& Exact) k-Disjoint-Shortest-Paths. | Rajesh Chitnis, Samuel Thomas, Anthony Wirth |
| 2023 | ICASSP | Effective Training of RNN Transducer Models on Diverse Sources of Speech and Text Data. | Takashi Fukuda, Samuel Thomas |
| 2023 | ICASSP | C2KD: Cross-Lingual Cross-Modal Knowledge Distillation for Multilingual Text-Video Retrieval. | Andrew Rouditchenko, Yung-Sung Chuang, Nina Shvetsova, Samuel Thomas, Rogrio Feris, Brian Kingsbury, Leonid Karlinsky, David Harwath, Hilde Kuehne, James R. Glass |
| 2023 | ICASSP | Fine-Grained Textual Knowledge Transfer to Improve RNN Transducers for Speech Recognition and Understanding. | Vishal Sunder, Samuel Thomas, Hong-Kwang Jeff Kuo, Brian Kingsbury, Eric Fosler-Lussier |
| 2023 | ICASSP | Multi-Speaker Data Augmentation for Improved end-to-end Automatic Speech Recognition. | Samuel Thomas, Hong-Kwang Jeff Kuo, George Saon, Brian Kingsbury |
| 2023 | Interspeech | Comparison of Multilingual Self-Supervised and Weakly-Supervised Speech Pre-Training for Adaptation to Unseen Languages. | Andrew Rouditchenko, Sameer Khurana, Samuel Thomas, Rogrio Feris, Leonid Karlinsky, Hilde Kuehne, David Harwath, Brian Kingsbury, James R. Glass |
| 2023 | Interspeech | ConvKT: Conversation-Level Knowledge Transfer for Context Aware End-to-End Spoken Language Understanding. | Vishal Sunder, Eric Fosler-Lussier, Samuel Thomas, Hong-Kwang Jeff Kuo, Brian Kingsbury |
| 2022 | CVPR | Everything at Once - Multi-modal Fusion Transformer for Video Retrieval. | Nina Shvetsova, Brian Chen, Andrew Rouditchenko, Samuel Thomas, Brian Kingsbury, Rogrio Feris, David Harwath, James R. Glass, Hilde Kuehne |
| 2022 | ICASSP | A New Data Augmentation Method for Intent Classification Enhancement and its Application on Spoken Conversation Datasets. | Zvi Kons, Aharon Satt, Hong-Kwang Kuo, Samuel Thomas, Boaz Carmeli, Ron Hoory, Brian Kingsbury |
| 2022 | ICASSP | Improving End-to-end Models for Set Prediction in Spoken Language Understanding. | Hong-Kwang Jeff Kuo, Zoltn Tske, Samuel Thomas, Brian Kingsbury, George Saon |
| 2022 | ICASSP | Towards End-to-End Integration of Dialog History for Improved Spoken Language Understanding. | Vishal Sunder, Samuel Thomas, Hong-Kwang Jeff Kuo, Jatin Ganhotra, Brian Kingsbury, Eric Fosler-Lussier |
| 2022 | ICASSP | Towards Reducing the Need for Speech Training Data to Build Spoken Language Understanding Systems. | Samuel Thomas, Hong-Kwang Jeff Kuo, Brian Kingsbury, George Saon |
| 2022 | ICASSP | Integrating Text Inputs for Training and Adapting RNN Transducer ASR Models. | Samuel Thomas, Brian Kingsbury, George Saon, Hong-Kwang Jeff Kuo |
| 2022 | Interspeech | Global RNN Transducer Models For Multi-dialect Speech Recognition. | Takashi Fukuda, Samuel Thomas, Masayuki Suzuki, Gakuto Kurata, George Saon, Brian Kingsbury |
| 2022 | Interspeech | Extending RNN-T-based speech recognition systems with emotion and language classification. | Zvi Kons, Hagai Aronowitz, Edmilson da Silva Morais, Matheus Damasceno, Hong-Kwang Kuo, Samuel Thomas, George Saon |
| 2022 | Interspeech | Tokenwise Contrastive Pretraining for Finer Speech-to-BERT Alignment in End-to-End Speech-to-Intent Systems. | Vishal Sunder, Eric Fosler-Lussier, Samuel Thomas, Hong-Kwang Kuo, Brian Kingsbury |
| 2021 | ASPLOS | A compiler infrastructure for accelerator generators. | Rachit Nigam, Samuel Thomas, Zhijing Li, Adrian Sampson |
| 2021 | ICASSP | RNN Transducer Models for Spoken Language Understanding. | Samuel Thomas, Hong-Kwang Jeff Kuo, George Saon, Zoltn Tske, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory |
| 2021 | ICASSP | End-to-End Spoken Language Understanding Using Transformer Networks and Self-Supervised Pre-Trained Features. | Edmilson da Silva Morais, Hong-Kwang Jeff Kuo, Samuel Thomas, Zoltn Tske, Brian Kingsbury |
| 2021 | ICCV | Multimodal Clustering Networks for Self-supervised Learning from Unlabeled Videos. | Brian Chen, Andrew Rouditchenko, Kevin Duarte, Hilde Kuehne, Samuel Thomas, Angie W. Boggust, Rameswar Panda, Brian Kingsbury, Rogrio Feris, David Harwath, James R. Glass, Michael Picheny, Shih-Fu Chang |
| 2021 | Interspeech | Speak or Chat with Me: End-to-End Spoken Language Understanding System with Flexible Inputs. | Sujeong Cha, Wangrui Hou, Hyun Jung, My Phung, Michael Picheny, Hong-Kwang Jeff Kuo, Samuel Thomas, Edmilson da Silva Morais |
| 2021 | Interspeech | Knowledge Distillation Based Training of Universal ASR Source Models for Cross-Lingual Transfer. | Takashi Fukuda, Samuel Thomas |
| 2021 | Interspeech | Integrating Dialog History into End-to-End Spoken Language Understanding Systems. | Jatin Ganhotra, Samuel Thomas, Hong-Kwang Jeff Kuo, Sachindra Joshi, George Saon, Zoltn Tske, Brian Kingsbury |
| 2021 | Interspeech | Cascaded Multilingual Audio-Visual Learning from Videos. | Andrew Rouditchenko, Angie W. Boggust, David Harwath, Samuel Thomas, Hilde Kuehne, Brian Chen, Rameswar Panda, Rogrio Feris, Brian Kingsbury, Michael Picheny, James R. Glass |
| 2021 | Interspeech | AVLnet: Learning Audio-Visual Language Representations from Instructional Videos. | Andrew Rouditchenko, Angie W. Boggust, David Harwath, Brian Chen, Dhiraj Joshi, Samuel Thomas, Kartik Audhkhasi, Hilde Kuehne, Rameswar Panda, Rogrio Schmidt Feris, Brian Kingsbury, Michael Picheny, Antonio Torralba, James R. Glass |
| 2020 | ICASSP | Leveraging Unpaired Text Data for Training End-To-End Speech-to-Intent Systems. | Yinghui Huang, Hong-Kwang Kuo, Samuel Thomas, Zvi Kons, Kartik Audhkhasi, Brian Kingsbury, Ron Hoory, Michael Picheny |
| 2020 | ICASSP | Audio-Assisted Image Inpainting for Talking Faces. | Alexandros Koumparoulis, Gerasimos Potamianos, Samuel Thomas, Edmilson da Silva Morais |
| 2020 | ICASSP | Training Spoken Language Understanding Systems with Non-Parallel Speech and Text. | Leda Sari, Samuel Thomas, Mark Hasegawa-Johnson |
| 2020 | Interspeech | Transliteration Based Data Augmentation for Training Multilingual ASR Acoustic Models in Low Resource Settings. | Samuel Thomas, Kartik Audhkhasi, Brian Kingsbury |
| 2020 | Interspeech | Implicit Transfer of Privileged Acoustic Information in a Generalized Knowledge Distillation Framework. | Takashi Fukuda, Samuel Thomas |
| 2020 | Interspeech | Resource-Adaptive Deep Learning for Visual Speech Recognition. | Alexandros Koumparoulis, Gerasimos Potamianos, Samuel Thomas, Edmilson da Silva Morais |
| 2020 | Interspeech | End-to-End Spoken Language Understanding Without Full Transcripts. | Hong-Kwang Jeff Kuo, Zoltn Tske, Samuel Thomas, Yinghui Huang, Kartik Audhkhasi, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory, Luis A. Lastras |
| 2020 | PLDI | Predictable accelerator design with time-sensitive affine types. | Rachit Nigam, Sachille Atapattu, Samuel Thomas, Zhijing Li, Theodore Bauer, Yuwei Ye, Apurva Koti, Adrian Sampson, Zhiru Zhang |
| 2020 | SBAC-PAD | Using Skip Graphs for Increased NUMA Locality. | Samuel Thomas, Roxana Hayne, Jonad Pulaj, Hammurabi Mendes |
| 2019 | ASRU | Mixed Bandwidth Acoustic Modeling Leveraging Knowledge Distillation. | Takashi Fukuda, Samuel Thomas |
| 2019 | ASRU | Semi-Supervised Training and Data Augmentation for Adaptation of Automatic Broadcast News Captioning Systems. | Yinghui Huang, Samuel Thomas, Masayuki Suzuki, Zoltn Tske, Larry Sansone, Michael Picheny |
| 2019 | ASRU | Simplified LSTMS for Speech Recognition. | George Saon, Zoltn Tske, Kartik Audhkhasi, Brian Kingsbury, Michael Picheny, Samuel Thomas |
| 2019 | CVPR | Grounding Spoken Words in Unlabeled Video. | Angie W. Boggust, Kartik Audhkhasi, Dhiraj Joshi, David Harwath, Samuel Thomas, Rogrio Schmidt Feris, Danny Gutfreund, Yang Zhang, Antonio Torralba, Michael Picheny, James R. Glass |
| 2019 | ICASSP | Pre-training of Speaker Embeddings for Low-latency Speaker Change Detection in Broadcast News. | Leda Sari, Samuel Thomas, Mark Hasegawa-Johnson, Michael Picheny |
| 2019 | ICASSP | Improvements to N-gram Language Model Using Text Generated from Neural Language Model. | Masayuki Suzuki, Nobuyasu Itoh, Tohru Nagano, Gakuto Kurata, Samuel Thomas |
| 2019 | ICASSP | English Broadcast News Speech Recognition by Humans and Machines. | Samuel Thomas, Masayuki Suzuki, Yinghui Huang, Gakuto Kurata, Zoltn Tske, George Saon, Brian Kingsbury, Michael Picheny, Tom Dibert, Alice Kaiser-Schatzlein, Bern Samko |
| 2019 | Interspeech | Learning Speaker Aware Offsets for Speaker Adaptation of Neural Networks. | Leda Sari, Samuel Thomas, Mark A. Hasegawa-Johnson |
| 2019 | Interspeech | Detection and Recovery of OOVs for Improved English Broadcast News Captioning. | Samuel Thomas, Kartik Audhkhasi, Zoltn Tske, Yinghui Huang, Michael Picheny |
| 2019 | PODC | Layering Data Structures over Skip Graphs for Increased NUMA Locality. | Samuel Thomas, Hammurabi Mendes |
| 2018 | ICASSP | Joint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition. | Xuesong Yang, Kartik Audhkhasi, Andrew Rosenberg, Samuel Thomas, Bhuvana Ramabhadran, Mark Hasegawa-Johnson |
| 2018 | Interspeech | Data Augmentation Improves Recognition of Foreign Accented Speech. | Takashi Fukuda, Raul Fernandez, Andrew Rosenberg, Samuel Thomas, Bhuvana Ramabhadran, Alexander Sorin, Gakuto Kurata |
| 2018 | Interspeech | Inference-Invariant Transformation of Batch Normalization for Domain Adaptation of Acoustic Models. | Masayuki Suzuki, Tohru Nagano, Gakuto Kurata, Samuel Thomas |
| 2018 | LREC | A Recorded Debating Dataset. | Shachar Mirkin, Michal Jacovi, Tamar Lavee, Hong-Kwang Kuo, Samuel Thomas, Leslie Sager, Lili Kotlerman, Elad Venezian, Noam Slonim |
| 2017 | ICASSP | Effective joint training of denoising feature space transforms and Neural Network based acoustic models. | Takashi Fukuda, Osamu Ichikawa, Gakuto Kurata, Ryuki Tachibana, Samuel Thomas, Bhuvana Ramabhadran |
| 2017 | Interspeech | Efficient Knowledge Distillation from an Ensemble of Teachers. | Takashi Fukuda, Masayuki Suzuki, Gakuto Kurata, Samuel Thomas, Jia Cui, Bhuvana Ramabhadran |
| 2017 | Interspeech | English Conversational Telephone Speech Recognition by Humans and Machines. | George Saon, Gakuto Kurata, Tom Sercu, Kartik Audhkhasi, Samuel Thomas, Dimitrios Dimitriadis, Xiaodong Cui, Bhuvana Ramabhadran, Michael Picheny, Lynn-Li Lim, Bergul Roomi, Phil Hall |
| 2016 | ICASSP | On the importance of event detection for ASR. | David Haws, Dimitrios Dimitriadis, George Saon, Samuel Thomas, Michael Picheny |
| 2016 | ICASSP | CNMF-based acoustic features for noise-robust ASR. | Colin Vaz, Dimitrios Dimitriadis, Samuel Thomas, Shrikanth S. Narayanan |
| 2016 | Interspeech | An Investigation on the Use of i-Vectors for Robust ASR. | Dimitrios Dimitriadis, Samuel Thomas, Sriram Ganapathy |
| 2016 | Interspeech | Domain Adaptation of CNN Based Acoustic Models Under Limited Resource Settings. | Masayuki Suzuki, Ryuki Tachibana, Samuel Thomas, Bhuvana Ramabhadran, George Saon |
| 2016 | Interspeech | Multilingual Data Selection for Low Resource Speech Recognition. | Samuel Thomas, Kartik Audhkhasi, Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran |
| 2015 | ICASSP | Improvements to the IBM speech activity detection system for the DARPA RATS program. | Samuel Thomas, George Saon, Maarten Van Segbroeck, Shrikanth S. Narayanan |
| 2015 | IJCAI | Compiling Constraint Networks into Multivalued Decomposable Decision Graphs. | Frdric Koriche, Jean-Marie Lagniez, Pierre Marquis, Samuel Thomas |
| 2015 | Interspeech | Investigating factor analysis features for deep neural networks in noisy speech recognition. | Sriram Ganapathy, Samuel Thomas, Dimitrios Dimitriadis, Steven J. Rennie |
| 2015 | Interspeech | The IBM BOLT speech transcription system. | Samuel Thomas, George Saon, Hong-Kwang Jeff Kuo, Lidia Mangu |
| 2014 | ICASSP | Analyzing convolutional neural networks for speech activity detection in mismatched acoustic conditions. | Samuel Thomas, Sriram Ganapathy, George Saon, Hagen Soltau |
| 2014 | Interspeech | Robust language identification using convolutional neural network features. | Sriram Ganapathy, Kyu Jeong Han, Samuel Thomas, Mohamed Kamal Omar, Maarten Van Segbroeck, Shrikanth S. Narayanan |
| 2013 | ICASSP | A summary of the 2012 JHU CLSP workshop on zero resource speech technologies and models of early language acquisition. | Aren Jansen, Emmanuel Dupoux, Sharon Goldwater, Mark Johnson, Sanjeev Khudanpur, Kenneth Church, Naomi Feldman, Hynek Hermansky, Florian Metze, Richard C. Rose, Mike Seltzer, Pascal Clark, Ian McGraw, Balakrishnan Varadarajan, Erin Bennett, Benjamin Brschinger, Justin T. Chiu, Ewan Dunbar, Abdellah Fourtassi, David Harwath, Chia-ying Lee, Keith D. Levin, Atta Norouzian, Vijayaditya Peddinti, Rachael Richardson, Thomas Schatz, Samuel Thomas |
| 2013 | ICASSP | Weak top-down constraints for unsupervised acoustic model training. | Aren Jansen, Samuel Thomas, Hynek Hermansky |
| 2013 | ICASSP | Developing a speaker identification system for the DARPA RATS project. | Oldrich Plchot, Spyros Matsoukas, Pavel Matejka, Najim Dehak, Jeff Z. Ma, Sandro Cumani, Ondrej Glembek, Hynek Hermansky, Sri Harish Reddy Mallidi, Nima Mesgarani, Richard M. Schwartz, Mehdi Soufifar, Zheng-Hua Tan, Samuel Thomas, Bing Zhang, Xinhui Zhou |
| 2013 | ICASSP | Deep neural network features and semi-supervised training for low resource speech recognition. | Samuel Thomas, Michael L. Seltzer, Kenneth Church, Hynek Hermansky |
| 2013 | IJCAI | Knowledge Compilation for Model Counting: Affine Decision Trees. | Frdric Koriche, Jean-Marie Lagniez, Pierre Marquis, Samuel Thomas |
| 2013 | Interspeech | The IBM speech activity detection system for the DARPA RATS program. | George Saon, Samuel Thomas, Hagen Soltau, Sriram Ganapathy, Brian Kingsbury |
| 2012 | ICASSP | The UMD-JHU 2011 speaker recognition system. | Daniel Garcia-Romero, Xinhui Zhou, Dmitry N. Zotkin, Balaji Vasan Srinivasan, Yuancheng Luo, Sriram Ganapathy, Samuel Thomas, Sridhar Krishna Nemala, Garimella S. V. S. Sivaram, Majid Mirbagheri, Sri Harish Reddy Mallidi, Thomas Janu, Padmanabhan Rajan, Nima Mesgarani, Mounya Elhilali, Hynek Hermansky, Shihab A. Shamma, Ramani Duraiswami |
| 2012 | ICASSP | Multilingual MLP features for low-resource LVCSR systems. | Samuel Thomas, Sriram Ganapathy, Hynek Hermansky |
| 2012 | Interspeech | Intrinsic Spectral Analysis for Zero and High Resource Speech Recognition. | Aren Jansen, Samuel Thomas, Hynek Hermansky |
| 2012 | Interspeech | Exploiting Discriminative Point Process Models for Spoken Term Detection. | Atta Norouzian, Aren Jansen, Richard C. Rose, Samuel Thomas |
| 2012 | Interspeech | Data-driven Posterior Features for Low Resource Speech Recognition Applications. | Samuel Thomas, Sriram Ganapathy, Aren Jansen, Hynek Hermansky |
| 2012 | Interspeech | Acoustic and Data-driven Features for Robust Speech Activity Detection. | Samuel Thomas, Sri Harish Reddy Mallidi, Thomas Janu, Hynek Hermansky, Nima Mesgarani, Xinhui Zhou, Shihab A. Shamma, Tim Ng, Bing Zhang, Long Nguyen, Spyros Matsoukas |
| 2011 | ICASSP | MLP based phoneme detectors for Automatic Speech Recognition. | Samuel Thomas, Patrick Nguyen, Geoffrey Zweig, Hynek Hermansky |
| 2011 | ICASSP | Speech recognitionwith segmental conditional random fields: A summary of the JHU CLSP 2010 Summer Workshop. | Geoffrey Zweig, Patrick Nguyen, Dirk Van Compernolle, Kris Demuynck, Les E. Atlas, Pascal Clark, Gregory Sell, Meihong Wang, Fei Sha, Hynek Hermansky, Damianos Karakos, Aren Jansen, Samuel Thomas, Sivaram G. S. V. S., Samuel R. Bowman, Justine T. Kao |
| 2011 | Interspeech | Rapid Evaluation of Speech Representations for Spoken Term Discovery. | Michael A. Carlin, Samuel Thomas, Aren Jansen, Hynek Hermansky |
| 2011 | Interspeech | Adaptive Stream Fusion in Multistream Recognition of Speech. | Nima Mesgarani, Samuel Thomas, Hynek Hermansky |
| 2011 | Interspeech | Mixture of Auto-Associative Neural Networks for Speaker Verification. | Garimella S. V. S. Sivaram, Samuel Thomas, Hynek Hermansky |
| 2010 | ICASSP | Multilingual acoustic modeling for speech recognition based on subspace Gaussian Mixture Models. | Luks Burget, Petr Schwarz, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Samuel Thomas |
| 2010 | ICASSP | Robust spectro-temporal features based on autoregressive models of Hilbert envelopes. | Sriram Ganapathy, Samuel Thomas, Hynek Hermansky |
| 2010 | ICASSP | Comparison of modulation features for phoneme recognition. | Sriram Ganapathy, Samuel Thomas, Hynek Hermansky |
| 2010 | ICASSP | A novel estimation of feature-space MLLR for full-covariance models. | Arnab Ghoshal, Daniel Povey, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Ondrej Glembek, Nagendra Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas |
| 2010 | ICASSP | Approaches to automatic lexicon learning with limited training examples. | Nagendra Goel, Samuel Thomas, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Petr Schwarz |
| 2010 | ICASSP | Subspace Gaussian Mixture Models for speech recognition. | Daniel Povey, Luks Burget, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas |
| 2010 | Interspeech | A multistream multiresolution framework for phoneme recognition. | Nima Mesgarani, Samuel Thomas, Hynek Hermansky |
| 2010 | Interspeech | Cross-lingual and multi-stream posterior features for low resource LVCSR systems. | Samuel Thomas, Sriram Ganapathy, Hynek Hermansky |
| 2010 | Interspeech | A phoneme recognition framework based on auditory spectro-temporal receptive fields. | Samuel Thomas, Kailash Patil, Sriram Ganapathy, Nima Mesgarani, Hynek Hermansky |
| 2009 | ASRU | Temporal envelope subtraction for robust speech recognition using modulation spectrum. | Sriram Ganapathy, Samuel Thomas, Hynek Hermansky |
| 2009 | ICASSP | Phoneme recognition using spectral envelope and modulation frequency features. | Samuel Thomas, Sriram Ganapathy, Hynek Hermansky |
| 2009 | Interspeech | Static and dynamic modulation spectrum for speech recognition. | Sriram Ganapathy, Samuel Thomas, Hynek Hermansky |
| 2009 | Interspeech | Tandem representations of spectral envelope and modulation frequency features for ASR. | Samuel Thomas, Sriram Ganapathy, Hynek Hermansky |
| 2008 | Interspeech | Front-end for far-field speech recognition based on frequency domain linear prediction. | Sriram Ganapathy, Samuel Thomas, Hynek Hermansky |
| 2008 | Interspeech | Hilbert envelope based spectro-temporal features for phoneme recognition in telephone speech. | Samuel Thomas, Sriram Ganapathy, Hynek Hermansky |
| 2007 | Interspeech | Language identification of person names using CF-IOF based weighing function. | Samuel Thomas, Ashish Verma |