Skip to content

Samuel Thomas

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

105

Venues

13

Active years

2007–2026

Best venue rank

A*

Where they publish

Papers

105 indexed papers, newest first.

YearVenueTitleAuthors
2026ISPASSCharacterizing and Optimizing Cache Placement for Secure Memory Metadata.Samuel Thomas, Blake Cragen, R. Iris Bahar, Tamara Silbergleit Lehman
2025ASRUOmni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM?Andrew Rouditchenko, Saurabhchand Bhati, Edson Araujo, Samuel Thomas, Hilde Kuehne, Rogrio Feris, James R. Glass
2025ASRUGranite-speech: open-source speech-aware LLMs with strong English ASR capabilities.George Saon, Avihu Dekel, Alexander Brooks, Tohru Nagano, Abraham Daniels, Aharon Satt, Ashish R. Mittal, Brian Kingsbury, David Haws, Edmilson da Silva Morais, Gakuto Kurata, Hagai Aronowitz, Ibrahim Ibrahim, Hong-Kwang Kuo, Kate Soule, Luis A. Lastras, Masayuki Suzuki, Ron Hoory, Samuel Thomas, Sashi Novitasari, Takashi Fukuda, Vishal Sunder, Xiaodong Cui, Zvi Kons
2025CVPRCAV-MAE Sync: Improving Contrastive Audio-Visual Mask Autoencoders via Fine-Grained Alignment.Edson Araujo, Andrew Rouditchenko, Yuan Gong, Saurabhchand Bhati, Samuel Thomas, Brian Kingsbury, Leonid Karlinsky, Rogrio Feris, James R. Glass, Hilde Kuehne
2025ICASSPLLM based Text Generation for Improved Low-resource Speech Recognition Models.Tohru Nagano, Gakuto Kurata, Samuel Thomas, Hong-Kwang Jeff Kuo, Daniel Bolaos, Hyun Jung, George Saon
2025ICASSPA Non-autoregressive Model for Joint STT and TTS.Vishal Sunder, Brian Kingsbury, George Saon, Samuel Thomas, Slava Shechtman, Hagai Aronowitz, Eric Fosler-Lussier, Luis A. Lastras
2024ASPLOSAutomatic Generation of Vectorizing Compilers for Customizable Digital Signal Processors.Samuel Thomas, James Bornholt
2024ASPLOSA Midsummer Night's Tree: Efficient and High Performance Secure SCM.Samuel Thomas, Kidus Workneh, Jac McCarty, Joseph Izraelevitz, Tamara Lehman, R. Iris Bahar
2024CVPRWhat, When, and Where? Self-Supervised Spatio- Temporal Grounding in Untrimmed Multi-Action Videos from Narrated Instructions.Brian Chen, Nina Shvetsova, Andrew Rouditchenko, Daniel Kondermann, Samuel Thomas, Shih-Fu Chang, Rogrio Feris, James R. Glass, Hilde Kuehne
2024InterspeechWhisper-Flamingo: Integrating Visual Features into Whisper for Audio-Visual Speech Recognition and Translation.Andrew Rouditchenko, Yuan Gong, Samuel Thomas, Leonid Karlinsky, Hilde Kuehne, Rogrio Feris, James Glass
2024WAOALower Bounds for Approximate (& Exact) k-Disjoint-Shortest-Paths.Rajesh Chitnis, Samuel Thomas, Anthony Wirth
2023ICASSPEffective Training of RNN Transducer Models on Diverse Sources of Speech and Text Data.Takashi Fukuda, Samuel Thomas
2023ICASSPC2KD: Cross-Lingual Cross-Modal Knowledge Distillation for Multilingual Text-Video Retrieval.Andrew Rouditchenko, Yung-Sung Chuang, Nina Shvetsova, Samuel Thomas, Rogrio Feris, Brian Kingsbury, Leonid Karlinsky, David Harwath, Hilde Kuehne, James R. Glass
2023ICASSPFine-Grained Textual Knowledge Transfer to Improve RNN Transducers for Speech Recognition and Understanding.Vishal Sunder, Samuel Thomas, Hong-Kwang Jeff Kuo, Brian Kingsbury, Eric Fosler-Lussier
2023ICASSPMulti-Speaker Data Augmentation for Improved end-to-end Automatic Speech Recognition.Samuel Thomas, Hong-Kwang Jeff Kuo, George Saon, Brian Kingsbury
2023InterspeechComparison of Multilingual Self-Supervised and Weakly-Supervised Speech Pre-Training for Adaptation to Unseen Languages.Andrew Rouditchenko, Sameer Khurana, Samuel Thomas, Rogrio Feris, Leonid Karlinsky, Hilde Kuehne, David Harwath, Brian Kingsbury, James R. Glass
2023InterspeechConvKT: Conversation-Level Knowledge Transfer for Context Aware End-to-End Spoken Language Understanding.Vishal Sunder, Eric Fosler-Lussier, Samuel Thomas, Hong-Kwang Jeff Kuo, Brian Kingsbury
2022CVPREverything at Once - Multi-modal Fusion Transformer for Video Retrieval.Nina Shvetsova, Brian Chen, Andrew Rouditchenko, Samuel Thomas, Brian Kingsbury, Rogrio Feris, David Harwath, James R. Glass, Hilde Kuehne
2022ICASSPA New Data Augmentation Method for Intent Classification Enhancement and its Application on Spoken Conversation Datasets.Zvi Kons, Aharon Satt, Hong-Kwang Kuo, Samuel Thomas, Boaz Carmeli, Ron Hoory, Brian Kingsbury
2022ICASSPImproving End-to-end Models for Set Prediction in Spoken Language Understanding.Hong-Kwang Jeff Kuo, Zoltn Tske, Samuel Thomas, Brian Kingsbury, George Saon
2022ICASSPTowards End-to-End Integration of Dialog History for Improved Spoken Language Understanding.Vishal Sunder, Samuel Thomas, Hong-Kwang Jeff Kuo, Jatin Ganhotra, Brian Kingsbury, Eric Fosler-Lussier
2022ICASSPTowards Reducing the Need for Speech Training Data to Build Spoken Language Understanding Systems.Samuel Thomas, Hong-Kwang Jeff Kuo, Brian Kingsbury, George Saon
2022ICASSPIntegrating Text Inputs for Training and Adapting RNN Transducer ASR Models.Samuel Thomas, Brian Kingsbury, George Saon, Hong-Kwang Jeff Kuo
2022InterspeechGlobal RNN Transducer Models For Multi-dialect Speech Recognition.Takashi Fukuda, Samuel Thomas, Masayuki Suzuki, Gakuto Kurata, George Saon, Brian Kingsbury
2022InterspeechExtending RNN-T-based speech recognition systems with emotion and language classification.Zvi Kons, Hagai Aronowitz, Edmilson da Silva Morais, Matheus Damasceno, Hong-Kwang Kuo, Samuel Thomas, George Saon
2022InterspeechTokenwise Contrastive Pretraining for Finer Speech-to-BERT Alignment in End-to-End Speech-to-Intent Systems.Vishal Sunder, Eric Fosler-Lussier, Samuel Thomas, Hong-Kwang Kuo, Brian Kingsbury
2021ASPLOSA compiler infrastructure for accelerator generators.Rachit Nigam, Samuel Thomas, Zhijing Li, Adrian Sampson
2021ICASSPRNN Transducer Models for Spoken Language Understanding.Samuel Thomas, Hong-Kwang Jeff Kuo, George Saon, Zoltn Tske, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory
2021ICASSPEnd-to-End Spoken Language Understanding Using Transformer Networks and Self-Supervised Pre-Trained Features.Edmilson da Silva Morais, Hong-Kwang Jeff Kuo, Samuel Thomas, Zoltn Tske, Brian Kingsbury
2021ICCVMultimodal Clustering Networks for Self-supervised Learning from Unlabeled Videos.Brian Chen, Andrew Rouditchenko, Kevin Duarte, Hilde Kuehne, Samuel Thomas, Angie W. Boggust, Rameswar Panda, Brian Kingsbury, Rogrio Feris, David Harwath, James R. Glass, Michael Picheny, Shih-Fu Chang
2021InterspeechSpeak or Chat with Me: End-to-End Spoken Language Understanding System with Flexible Inputs.Sujeong Cha, Wangrui Hou, Hyun Jung, My Phung, Michael Picheny, Hong-Kwang Jeff Kuo, Samuel Thomas, Edmilson da Silva Morais
2021InterspeechKnowledge Distillation Based Training of Universal ASR Source Models for Cross-Lingual Transfer.Takashi Fukuda, Samuel Thomas
2021InterspeechIntegrating Dialog History into End-to-End Spoken Language Understanding Systems.Jatin Ganhotra, Samuel Thomas, Hong-Kwang Jeff Kuo, Sachindra Joshi, George Saon, Zoltn Tske, Brian Kingsbury
2021InterspeechCascaded Multilingual Audio-Visual Learning from Videos.Andrew Rouditchenko, Angie W. Boggust, David Harwath, Samuel Thomas, Hilde Kuehne, Brian Chen, Rameswar Panda, Rogrio Feris, Brian Kingsbury, Michael Picheny, James R. Glass
2021InterspeechAVLnet: Learning Audio-Visual Language Representations from Instructional Videos.Andrew Rouditchenko, Angie W. Boggust, David Harwath, Brian Chen, Dhiraj Joshi, Samuel Thomas, Kartik Audhkhasi, Hilde Kuehne, Rameswar Panda, Rogrio Schmidt Feris, Brian Kingsbury, Michael Picheny, Antonio Torralba, James R. Glass
2020ICASSPLeveraging Unpaired Text Data for Training End-To-End Speech-to-Intent Systems.Yinghui Huang, Hong-Kwang Kuo, Samuel Thomas, Zvi Kons, Kartik Audhkhasi, Brian Kingsbury, Ron Hoory, Michael Picheny
2020ICASSPAudio-Assisted Image Inpainting for Talking Faces.Alexandros Koumparoulis, Gerasimos Potamianos, Samuel Thomas, Edmilson da Silva Morais
2020ICASSPTraining Spoken Language Understanding Systems with Non-Parallel Speech and Text.Leda Sari, Samuel Thomas, Mark Hasegawa-Johnson
2020InterspeechTransliteration Based Data Augmentation for Training Multilingual ASR Acoustic Models in Low Resource Settings.Samuel Thomas, Kartik Audhkhasi, Brian Kingsbury
2020InterspeechImplicit Transfer of Privileged Acoustic Information in a Generalized Knowledge Distillation Framework.Takashi Fukuda, Samuel Thomas
2020InterspeechResource-Adaptive Deep Learning for Visual Speech Recognition.Alexandros Koumparoulis, Gerasimos Potamianos, Samuel Thomas, Edmilson da Silva Morais
2020InterspeechEnd-to-End Spoken Language Understanding Without Full Transcripts.Hong-Kwang Jeff Kuo, Zoltn Tske, Samuel Thomas, Yinghui Huang, Kartik Audhkhasi, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory, Luis A. Lastras
2020PLDIPredictable accelerator design with time-sensitive affine types.Rachit Nigam, Sachille Atapattu, Samuel Thomas, Zhijing Li, Theodore Bauer, Yuwei Ye, Apurva Koti, Adrian Sampson, Zhiru Zhang
2020SBAC-PADUsing Skip Graphs for Increased NUMA Locality.Samuel Thomas, Roxana Hayne, Jonad Pulaj, Hammurabi Mendes
2019ASRUMixed Bandwidth Acoustic Modeling Leveraging Knowledge Distillation.Takashi Fukuda, Samuel Thomas
2019ASRUSemi-Supervised Training and Data Augmentation for Adaptation of Automatic Broadcast News Captioning Systems.Yinghui Huang, Samuel Thomas, Masayuki Suzuki, Zoltn Tske, Larry Sansone, Michael Picheny
2019ASRUSimplified LSTMS for Speech Recognition.George Saon, Zoltn Tske, Kartik Audhkhasi, Brian Kingsbury, Michael Picheny, Samuel Thomas
2019CVPRGrounding Spoken Words in Unlabeled Video.Angie W. Boggust, Kartik Audhkhasi, Dhiraj Joshi, David Harwath, Samuel Thomas, Rogrio Schmidt Feris, Danny Gutfreund, Yang Zhang, Antonio Torralba, Michael Picheny, James R. Glass
2019ICASSPPre-training of Speaker Embeddings for Low-latency Speaker Change Detection in Broadcast News.Leda Sari, Samuel Thomas, Mark Hasegawa-Johnson, Michael Picheny
2019ICASSPImprovements to N-gram Language Model Using Text Generated from Neural Language Model.Masayuki Suzuki, Nobuyasu Itoh, Tohru Nagano, Gakuto Kurata, Samuel Thomas
2019ICASSPEnglish Broadcast News Speech Recognition by Humans and Machines.Samuel Thomas, Masayuki Suzuki, Yinghui Huang, Gakuto Kurata, Zoltn Tske, George Saon, Brian Kingsbury, Michael Picheny, Tom Dibert, Alice Kaiser-Schatzlein, Bern Samko
2019InterspeechLearning Speaker Aware Offsets for Speaker Adaptation of Neural Networks.Leda Sari, Samuel Thomas, Mark A. Hasegawa-Johnson
2019InterspeechDetection and Recovery of OOVs for Improved English Broadcast News Captioning.Samuel Thomas, Kartik Audhkhasi, Zoltn Tske, Yinghui Huang, Michael Picheny
2019PODCLayering Data Structures over Skip Graphs for Increased NUMA Locality.Samuel Thomas, Hammurabi Mendes
2018ICASSPJoint Modeling of Accents and Acoustics for Multi-Accent Speech Recognition.Xuesong Yang, Kartik Audhkhasi, Andrew Rosenberg, Samuel Thomas, Bhuvana Ramabhadran, Mark Hasegawa-Johnson
2018InterspeechData Augmentation Improves Recognition of Foreign Accented Speech.Takashi Fukuda, Raul Fernandez, Andrew Rosenberg, Samuel Thomas, Bhuvana Ramabhadran, Alexander Sorin, Gakuto Kurata
2018InterspeechInference-Invariant Transformation of Batch Normalization for Domain Adaptation of Acoustic Models.Masayuki Suzuki, Tohru Nagano, Gakuto Kurata, Samuel Thomas
2018LRECA Recorded Debating Dataset.Shachar Mirkin, Michal Jacovi, Tamar Lavee, Hong-Kwang Kuo, Samuel Thomas, Leslie Sager, Lili Kotlerman, Elad Venezian, Noam Slonim
2017ICASSPEffective joint training of denoising feature space transforms and Neural Network based acoustic models.Takashi Fukuda, Osamu Ichikawa, Gakuto Kurata, Ryuki Tachibana, Samuel Thomas, Bhuvana Ramabhadran
2017InterspeechEfficient Knowledge Distillation from an Ensemble of Teachers.Takashi Fukuda, Masayuki Suzuki, Gakuto Kurata, Samuel Thomas, Jia Cui, Bhuvana Ramabhadran
2017InterspeechEnglish Conversational Telephone Speech Recognition by Humans and Machines.George Saon, Gakuto Kurata, Tom Sercu, Kartik Audhkhasi, Samuel Thomas, Dimitrios Dimitriadis, Xiaodong Cui, Bhuvana Ramabhadran, Michael Picheny, Lynn-Li Lim, Bergul Roomi, Phil Hall
2016ICASSPOn the importance of event detection for ASR.David Haws, Dimitrios Dimitriadis, George Saon, Samuel Thomas, Michael Picheny
2016ICASSPCNMF-based acoustic features for noise-robust ASR.Colin Vaz, Dimitrios Dimitriadis, Samuel Thomas, Shrikanth S. Narayanan
2016InterspeechAn Investigation on the Use of i-Vectors for Robust ASR.Dimitrios Dimitriadis, Samuel Thomas, Sriram Ganapathy
2016InterspeechDomain Adaptation of CNN Based Acoustic Models Under Limited Resource Settings.Masayuki Suzuki, Ryuki Tachibana, Samuel Thomas, Bhuvana Ramabhadran, George Saon
2016InterspeechMultilingual Data Selection for Low Resource Speech Recognition.Samuel Thomas, Kartik Audhkhasi, Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran
2015ICASSPImprovements to the IBM speech activity detection system for the DARPA RATS program.Samuel Thomas, George Saon, Maarten Van Segbroeck, Shrikanth S. Narayanan
2015IJCAICompiling Constraint Networks into Multivalued Decomposable Decision Graphs.Frdric Koriche, Jean-Marie Lagniez, Pierre Marquis, Samuel Thomas
2015InterspeechInvestigating factor analysis features for deep neural networks in noisy speech recognition.Sriram Ganapathy, Samuel Thomas, Dimitrios Dimitriadis, Steven J. Rennie
2015InterspeechThe IBM BOLT speech transcription system.Samuel Thomas, George Saon, Hong-Kwang Jeff Kuo, Lidia Mangu
2014ICASSPAnalyzing convolutional neural networks for speech activity detection in mismatched acoustic conditions.Samuel Thomas, Sriram Ganapathy, George Saon, Hagen Soltau
2014InterspeechRobust language identification using convolutional neural network features.Sriram Ganapathy, Kyu Jeong Han, Samuel Thomas, Mohamed Kamal Omar, Maarten Van Segbroeck, Shrikanth S. Narayanan
2013ICASSPA summary of the 2012 JHU CLSP workshop on zero resource speech technologies and models of early language acquisition.Aren Jansen, Emmanuel Dupoux, Sharon Goldwater, Mark Johnson, Sanjeev Khudanpur, Kenneth Church, Naomi Feldman, Hynek Hermansky, Florian Metze, Richard C. Rose, Mike Seltzer, Pascal Clark, Ian McGraw, Balakrishnan Varadarajan, Erin Bennett, Benjamin Brschinger, Justin T. Chiu, Ewan Dunbar, Abdellah Fourtassi, David Harwath, Chia-ying Lee, Keith D. Levin, Atta Norouzian, Vijayaditya Peddinti, Rachael Richardson, Thomas Schatz, Samuel Thomas
2013ICASSPWeak top-down constraints for unsupervised acoustic model training.Aren Jansen, Samuel Thomas, Hynek Hermansky
2013ICASSPDeveloping a speaker identification system for the DARPA RATS project.Oldrich Plchot, Spyros Matsoukas, Pavel Matejka, Najim Dehak, Jeff Z. Ma, Sandro Cumani, Ondrej Glembek, Hynek Hermansky, Sri Harish Reddy Mallidi, Nima Mesgarani, Richard M. Schwartz, Mehdi Soufifar, Zheng-Hua Tan, Samuel Thomas, Bing Zhang, Xinhui Zhou
2013ICASSPDeep neural network features and semi-supervised training for low resource speech recognition.Samuel Thomas, Michael L. Seltzer, Kenneth Church, Hynek Hermansky
2013IJCAIKnowledge Compilation for Model Counting: Affine Decision Trees.Frdric Koriche, Jean-Marie Lagniez, Pierre Marquis, Samuel Thomas
2013InterspeechThe IBM speech activity detection system for the DARPA RATS program.George Saon, Samuel Thomas, Hagen Soltau, Sriram Ganapathy, Brian Kingsbury
2012ICASSPThe UMD-JHU 2011 speaker recognition system.Daniel Garcia-Romero, Xinhui Zhou, Dmitry N. Zotkin, Balaji Vasan Srinivasan, Yuancheng Luo, Sriram Ganapathy, Samuel Thomas, Sridhar Krishna Nemala, Garimella S. V. S. Sivaram, Majid Mirbagheri, Sri Harish Reddy Mallidi, Thomas Janu, Padmanabhan Rajan, Nima Mesgarani, Mounya Elhilali, Hynek Hermansky, Shihab A. Shamma, Ramani Duraiswami
2012ICASSPMultilingual MLP features for low-resource LVCSR systems.Samuel Thomas, Sriram Ganapathy, Hynek Hermansky
2012InterspeechIntrinsic Spectral Analysis for Zero and High Resource Speech Recognition.Aren Jansen, Samuel Thomas, Hynek Hermansky
2012InterspeechExploiting Discriminative Point Process Models for Spoken Term Detection.Atta Norouzian, Aren Jansen, Richard C. Rose, Samuel Thomas
2012InterspeechData-driven Posterior Features for Low Resource Speech Recognition Applications.Samuel Thomas, Sriram Ganapathy, Aren Jansen, Hynek Hermansky
2012InterspeechAcoustic and Data-driven Features for Robust Speech Activity Detection.Samuel Thomas, Sri Harish Reddy Mallidi, Thomas Janu, Hynek Hermansky, Nima Mesgarani, Xinhui Zhou, Shihab A. Shamma, Tim Ng, Bing Zhang, Long Nguyen, Spyros Matsoukas
2011ICASSPMLP based phoneme detectors for Automatic Speech Recognition.Samuel Thomas, Patrick Nguyen, Geoffrey Zweig, Hynek Hermansky
2011ICASSPSpeech recognitionwith segmental conditional random fields: A summary of the JHU CLSP 2010 Summer Workshop.Geoffrey Zweig, Patrick Nguyen, Dirk Van Compernolle, Kris Demuynck, Les E. Atlas, Pascal Clark, Gregory Sell, Meihong Wang, Fei Sha, Hynek Hermansky, Damianos Karakos, Aren Jansen, Samuel Thomas, Sivaram G. S. V. S., Samuel R. Bowman, Justine T. Kao
2011InterspeechRapid Evaluation of Speech Representations for Spoken Term Discovery.Michael A. Carlin, Samuel Thomas, Aren Jansen, Hynek Hermansky
2011InterspeechAdaptive Stream Fusion in Multistream Recognition of Speech.Nima Mesgarani, Samuel Thomas, Hynek Hermansky
2011InterspeechMixture of Auto-Associative Neural Networks for Speaker Verification.Garimella S. V. S. Sivaram, Samuel Thomas, Hynek Hermansky
2010ICASSPMultilingual acoustic modeling for speech recognition based on subspace Gaussian Mixture Models.Luks Burget, Petr Schwarz, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Samuel Thomas
2010ICASSPRobust spectro-temporal features based on autoregressive models of Hilbert envelopes.Sriram Ganapathy, Samuel Thomas, Hynek Hermansky
2010ICASSPComparison of modulation features for phoneme recognition.Sriram Ganapathy, Samuel Thomas, Hynek Hermansky
2010ICASSPA novel estimation of feature-space MLLR for full-covariance models.Arnab Ghoshal, Daniel Povey, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Ondrej Glembek, Nagendra Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas
2010ICASSPApproaches to automatic lexicon learning with limited training examples.Nagendra Goel, Samuel Thomas, Mohit Agarwal, Pinar Akyazi, Luks Burget, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Martin Karafit, Daniel Povey, Ariya Rastrow, Richard C. Rose, Petr Schwarz
2010ICASSPSubspace Gaussian Mixture Models for speech recognition.Daniel Povey, Luks Burget, Mohit Agarwal, Pinar Akyazi, Kai Feng, Arnab Ghoshal, Ondrej Glembek, Nagendra K. Goel, Martin Karafit, Ariya Rastrow, Richard C. Rose, Petr Schwarz, Samuel Thomas
2010InterspeechA multistream multiresolution framework for phoneme recognition.Nima Mesgarani, Samuel Thomas, Hynek Hermansky
2010InterspeechCross-lingual and multi-stream posterior features for low resource LVCSR systems.Samuel Thomas, Sriram Ganapathy, Hynek Hermansky
2010InterspeechA phoneme recognition framework based on auditory spectro-temporal receptive fields.Samuel Thomas, Kailash Patil, Sriram Ganapathy, Nima Mesgarani, Hynek Hermansky
2009ASRUTemporal envelope subtraction for robust speech recognition using modulation spectrum.Sriram Ganapathy, Samuel Thomas, Hynek Hermansky
2009ICASSPPhoneme recognition using spectral envelope and modulation frequency features.Samuel Thomas, Sriram Ganapathy, Hynek Hermansky
2009InterspeechStatic and dynamic modulation spectrum for speech recognition.Sriram Ganapathy, Samuel Thomas, Hynek Hermansky
2009InterspeechTandem representations of spectral envelope and modulation frequency features for ASR.Samuel Thomas, Sriram Ganapathy, Hynek Hermansky
2008InterspeechFront-end for far-field speech recognition based on frequency domain linear prediction.Sriram Ganapathy, Samuel Thomas, Hynek Hermansky
2008InterspeechHilbert envelope based spectro-temporal features for phoneme recognition in telephone speech.Samuel Thomas, Sriram Ganapathy, Hynek Hermansky
2007InterspeechLanguage identification of person names using CF-IOF based weighing function.Samuel Thomas, Ashish Verma