| 2024 | ICASSP | SADA: Saudi Audio Dataset for Arabic. | Sadeen Alharbi, Areeb Alowisheq, Zoltn Tske, Kareem Darwish, Abdullah Alrajeh, Abdulmajeed Alrowithi, Aljawharah Bin Tamran, Asma Ibrahim, Raghad Aloraini, Raneem Alnajim, Ranya A. Alkahtani, Renad Almuasaad, Sara Alrasheed, Shaykhah Alsubaie, Yaser Alonaizan |
| 2022 | ICASSP | Improving End-to-end Models for Set Prediction in Spoken Language Understanding. | Hong-Kwang Jeff Kuo, Zoltn Tske, Samuel Thomas, Brian Kingsbury, George Saon |
| 2021 | ICASSP | RNN Transducer Models for Spoken Language Understanding. | Samuel Thomas, Hong-Kwang Jeff Kuo, George Saon, Zoltn Tske, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory |
| 2021 | ICASSP | End-to-End Spoken Language Understanding Using Transformer Networks and Self-Supervised Pre-Trained Features. | Edmilson da Silva Morais, Hong-Kwang Jeff Kuo, Samuel Thomas, Zoltn Tske, Brian Kingsbury |
| 2021 | ICASSP | Advancing RNN Transducer Technology for Speech Recognition. | George Saon, Zoltn Tske, Daniel Bolaos, Brian Kingsbury |
| 2021 | Interspeech | Reducing Exposure Bias in Training Recurrent Neural Network Transducers. | Xiaodong Cui, Brian Kingsbury, George Saon, David Haws, Zoltn Tske |
| 2021 | Interspeech | 4-Bit Quantization of LSTM-Based Speech Recognition Models. | Andrea Fasoli, Chia-Yu Chen, Mauricio J. Serrano, Xiao Sun, Naigang Wang, Swagath Venkataramani, George Saon, Xiaodong Cui, Brian Kingsbury, Wei Zhang, Zoltn Tske, Kailash Gopalakrishnan |
| 2021 | Interspeech | Integrating Dialog History into End-to-End Spoken Language Understanding Systems. | Jatin Ganhotra, Samuel Thomas, Hong-Kwang Jeff Kuo, Sachindra Joshi, George Saon, Zoltn Tske, Brian Kingsbury |
| 2021 | Interspeech | Improving Customization of Neural Transducers by Mitigating Acoustic Mismatch of Synthesized Audio. | Gakuto Kurata, George Saon, Brian Kingsbury, David Haws, Zoltn Tske |
| 2021 | Interspeech | On the Limit of English Conversational Speech Recognition. | Zoltn Tske, George Saon, Brian Kingsbury |
| 2020 | ICASSP | Alignment-Length Synchronous Decoding for RNN Transducer. | George Saon, Zoltn Tske, Kartik Audhkhasi |
| 2020 | Interspeech | End-to-End Spoken Language Understanding Without Full Transcripts. | Hong-Kwang Jeff Kuo, Zoltn Tske, Samuel Thomas, Yinghui Huang, Kartik Audhkhasi, Brian Kingsbury, Gakuto Kurata, Zvi Kons, Ron Hoory, Luis A. Lastras |
| 2020 | Interspeech | Single Headed Attention Based Sequence-to-Sequence Model for State-of-the-Art Results on Switchboard. | Zoltn Tske, George Saon, Kartik Audhkhasi, Brian Kingsbury |
| 2019 | ASRU | Semi-Supervised Training and Data Augmentation for Adaptation of Automatic Broadcast News Captioning Systems. | Yinghui Huang, Samuel Thomas, Masayuki Suzuki, Zoltn Tske, Larry Sansone, Michael Picheny |
| 2019 | ASRU | Simplified LSTMS for Speech Recognition. | George Saon, Zoltn Tske, Kartik Audhkhasi, Brian Kingsbury, Michael Picheny, Samuel Thomas |
| 2019 | ICASSP | Sequence Noise Injected Training for End-to-end Speech Recognition. | George Saon, Zoltn Tske, Kartik Audhkhasi, Brian Kingsbury |
| 2019 | ICASSP | English Broadcast News Speech Recognition by Humans and Machines. | Samuel Thomas, Masayuki Suzuki, Yinghui Huang, Gakuto Kurata, Zoltn Tske, George Saon, Brian Kingsbury, Michael Picheny, Tom Dibert, Alice Kaiser-Schatzlein, Bern Samko |
| 2019 | Interspeech | Forget a Bit to Learn Better: Soft Forgetting for CTC-Based Automatic Speech Recognition. | Kartik Audhkhasi, George Saon, Zoltn Tske, Brian Kingsbury, Michael Picheny |
| 2019 | Interspeech | Challenging the Boundaries of Speech Recognition: The MALACH Corpus. | Michael Picheny, Zoltn Tske, Brian Kingsbury, Kartik Audhkhasi, Xiaodong Cui, George Saon |
| 2019 | Interspeech | Detection and Recovery of OOVs for Improved English Broadcast News Captioning. | Samuel Thomas, Kartik Audhkhasi, Zoltn Tske, Yinghui Huang, Michael Picheny |
| 2019 | Interspeech | Advancing Sequence-to-Sequence Based Speech Recognition. | Zoltn Tske, Kartik Audhkhasi, George Saon |
| 2018 | ICASSP | Acoustic Modeling of Speech Waveform Based on Multi-Resolution, Neural Network Signal Processing. | Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2018 | Interspeech | Investigation on LSTM Recurrent N-gram Language Models for Speech Recognition. | Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2017 | Interspeech | Parallel Neural Network Features for Improved Tandem Acoustic Modeling. | Zoltn Tske, Wilfried Michel, Ralf Schlter, Hermann Ney |
| 2016 | ICASSP | Investigation on log-linear interpolation of multi-domain neural network language model. | Zoltn Tske, Kazuki Irie, Ralf Schlter, Hermann Ney |
| 2016 | Interspeech | LSTM, GRU, Highway and a Bit of Attention: An Empirical Overview for Language Modeling in Speech Recognition. | Kazuki Irie, Zoltn Tske, Tamer Alkhouli, Ralf Schlter, Hermann Ney |
| 2015 | ASRU | Multilingual representations for low resource speech recognition and keyword search. | Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran, Abhinav Sethy, Kartik Audhkhasi, Xiaodong Cui, Ellen Kislal, Lidia Mangu, Markus Nubaum-Thom, Michael Picheny, Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney, Mark J. F. Gales, Kate M. Knill, Anton Ragni, Haipeng Wang, Philip C. Woodland |
| 2015 | ASRU | Speaker adaptive joint training of Gaussian mixture models and bottleneck features. | Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney |
| 2015 | ICASSP | Integrating Gaussian mixtures into deep neural networks: Softmax layer with hidden variables. | Zoltn Tske, Muhammad Ali Tahir, Ralf Schlter, Hermann Ney |
| 2015 | Interspeech | Convolutional neural networks for acoustic modeling of raw time signal in LVCSR. | Pavel Golik, Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2015 | Interspeech | Multilingual features based keyword search for very low-resource languages. | Pavel Golik, Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2015 | Interspeech | Improvements in RWTH LVCSR evaluation systems for Polish, Portuguese, English, urdu, and Arabic. | M. Ali Basha Shaik, Zoltn Tske, Muhammad Ali Tahir, Markus Nubaum-Thom, Ralf Schlter, Hermann Ney |
| 2014 | ICASSP | Multilingual MRASTA features for low-resource keyword search and speech recognition systems. | Zoltn Tske, David Nolden, Ralf Schlter, Hermann Ney |
| 2014 | ICASSP | The RWTH English lecture recognition system. | Simon Wiesler, Kazuki Irie, Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2014 | Interspeech | RWTH LVCSR systems for quaero and EU-bridge: German, Polish, Spanish and Portuguese. | M. Ali Basha Shaik, Zoltn Tske, Muhammad Ali Tahir, Markus Nubaum-Thom, Ralf Schlter, Hermann Ney |
| 2014 | Interspeech | Lattice decoding and rescoring with long-Span neural network language models. | Martin Sundermeyer, Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2014 | Interspeech | Data augmentation, feature combination, and multilingual neural networks to improve ASR and KWS performance for low-resource languages. | Zoltn Tske, Pavel Golik, David Nolden, Ralf Schlter, Hermann Ney |
| 2014 | Interspeech | Acoustic modeling with deep neural networks using raw time signal for LVCSR. | Zoltn Tske, Pavel Golik, Ralf Schlter, Hermann Ney |
| 2013 | ICASSP | Investigation on cross- and multilingual MLP features under matched and mismatched acoustical conditions. | Zoltn Tske, Joel Pinto, Daniel Willett, Ralf Schlter |
| 2013 | ICASSP | Deep hierarchical bottleneck MRASTA features for LVCSR. | Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2013 | Interspeech | Development of the RWTH transcription system for slovenian. | Pavel Golik, Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2013 | Interspeech | Multilingual hierarchical MRASTA features for ASR. | Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2012 | ICASSP | Comparison and combination of different CRBE based MLP features for LVCSR. | Zoltn Tske, Ralf Schlter, Hermann Ney |
| 2012 | Interspeech | Posterior-Scaled MPE: Novel Discriminative Training Criteria. | Markus Nubaum-Thom, Zoltn Tske, Georg Heigold, Ralf Schlter, Hermann Ney |
| 2012 | Interspeech | Non-stationary signal processing and its application in speech recognition. | Zoltn Tske, Friedhelm R. Drepper, Ralf Schlter |
| 2012 | Interspeech | Context-Dependent MLPs for LVCSR: TANDEM, Hybrid or Both? | Zoltn Tske, Ralf Schlter, Hermann Ney, Martin Sundermeyer |
| 2011 | ICASSP | Non-stationary feature extraction for automatic speech recognition. | Zoltn Tske, Pavel Golik, Ralf Schlter, Friedhelm R. Drepper |
| 2011 | Interspeech | A Study on Speaker Normalized MLP Features in LVCSR. | Zoltn Tske, Christian Plahl, Ralf Schlter |
| 2009 | Interspeech | Investigation of morph-based speech recognition improvements across speech genres. | Pter Mihajlik, Balzs Tarjn, Zoltn Tske, Tibor Fegy |
| 2007 | Interspeech | A morpho-graphemic approach for the recognition of spontaneous speech in agglutinative languages - like Hungarian. | Pter Mihajlik, Tibor Fegy, Zoltn Tske, Pavel Ircing |
| 2005 | Interspeech | Evaluation and optimization of noise robust front-end technologies for the automatic recognition of Hungarian telephone speech. | Pter Mihajlik, Zoltn Tobler, Zoltn Tske, Gza Gordos |
| 2005 | Interspeech | Robust voice activity detection based on the entropy of noise-suppressed spectrum. | Zoltn Tske, Pter Mihajlik, Zoltn Tobler, Tibor Fegy |