| 2025 | ICASSP | Semi-intrusive audio evaluation: Casting non-intrusive assessment as a multi-modal text prediction task. | Jozef Coldenhoff, Milos Cernak |
| 2025 | ICASSP | OpenACE: An Open Benchmark for Evaluating Audio Coding Performance. | Jozef Coldenhoff, Niclas Granqvist, Milos Cernak |
| 2025 | Interspeech | Model as Loss: A Self-Consistent Training Paradigm. | Saisamarth Rajesh Phaye, Milos Cernak, Andrew Harper |
| 2025 | Interspeech | DeepFilterGAN: A Full-band Real-time Speech Enhancement System with GAN-based Stochastic Regeneration. | Sanberk Serbest, Tijana Stojkovic, Milos Cernak, Andrew Harper |
| 2024 | ICASSP | Multi-Channel Mosra: Mean Opinion Score and Room Acoustics Estimation Using Simulated Data and A Teacher Model. | Jozef Coldenhoff, Andrew Harper, Paul Kendrick, Tijana Stojkovic, Milos Cernak |
| 2024 | ICASSP | On Real-Time Multi-Stage Speech Enhancement Systems. | Lingjun Meng, Jozef Coldenhoff, Paul Kendrick, Tijana Stojkovic, Andrew Harper, Kiril Ratmanski, Milos Cernak |
| 2023 | ICASSP | Efficient Speech Quality Assessment Using Self-Supervised Framewise Embeddings. | Karl El Hajal, Zihan Wu, Neil Scheidwasser-Clow, Gasser Elbanna, Milos Cernak |
| 2023 | ICASSP | Personalized Task Load Prediction in Speech Communication. | Robert P. Spang, Karl El Hajal, Sebastian Mller, Milos Cernak |
| 2023 | Interspeech | Speaker Embeddings as Individuality Proxy for Voice Stress Detection. | Zihan Wu, Neil Scheidwasser-Clow, Karl El Hajal, Milos Cernak |
| 2023 | Interspeech | ALO-VC: Any-to-any Low-latency One-shot Voice Conversion. | Bohan Wang, Damien Ronssin, Milos Cernak |
| 2022 | ICASSP | SERAB: A Multi-Lingual Benchmark for Speech Emotion Recognition. | Neil Scheidwasser-Clow, Mikolaj Kegler, Pierre Beckmann, Milos Cernak |
| 2022 | Interspeech | PEAF: Learnable Power Efficient Analog Acoustic Features for Audio Recognition. | Boris Bergsma, Minhao Yang, Milos Cernak |
| 2022 | Interspeech | Hybrid Handcrafted and Learnable Audio Representation for Analysis of Speech Under Cognitive and Physical Load. | Gasser Elbanna, Alice Biryukov, Neil Scheidwasser-Clow, Lara Orlandic, Pablo Mainar, Mikolaj Kegler, Pierre Beckmann, Milos Cernak |
| 2022 | Interspeech | MOSRA: Joint Mean Opinion Score and Room Acoustics Speech Quality Assessment. | Karl El Hajal, Milos Cernak, Pablo Mainar |
| 2022 | Interspeech | Application for Real-time Personalized Speaker Extraction. | Damien Ronssin, Milos Cernak |
| 2021 | ASRU | AC-VC: Non-Parallel Low Latency Phonetic Posteriorgrams Based Voice Conversion. | Damien Ronssin, Milos Cernak |
| 2021 | Interspeech | Non-Intrusive Speech Quality Assessment with Transfer Learning and Subject-Specific Scaling. | Natalia Nessler, Milos Cernak, Paolo Prandoni, Pablo Mainar |
| 2020 | ICASSP | A Bin Encoding Training of a Spiking Neural Network Based Voice Activity Detection. | Giorgia Dellaferrera, Flavio Martinelli, Milos Cernak |
| 2020 | ICASSP | Spiking Neural Networks Trained With Backpropagation for Low Power Neuromorphic Implementation of Voice Activity Detection. | Flavio Martinelli, Giorgia Dellaferrera, Pablo Mainar, Milos Cernak |
| 2020 | Interspeech | Deep Speech Inpainting of Time-Frequency Masks. | Mikolaj Kegler, Pierre Beckmann, Milos Cernak |
| 2019 | Interspeech | Phone-Attribute Posteriors to Evaluate the Speech of Cochlear Implant Users. | Toms Arias-Vergara, Juan Rafael Orozco-Arroyave, Milos Cernak, Sandra Gollwitzer, Maria Schuster, Elmar Nth |
| 2019 | Interspeech | Evaluating Audiovisual Source Separation in the Context of Video Conferencing. | Berkay Inan, Milos Cernak, Helmut Grabner, Helena Peic Tukuljac, Rodrigo C. G. Pena, Benjamin Ricaud |
| 2019 | Interspeech | Open-Vocabulary Keyword Spotting with Audio and Text Embeddings. | Niccol Sacchi, Alexandre Nanchen, Martin Jaggi, Milos Cernak |
| 2019 | Interspeech | End-to-End Accented Speech Recognition. | Thibault Viglino, Petr Motlcek, Milos Cernak |
| 2018 | ICASSP | Nasal Speech Sounds Detection Using Connectionist Temporal Classification. | Milos Cernak, Sibo Tong |
| 2017 | ICASSP | On the impact of non-modal phonation on phonological features. | Milos Cernak, Elmar Nth, Frank Rudzicz, Heidi Christensen, Juan Rafael Orozco-Arroyave, Raman Arora, Tobias Bocklet, Hamidreza Chinaei, Julius Hannink, Phani Sankar Nidadavolu, Juan Camilo Vsquez-Correa, Maria Yancheva, Alyssa Vann, Nikolai Vogler |
| 2017 | ICASSP | Multi-view representation learning via gcca for multimodal analysis of Parkinson's disease. | Juan Camilo Vsquez-Correa, Juan Rafael Orozco-Arroyave, Raman Arora, Elmar Nth, Najim Dehak, Heidi Christensen, Frank Rudzicz, Tobias Bocklet, Milos Cernak, Hamid R. Chinaei, Julius Hannink, Phani Sankar Nidadavolu, Maria Yancheva, Alyssa Vann, Nikolai Vogler |
| 2017 | Interspeech | Bob Speaks Kaldi. | Milos Cernak, Alain Komaty, Amir Mohammadi, Andr Anjos, Sbastien Marcel |
| 2016 | Interspeech | Phonetic and Phonological Posterior Search Space Hashing Exploiting Class-Specific Sparsity Structures. | Afsaneh Asaei, Gil Luyet, Milos Cernak, Herv Bourlard |
| 2016 | Interspeech | Sound Pattern Matching for Automatic Prosodic Event Detection. | Milos Cernak, Afsaneh Asaei, Pierre-Edouard Honnet, Philip N. Garner, Herv Bourlard |
| 2016 | Interspeech | PhonVoc: A Phonetic and Phonological Vocoding Toolkit. | Milos Cernak, Philip N. Garner |
| 2016 | Interspeech | Probabilistic Amplitude Demodulation Features in Speech Synthesis for Improving Prosody. | Alexandros Lazaridis, Milos Cernak, Philip N. Garner |
| 2016 | Interspeech | HMM-Based Non-Native Accent Assessment Using Posterior Features. | Ramya Rasipuram, Milos Cernak, Mathew Magimai-Doss |
| 2015 | ICASSP | Phonological vocoding using artificial neural networks. | Milos Cernak, Blaise Potard, Philip N. Garner |
| 2015 | Interspeech | On compressibility of neural network phonological features for low bit rate speech coding. | Afsaneh Asaei, Milos Cernak, Herv Bourlard |
| 2015 | Interspeech | An empirical model of emphatic word detection. | Milos Cernak, Pierre-Edouard Honnet |
| 2015 | Interspeech | Neuromorphic based oscillatory device for incremental syllable boundary detection. | Alexandre Hyafil, Milos Cernak |
| 2015 | Interspeech | Automatic accentedness evaluation of non-native speech using phonetic and sub-phonetic posterior probabilities. | Ramya Rasipuram, Milos Cernak, Alexandre Nanchen, Mathew Magimai-Doss |
| 2014 | Interspeech | Stress and accent transmission in HMM-based syllable-context very low bit rate speech coding. | Milos Cernak, Alexandros Lazaridis, Philip N. Garner, Petr Motlcek |
| 2014 | Interspeech | Development of bilingual ASR system for MediaParl corpus. | Petr Motlcek, David Imseng, Milos Cernak, Namhoon Kim |
| 2013 | ACII | Automatic Staging of Audio with Emotions. | Lakshmi Babu Saheer, Milos Cernak |
| 2013 | ICASSP | On the (UN)importance of the contextual factors in HMM-based speech synthesis and coding. | Milos Cernak, Petr Motlcek, Philip N. Garner |
| 2013 | Interspeech | Syllable-based pitch encoding for low bit rate speech coding with recognition/synthesis architecture. | Milos Cernak, Xingyu Na, Philip N. Garner |
| 2012 | Interspeech | Robust triphone mapping for acoustic modeling. | Milos Cernak, David Imseng, Herv Bourlard |
| 2011 | Interspeech | Effective Triphone Mapping for Acoustic Modeling in Speech Recognition. | Sakhia Darjaa, Milos Cernak, Marin Trnka, Milan Rusko, Rbert Sabo |
| 2006 | ICASSP | Unit Selection Speech Synthesis in Noise. | Milos Cernak |
| 2005 | ICASSP | TTSBOX: a MATLAB toolbox for teaching text-to-speech synthesis. | Thierry Dutoit, Milos Cernak |