| 2025 | Interspeech | Multi-Teacher Language-Aware Knowledge Distillation for Multilingual Speech Emotion Recognition. | Mehedi Hasan Bijoy, Dejan Porjazovski, Tams Grsz, Mikko Kurimo |
| 2025 | Interspeech | Is your model big enough? Training and interpreting large-scale monolingual speech foundation models. | Yaroslav Getman, Tams Grsz, Tommi Lehtonen, Mikko Kurimo |
| 2025 | Interspeech | Mispronunciation Detection Without L2 Pronunciation Dataset in Low-Resource Setting: A Case Study in Finland Swedish. | Nhan Phan, Mikko Kuronen, Maria Kautonen, Riikka Ullakonoja, Anna von Zansen, Yaroslav Getman, Ekaterina Voskoboinik, Tams Grsz, Mikko Kurimo |
| 2024 | COLING | Collecting Linguistic Resources for Assessing Children's Pronunciation of Nordic Languages. | Anne Marte Haug Olstad, Anna Smolander, Sofia Strmbergsson, Sari Ylinen, Minna Lehtonen, Mikko Kurimo, Yaroslav Getman, Tams Grsz, Xinwei Cao, Torbjrn Svendsen, Giampiero Salvi |
| 2024 | ICASSP | Investigating the Clusters Discovered By Pre-Trained AV-HuBERT. | Anja Virkkunen, Marek Sarvas, Guangpu Huang, Tams Grsz, Mikko Kurimo |
| 2024 | Interspeech | Exploring adaptation techniques of large speech foundation models for low-resource ASR: a case study on Northern Smi. | Yaroslav Getman, Tams Grsz, Katri Hiovain-Asikainen, Mikko Kurimo |
| 2024 | Interspeech | What happens in continued pre-training? Analysis of self-supervised speech models with continued pre-training for colloquial Finnish ASR. | Yaroslav Getman, Tams Grsz, Mikko Kurimo |
| 2024 | Interspeech | Oversampling, Augmentation and Curriculum Learning for Speaking Assessment with Limited Training Data. | Tin Mei Lun, Ekaterina Voskoboinik, Ragheb Al-Ghezi, Tams Grsz, Mikko Kurimo |
| 2024 | Interspeech | CaptainA self-study mobile app for practising speaking: task completion assessment and feedback with generative AI. | Nhan Phan, Anna von Zansen, Maria Kautonen, Tams Grsz, Mikko Kurimo |
| 2024 | Interspeech | Automated content assessment and feedback for Finnish L2 learners in a picture description speaking task. | Nhan Phan, Anna von Zansen, Maria Kautonen, Ekaterina Voskoboinik, Tams Grsz, Raili Hildn, Mikko Kurimo |
| 2023 | Interspeech | Investigating wav2vec2 context representations and the effects of fine-tuning, a case-study of a Finnish model. | Tams Grsz, Yaroslav Getman, Ragheb Al-Ghezi, Aku Rouhe, Mikko Kurimo |
| 2022 | Interspeech | wav2vec2-based Speech Rating System for Children with Speech Sound Disorder. | Yaroslav Getman, Ragheb Al-Ghezi, Katja Voskoboinik, Tams Grsz, Mikko Kurimo, Giampiero Salvi, Torbjrn Svendsen, Sofia Strmbergsson |
| 2022 | Interspeech | Comparison and Analysis of New Curriculum Criteria for End-to-End ASR. | Georgios Karakasidis, Tams Grsz, Mikko Kurimo |
| 2020 | Interspeech | Data Augmentation Using Prosody and False Starts to Recognize Non-Native Children's Speech. | Hemant Kumar Kathania, Mittul Singh, Tams Grsz, Mikko Kurimo |
| 2019 | IJCNN | Autoencoder-Based Articulatory-to-Acoustic Mapping for Ultrasound Silent Speech Interfaces. | Gbor Gosztolya, dm Pintr, Lszl Tth, Tams Grsz, Alexandra Mark, Tams Gbor Csap |
| 2019 | Interspeech | Ultrasound-Based Silent Speech Interface Built on a Continuous Vocoder. | Tams Gbor Csap, Mohammed Salah Al-Radhi, Gza Nmeth, Gbor Gosztolya, Tams Grsz, Lszl Tth, Alexandra Mark |
| 2018 | ICASSP | F0 Estimation for DNN-Based Ultrasound Silent Speech Interfaces. | Tams Grsz, Gbor Gosztolya, Lszl Tth, Tams Gbor Csap, Alexandra Mark |
| 2018 | Interspeech | General Utterance-Level Feature Extraction for Classifying Crying Sounds, Atypical & Self-Assessed Affect and Heart Beats. | Gbor Gosztolya, Tams Grsz, Lszl Tth |
| 2018 | Interspeech | Multi-Task Learning of Speech Recognition and Speech Synthesis Parameters for Ultrasound-based Silent Speech Interfaces. | Lszl Tth, Gbor Gosztolya, Tams Grsz, Alexandra Mark, Tams Gbor Csap |
| 2017 | Interspeech | DNN-Based Ultrasound-to-Speech Conversion for a Silent Speech Interface. | Tams Gbor Csap, Tams Grsz, Gbor Gosztolya, Lszl Tth, Alexandra Mark |
| 2017 | Interspeech | DNN-Based Feature Extraction and Classifier Combination for Child-Directed Speech, Cold and Snoring Identification. | Gbor Gosztolya, Rbert Busa-Fekete, Tams Grsz, Lszl Tth |
| 2017 | Interspeech | Training Context-Dependent DNN Acoustic Models Using Probabilistic Sampling. | Tams Grsz, Gbor Gosztolya, Lszl Tth |
| 2017 | Interspeech | A Comparative Evaluation of GMM-Free State Tying Methods for ASR. | Tams Grsz, Gbor Gosztolya, Lszl Tth |
| 2016 | Interspeech | Determining Native Language and Deception Using Phonetic Features and Classifier Combination. | Gbor Gosztolya, Tams Grsz, Rbert Busa-Fekete, Lszl Tth |
| 2016 | Interspeech | Estimating the Sincerity of Apologies in Speech by DNN Rank Learning and Prosodic Analysis. | Gbor Gosztolya, Tams Grsz, Gyrgy Szaszk, Lszl Tth |
| 2016 | Interspeech | GMM-Free Flat Start Sequence-Discriminative DNN Training. | Gbor Gosztolya, Tams Grsz, Lszl Tth |
| 2016 | Interspeech | Detecting Mild Cognitive Impairment from Spontaneous Speech by Correlation-Based Phonetic Feature Selection. | Gbor Gosztolya, Lszl Tth, Tams Grsz, Veronika Vincze, Ildik Hoffmann, Grta Szatlczki, Magdolna Pkski, Jnos Klmn |
| 2015 | ICASSP | Building context-dependent DNN acoustic models using Kullback-Leibler divergence-based state tying. | Gbor Gosztolya, Tams Grsz, Lszl Tth, David Imseng |
| 2015 | Interspeech | Assessing the degree of nativeness and parkinson's condition using Gaussian processes and deep rectifier neural networks. | Tams Grsz, Rbert Busa-Fekete, Gbor Gosztolya, Lszl Tth |
| 2014 | Interspeech | Detecting the intensity of cognitive and physical load using AdaBoost and deep rectifier neural networks. | Gbor Gosztolya, Tams Grsz, Rbert Busa-Fekete, Lszl Tth |