| 2025 | Interspeech | Investigating Affect Mining Techniques for Annotation Sample Selection in the Creation of Finnish Affective Speech Corpus. | Kalle Lahtinen, Einari Vaaras, Liisa Mustanoja, Okko Rsnen |
| 2024 | CogSci | Age-Dependent Analysis and Stochastic Generation of Child-Directed Speech. | Okko Rsnen, Daniil Kocharov |
| 2024 | Interspeech | The Difficulty and Importance of Estimating the Lower and Upper Bounds of Infant Speech Exposure. | Joseph Coffey, Okko Rsnen, Camila Scaff, Alejandrina Cristi |
| 2023 | CogSci | Analysing the Impact of Audio Quality on the Use of Naturalistic Long-Form Recordings for Infant-Directed Speech Research. | Mara Andrea Cruz Blandn, Alejandrina Cristi, Okko Rsnen |
| 2023 | CogSci | Computational Insights to Acquisition of Phonemes, Words, and Word Meanings in Early Language: Sequential or Parallel Acquisition? | Khazar Khorrami, Mara Andrea Cruz Blandn, Okko Rsnen |
| 2023 | CogSci | Is Reliability of Cognitive Measures in Children Dependent on Participant Age? A Case Study with Two Large-Scale Datasets. | Okko Rsnen, Mara Andrea Cruz Blandn, Jukka Leppnen |
| 2023 | ICASSP | On Negative Sampling for Contrastive Audio-Text Retrieval. | Huang Xie, Okko Rsnen, Tuomas Virtanen |
| 2023 | Interspeech | BabySLM: language-acquisition-friendly benchmark of self-supervised spoken language models. | Marvin Lavechin, Yaya Sy, Hadrien Titeux, Mara Andrea Cruz Blandn, Okko Rsnen, Herv Bredin, Emmanuel Dupoux, Alejandrina Cristi |
| 2023 | Interspeech | Syllable Discovery and Cross-Lingual Generalization in a Visually Grounded, Self-Supervised Speech Model. | Puyuan Peng, Shang-Wen Li, Okko Rsnen, Abdelrahman Mohamed, David Harwath |
| 2022 | ICASSP | Unsupervised Audio-Caption Aligning Learns Correspondences Between Individual Sound Events and Textual Phrases. | Huang Xie, Okko Rsnen, Konstantinos Drossos, Tuomas Virtanen |
| 2022 | Interspeech | Analysis of Self-Supervised Learning and Dimensionality Reduction Methods in Clustering-Based Active Learning for Speech Emotion Recognition. | Einari Vaaras, Manu Airaksinen, Okko Rsnen |
| 2021 | ICASSP | Zero-Shot Audio Classification with Factored Linear and Nonlinear Acoustic-Semantic Projections. | Huang Xie, Okko Rsnen, Tuomas Virtanen |
| 2021 | Interspeech | Evaluation of Audio-Visual Alignments in Visually Grounded Speech Models. | Khazar Khorrami, Okko Rsnen |
| 2021 | Interspeech | Automatic Analysis of the Emotional Content of Speech in Daylong Child-Centered Recordings from a Neonatal Intensive Care Unit. | Einari Vaaras, Sari Ahlqvist-Bjrkroth, Konstantinos Drossos, Okko Rsnen |
| 2020 | CogSci | Measuring prosodic predictability in children's home language environments. | Kyle MacDonald, Marisa Casillas, Okko Rsnen, Anne S. Warlaumont |
| 2020 | Interspeech | Unsupervised Discovery of Recurring Speech Patterns Using Probabilistic Adaptive Metrics. | Okko Rsnen, Mara Andrea Cruz Blandn |
| 2019 | ICASSP | Data Augmentation Strategies for Neural Network F0 Estimation. | Manu Airaksinen, Lauri Juvela, Paavo Alku, Okko Rsnen |
| 2019 | ICASSP | Cycle-consistent Adversarial Networks for Non-parallel Vocal Effort Based Speaking Style Conversion. | Shreyas Seshadri, Lauri Juvela, Junichi Yamagishi, Okko Rsnen, Paavo Alku |
| 2019 | Interspeech | A Computational Model of Early Language Acquisition from Audiovisual Experiences of Young Infants. | Okko Rsnen, Khazar Khorrami |
| 2019 | Interspeech | Augmented CycleGANs for Continuous Scale Normal-to-Lombard Speaking Style Conversion. | Shreyas Seshadri, Lauri Juvela, Paavo Alku, Okko Rsnen |
| 2018 | Interspeech | Time-regularized Linear Prediction for Noise-robust Extraction of the Spectral Envelope of Speech. | Manu Airaksinen, Lauri Juvela, Okko Rsnen, Paavo Alku |
| 2018 | Interspeech | Comparison of Syllabification Algorithms and Training Strategies for Robust Word Count Estimation across Different Languages and Recording Conditions. | Okko Rsnen, Shreyas Seshadri, Marisa Casillas |
| 2017 | ACL | Blind Phoneme Segmentation With Temporal Prediction Errors. | Paul Michel, Okko Rsnen, Roland Thiollire, Emmanuel Dupoux |
| 2017 | CogSci | Connecting stimulus-driven attention to the properties of infant-directed speech - Is exaggerated intonation also more surprising? | Okko Rsnen, Sofoklis Kakouros, Melanie Soderstrom |
| 2017 | ICASSP | Dirichlet process mixture models for clustering i-vector data. | Shreyas Seshadri, Ulpu Remes, Okko Rsnen |
| 2017 | Interspeech | Evaluation of Spectral Tilt Measures for Sentence Prominence Under Different Noise Conditions. | Sofoklis Kakouros, Okko Rsnen, Paavo Alku |
| 2017 | Interspeech | Speaking Style Conversion from Normal to Lombard Speech Using a Glottal Vocoder and Bayesian GMMs. | Ana Ramrez Lpez, Shreyas Seshadri, Lauri Juvela, Okko Rsnen, Paavo Alku |
| 2017 | Interspeech | Comparison of Non-Parametric Bayesian Mixture Models for Syllable Clustering and Zero-Resource Speech Processing. | Shreyas Seshadri, Ulpu Remes, Okko Rsnen |
| 2016 | CogSci | Statistical Learning of Prosodic Patterns and Reversal of Perceptual Cues for Sentence Prominence. | Sofoklis Kakouros, Okko Rsnen |
| 2016 | CogSci | A Cognitive Approach to Modeling Sentence Level Prominence Based on Stimulus Unpredictability. | Sofoklis Kakouros, Okko Rsnen |
| 2016 | CogSci | Analyzing distributional learning of phonemic categories in unsupervised deep neural networks. | Okko Rsnen, Tasha Nagamine, Nima Mesgarani |
| 2016 | Interspeech | Analyzing the Contribution of Top-Down Lexical and Bottom-Up Acoustic Cues in the Detection of Sentence Prominence. | Sofoklis Kakouros, Joris Pelemans, Lyan Verwimp, Patrick Wambacq, Okko Rsnen |
| 2015 | Interspeech | Automatic detection of sentence prominence in speech using predictability of word-level acoustic features. | Sofoklis Kakouros, Okko Rsnen |
| 2015 | Interspeech | Unsupervised word discovery from speech using automatic segmentation into syllable-like units. | Okko Rsnen, Gabriel Doyle, Michael C. Frank |
| 2015 | Interspeech | Weakly-supervised word learning is improved by an active online algorithm. | Heikki Rasilo, Okko Rsnen |
| 2014 | CogSci | Basic cuts revisited: Temporal segmentation of speech into phone-like units with statistical learning at a pre-linguistic level. | Okko Rsnen |
| 2014 | Interspeech | Perception of sentence stress in English infant directed speech. | Sofoklis Kakouros, Okko Rsnen |
| 2013 | Interspeech | Automatic self-supervised learning of associations between speech and text. | Juha Knuuttila, Okko Rsnen, Unto K. Laine |
| 2013 | Interspeech | Random subset feature selection in automatic recognition of developmental disorders, affective states, and level of conflict from speech. | Okko Rsnen, Jouni Pohjalainen |
| 2012 | CogSci | Acoustic analysis supports the existence of a single distributional learning mechanism in structural rule learning from an artificial language. | Okko Rsnen, Heikki Rasilo |
| 2012 | ICASSP | Hierarchical unsupervised discovery of user context from multivariate sensory data. | Okko Rsnen |
| 2012 | ICASSP | Context induced merging of synonymous word models in computational modeling of early language acquisition. | Okko Rsnen |
| 2012 | Interspeech | Feature Selection for Speaker Traits. | Jouni Pohjalainen, Serdar Kadioglu, Okko Rsnen |
| 2012 | Interspeech | Non-auditory cognitive capabilities in computational modeling of early language acquisition. | Okko Rsnen |
| 2012 | Interspeech | Average Spectrotemporal Structure of Continuous Speech Matches with the Frequency Resolution of Human Hearing. | Okko Rsnen |
| 2012 | Interspeech | Modeling spoken language acquisition with a generic cognitive architecture for associative learning. | Okko Rsnen, Heikki Rasilo, Unto K. Laine |