| 2025 | ICASSP | MARS6: A Small and Robust Hierarchical-Codec Text-to-Speech Model. | Matthew Baas, Pieter Scholtz, Arnav Mehta, Elliott Dyson, Akshat Prakash, Herman Kamper |
| 2025 | ICASSP | Speech Recognition for Automatically Assessing Afrikaans and isiXhosa Preschool Oral Narratives. | Christiaan Jacobs, Annelien Smith, Daleen Klop, Ondrej Klejch, Febe de Wet, Herman Kamper |
| 2025 | ICASSP | Unsupervised Word Discovery: Boundary Detection with Clustering vs. Dynamic Programming. | Simon Malan, Benjamin van Niekerk, Herman Kamper |
| 2025 | Interspeech | LinearVC: Linear Transformations of Self-Supervised Features Through the Lens of Voice Conversion. | Herman Kamper, Benjamin van Niekerk, Julian Zadi, Marc-Andr Carbonneau |
| 2025 | Interspeech | The mutual exclusivity bias of bilingual visually grounded speech models. | Dan Oneata, Leanne Nortje, Yevgen Matusevych, Herman Kamper |
| 2025 | Interspeech | Spoken Language Modeling with Duration-Penalized Self-Supervised Units. | Nicol Visser, Herman Kamper |
| 2024 | Interspeech | Spoken-Term Discovery using Discrete Speech Units. | Benjamin van Niekerk, Julian Zadi, Marc-Andr Carbonneau, Herman Kamper |
| 2024 | Interspeech | Translating speech with just images. | Dan Oneata, Herman Kamper |
| 2023 | Interspeech | Voice Conversion With Just Nearest Neighbors. | Matthew Baas, Benjamin van Niekerk, Herman Kamper |
| 2023 | Interspeech | Towards hate speech detection in low-resource languages: Comparing ASR to acoustic word embeddings on Wolof and Swahili. | Christiaan Jacobs, Nathanal Carraz Rakotonirina, Everlyn Asiko Chimoto, Bruce A. Bassett, Herman Kamper |
| 2023 | Interspeech | Mitigating Catastrophic Forgetting for Few-Shot Spoken Word Classification Through Meta-Learning. | Ruan van der Merwe, Herman Kamper |
| 2023 | Interspeech | Visually grounded few-shot word acquisition with fewer shots. | Leanne Nortje, Benjamin van Niekerk, Herman Kamper |
| 2022 | ICASSP | A Comparison of Discrete and Soft Speech Units for Improved Voice Conversion. | Benjamin van Niekerk, Marc-Andr Carbonneau, Julian Zadi, Matthew Baas, Hugo Seut, Herman Kamper |
| 2022 | Interspeech | Voice Conversion Can Improve ASR in Very Low-Resource Settings. | Matthew Baas, Herman Kamper |
| 2022 | Interspeech | A Temporal Extension of Latent Dirichlet Allocation for Unsupervised Acoustic Unit Discovery. | Werner van der Merwe, Herman Kamper, Johan Adam du Preez |
| 2021 | EACL | A phonetic model of non-native spoken word processing. | Yevgen Matusevych, Herman Kamper, Thomas Schatz, Naomi Feldman, Sharon Goldwater |
| 2021 | Interspeech | Multilingual Transfer of Acoustic Word Embeddings Improves When Training on Languages Related to the Target Zero-Resource Language. | Christiaan Jacobs, Herman Kamper |
| 2021 | Interspeech | Towards Unsupervised Phone and Word Segmentation Using Self-Supervised Vector-Quantized Neural Networks. | Herman Kamper, Benjamin van Niekerk |
| 2021 | Interspeech | Analyzing Speaker Information in Self-Supervised Models to Improve Zero-Resource Speech Processing. | Benjamin van Niekerk, Leanne Nortje, Matthew Baas, Herman Kamper |
| 2021 | Interspeech | Direct Multimodal Few-Shot Learning of Speech and Images. | Leanne Nortje, Herman Kamper |
| 2021 | Interspeech | Attention-Based Keyword Localisation in Speech Using Visual Grounding. | Kayode Olaleye, Herman Kamper |
| 2020 | CogSci | Evaluating computational models of infant phonetic learning across languages. | Yevgen Matusevych, Thomas Schatz, Herman Kamper, Naomi Feldman, Sharon Goldwater |
| 2020 | EMNLP | Participatory Research for Low-resourced Machine Translation: A Case Study in African Languages. | Wilhelmina Nekoto, Vukosi Marivate, Tshinondiwa Matsila, Timi E. Fasubaa, Taiwo Fagbohungbe, Solomon Oluwole Akinola, Shamsuddeen Hassan Muhammad, Salomon Kabongo Kabenamualu, Salomey Osei, Freshia Sackey, Rubungo Andre Niyongabo, Ricky Macharm, Perez Ogayo, Orevaoghene Ahia, Musie Meressa Berhe, Mofetoluwa Adeyemi, Masabata Mokgesi-Selinga, Lawrence Okegbemi, Laura Martinus, Kolawole Tajudeen, Kevin Degila, Kelechi Ogueji, Kathleen Siminyu, Julia Kreutzer, Jason Webster, Jamiil Toure Ali, Jade Z. Abbott, Iroro Orife, Ignatius Ezeani, Idris Abdulkabir Dangana, Herman Kamper, Hady Elsahar, Goodness Duru, Ghollah Kioko, Espoir Murhabazi, Elan Van Biljon, Daniel Whitenack, Christopher Onyefuluchi, Chris Chinenye Emezue, Bonaventure F. P. Dossou, Blessing K. Sibanda, Blessing Itoro Bassey, Ayodele Olabiyi, Arshath Ramkilowan, Alp ktem, Adewale Akinfaderin, Abdallah Bashir |
| 2020 | ICASSP | Cross-Lingual Topic Prediction For Speech Using Translations. | Sameer Bansal, Herman Kamper, Adam Lopez, Sharon Goldwater |
| 2020 | ICASSP | Multilingual Acoustic Word Embedding Models for Processing Zero-resource Languages. | Herman Kamper, Yevgen Matusevych, Sharon Goldwater |
| 2020 | Interspeech | Vector-Quantized Neural Networks for Acoustic Unit Discovery in the ZeroSpeech 2020 Challenge. | Benjamin van Niekerk, Leanne Nortje, Herman Kamper |
| 2020 | Interspeech | Unsupervised vs. Transfer Learning for Multimodal One-Shot Matching of Speech and Images. | Leanne Nortje, Herman Kamper |
| 2019 | ICASSP | Multimodal One-shot Learning of Speech and Images. | Ryan Eloff, Herman A. Engelbrecht, Herman Kamper |
| 2019 | ICASSP | Truly Unsupervised Acoustic Word Embeddings Using Weak Top-down Constraints in Encoder-decoder Models. | Herman Kamper |
| 2019 | ICASSP | Semantic Query-by-example Speech Search Using Visual Grounding. | Herman Kamper, Aristotelis Anastassiou, Karen Livescu |
| 2019 | Interspeech | Unsupervised Acoustic Unit Discovery for Speech Synthesis Using Discrete Latent-Variable Neural Networks. | Ryan Eloff, Andr Nortje, Benjamin van Niekerk, Avashna Govender, Leanne Nortje, Arnu Pretorius, Elan Van Biljon, Ewald van der Westhuizen, Lisa van Staden, Herman Kamper |
| 2019 | Interspeech | Feature Exploration for Almost Zero-Resource ASR-Free Keyword Spotting Using a Multilingual Bottleneck Extractor and Correspondence Autoencoders. | Raghav Menon, Herman Kamper, Ewald van der Westhuizen, John A. Quinn, Thomas Niesler |
| 2019 | Interspeech | On the Contributions of Visual and Textual Supervision in Low-Resource Semantic Speech Retrieval. | Ankita Pasad, Bowen Shi, Herman Kamper, Karen Livescu |
| 2019 | NAACL | Pre-training on high-resource speech recognition improves low-resource speech-to-text translation. | Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, Sharon Goldwater |
| 2018 | CVPR | Semantic Speech Retrieval With a Visually Grounded Model of Untranscribed Speech. | Herman Kamper, Gregory Shakhnarovich, Karen Livescu |
| 2018 | ICASSP | Phoneme Based Embedded Segmental K-Means for Unsupervised Term Discovery. | Saurabchiand Bhati, Herman Kamper, K. Sri Rama Murty |
| 2018 | ICML | Learning Dynamics of Linear Denoising Autoencoders. | Arnu Pretorius, Steve Kroon, Herman Kamper |
| 2018 | Interspeech | Low-Resource Speech-to-Text Translation. | Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, Sharon Goldwater |
| 2018 | Interspeech | Fast ASR-free and Almost Zero-resource Keyword Spotting Using DTW and CNNs for Humanitarian Monitoring. | Raghav Menon, Herman Kamper, John A. Quinn, Thomas Niesler |
| 2017 | ASRU | An embedded segmental K-means model for unsupervised segmentation and clustering of speech. | Herman Kamper, Karen Livescu, Sharon Goldwater |
| 2017 | EACL | Towards speech-to-text translation without speech recognition. | Sameer Bansal, Herman Kamper, Adam Lopez, Sharon Goldwater |
| 2017 | ICASSP | Weakly supervised spoken term discovery using cross-lingual side information. | Sameer Bansal, Herman Kamper, Sharon Goldwater, Adam Lopez |
| 2017 | Interspeech | Visually Grounded Learning of Keyword Prediction from Untranscribed Speech. | Herman Kamper, Shane Settle, Gregory Shakhnarovich, Karen Livescu |
| 2017 | Interspeech | Query-by-Example Search with Discriminative Neural Acoustic Word Embeddings. | Shane Settle, Keith D. Levin, Herman Kamper, Karen Livescu |
| 2016 | ICASSP | Deep convolutional acoustic word embeddings using word-pair side information. | Herman Kamper, Weiran Wang, Karen Livescu |
| 2015 | ICASSP | Unsupervised neural network based feature extraction using weak top-down constraints. | Herman Kamper, Micha Elsner, Aren Jansen, Sharon Goldwater |
| 2015 | Interspeech | Fully unsupervised small-vocabulary speech recognition using a segmental Bayesian model. | Herman Kamper, Aren Jansen, Sharon Goldwater |
| 2015 | Interspeech | A comparison of neural network methods for unsupervised representation learning on the zero resource speech challenge. | Daniel Renshaw, Herman Kamper, Aren Jansen, Sharon Goldwater |
| 2011 | Interspeech | Multi-Accent Speech Recognition of Afrikaans, Black and White Varieties of South African English. | Herman Kamper, Thomas Niesler |