| 2025 | Interspeech | From Static to Dynamic: Enhancing AAC with Generative Imagery and Zero-Shot TTS. | Juliana Francis, Joakim Gustafson, va Szkely |
| 2025 | Interspeech | VoiceQualityVC: A Voice Conversion System for Studying the Perceptual Effects of Voice Quality in Speech. | Harm Lameris, Joakim Gustafson, va Szkely |
| 2025 | Interspeech | Towards Adaptable and Intelligible Speech Synthesis in Noisy Environments. | Lubos Marcinek, Jonas Beskow, Joakim Gustafson |
| 2025 | SIGdial | Role of Reasoning in LLM Enjoyment Detection: Evaluation Across Conversational Levels for Human-Robot Interaction. | Lubos Marcinek, Bahar Irfan, Gabriel Skantze, Andr Pereira, Joakim Gustafson |
| 2024 | COLING | The Role of Creaky Voice in Turn Taking and the Perception of Speaker Stance: Experiments Using Controllable TTS. | Harm Lameris, va Szkely, Joakim Gustafson |
| 2024 | COLING | Revisiting Three Text-to-Speech Synthesis Experiments with a Web-Based Audience Response System. | Christina Tnnander, Jens Edlund, Joakim Gustafson |
| 2024 | ICMI | Multimodal User Enjoyment Detection in Human-Robot Conversation: The Power of Large Language Models. | Andr Pereira, Lubos Marcinek, Jura Miniota, Sofia Thunberg, Erik Lagerstedt, Joakim Gustafson, Gabriel Skantze, Bahar Irfan |
| 2024 | Interspeech | ConnecTone: a modular AAC system prototype with contextual generative text prediction and style-adaptive conversational TTS. | Juliana Francis, va Szkely, Joakim Gustafson |
| 2024 | Interspeech | CreakVC: a voice conversion tool for modulating creaky voice. | Harm Lameris, Joakim Gustafson, va Szkely |
| 2024 | Interspeech | Contextual Interactive Evaluation of TTS Models in Dialogue Systems. | Siyang Wang, va Szkely, Joakim Gustafson |
| 2023 | CHI | A Special Interest Group on Developing Theories of Language Use in Interaction with Conversational User Interfaces. | Paola Raquel Pea, Philip R. Doyle, Emily Yj Ip, Giovanni M. Di Liberto, Darragh Higgins, Rachel McDonnell, Holly P. Branigan, Joakim Gustafson, Donald McMillan, Robert J. Moore, Benjamin R. Cowan |
| 2023 | ICASSP | Prosody-Controllable Spontaneous TTS with Neural HMMS. | Harm Lameris, Shivam Mehta, Gustav Eje Henter, Joakim Gustafson, va Szkely |
| 2023 | ICASSP | A Comparative Study of Self-Supervised Speech Representations in Read and Spontaneous TTS. | Siyang Wang, Gustav Eje Henter, Joakim Gustafson, va Szkely |
| 2023 | Interspeech | Automatic Evaluation of Turn-taking Cues in Conversational Speech Synthesis. | Erik Ekstedt, Siyang Wang, va Szkely, Joakim Gustafson, Gabriel Skantze |
| 2023 | Interspeech | Pardon my disfluency: The impact of disfluency effects on the perception of speaker competence and confidence. | Ambika Kirkland, Joakim Gustafson, va Szkely |
| 2023 | Interspeech | Beyond Style: Synthesizing Speech with Pragmatic Functions. | Harm Lameris, Joakim Gustafson, va Szkely |
| 2023 | Interspeech | Prosody-controllable Gender-ambiguous Speech Synthesis: A Tool for Investigating Implicit Bias in Speech Perception. | va Szkely, Joakim Gustafson, Ilaria Torre |
| 2023 | Interspeech | So-to-Speak: An Exploratory Platform for Investigating the Interplay between Style and Prosody in TTS. | va Szkely, Siyang Wang, Joakim Gustafson |
| 2023 | IVA | Generation of speech and facial animation with controllable articulatory effort for amusing conversational characters. | Joakim Gustafson, va Szkely, Jonas Beskow |
| 2023 | RO-MAN | Hi robot, it's not what you say, it's how you say it. | Jura Miniota, Siyang Wang, Jonas Beskow, Joakim Gustafson, va Szkely, Andr Pereira |
| 2022 | Interspeech | Where's the uh, hesitation? The interplay between filled pause location, speech rate and fundamental frequency in perception of confidence. | Ambika Kirkland, Harm Lameris, va Szkely, Joakim Gustafson |
| 2022 | LREC | Evaluating Sampling-based Filler Insertion with Spontaneous TTS. | Siyang Wang, Joakim Gustafson, va Szkely |
| 2021 | ICMI | A Systematic Cross-Corpus Analysis of Human Reactions to Robot Conversational Failures. | Dimosthenis Kontogiorgos, Minh Tran, Joakim Gustafson, Mohammad Soleymani |
| 2021 | ICMI | Integrated Speech and Gesture Synthesis. | Siyang Wang, Simon Alexanderson, Joakim Gustafson, Jonas Beskow, Gustav Eje Henter, va Szkely |
| 2020 | CHI | Embodiment Effects in Interactions with Failing Robots. | Dimosthenis Kontogiorgos, Sanne van Waveren, Olle Wallberg, Andr Pereira, Iolanda Leite, Joakim Gustafson |
| 2020 | HRI | Behavioural Responses to Robot Conversational Failures. | Dimosthenis Kontogiorgos, Andr Pereira, Boran Sahindal, Sanne van Waveren, Joakim Gustafson |
| 2020 | HRI | Effects of Different Interaction Contexts when Evaluating Gaze Models in HRI. | Andr Pereira, Catharine Oertel, Leonor Fermoselle, Joseph Mendelson, Joakim Gustafson |
| 2020 | ICASSP | Breathing and Speech Planning in Spontaneous Speech Synthesis. | va Szkely, Gustav Eje Henter, Jonas Beskow, Joakim Gustafson |
| 2020 | LREC | Chinese Whispers: A Multimodal Dataset for Embodied Language Grounding. | Dimosthenis Kontogiorgos, Elena Sibirtseva, Joakim Gustafson |
| 2020 | LREC | Augmented Prompt Selection for Evaluation of Spontaneous Speech Synthesis. | va Szkely, Jens Edlund, Joakim Gustafson |
| 2019 | CogSci | The Effects of Embodiment and Social Eye-Gaze in Conversational Agents. | Dimosthenis Kontogiorgos, Gabriel Skantze, Andr Pereira, Joakim Gustafson |
| 2019 | ICASSP | Casting to Corpus: Segmenting and Selecting Spontaneous Dialogue for Tts with a Cnn-lstm Speaker-dependent Breath Detector. | va Szkely, Gustav Eje Henter, Joakim Gustafson |
| 2019 | ICMI | Estimating Uncertainty in Task-Oriented Dialogue. | Dimosthenis Kontogiorgos, Andr Pereira, Joakim Gustafson |
| 2019 | Interspeech | Off the Cuff: Exploring Extemporaneous Speech Delivery with TTS. | va Szkely, Gustav Eje Henter, Jonas Beskow, Joakim Gustafson |
| 2019 | Interspeech | Spontaneous Conversational Speech Synthesis from Found Data. | va Szkely, Gustav Eje Henter, Jonas Beskow, Joakim Gustafson |
| 2019 | IROS | Responsive Joint Attention in Human-Robot Interaction. | Andr Pereira, Catharine Oertel, Leonor Fermoselle, Joe Mendelson, Joakim Gustafson |
| 2019 | IVA | The Effects of Anthropomorphism and Non-verbal Social Behaviour in Virtual Assistants. | Dimosthenis Kontogiorgos, Andr Pereira, Olle Andersson, Marco Koivisto, Elena Gonzalez Rabal, Ville Vartiainen, Joakim Gustafson |
| 2018 | ICMI | Multimodal Reference Resolution In Collaborative Assembly Tasks. | Dimosthenis Kontogiorgos, Elena Sibirtseva, Andr Pereira, Gabriel Skantze, Joakim Gustafson |
| 2018 | IJCAI | Interactive, Collaborative Robots: Challenges and Opportunities. | Danica Kragic, Joakim Gustafson, Hakan Karaoguz, Patric Jensfelt, Robert Krug |
| 2018 | LREC | Crowdsourced Multimodal Corpora Collection Tool. | Patrik Jonell, Catharine Oertel, Dimosthenis Kontogiorgos, Jonas Beskow, Joakim Gustafson |
| 2018 | LREC | A Multimodal Corpus for Mutual Gaze and Joint Attention in Multiparty Situated Interaction. | Dimosthenis Kontogiorgos, Vanya Avramova, Simon Alexandersson, Patrik Jonell, Catharine Oertel, Jonas Beskow, Gabriel Skantze, Joakim Gustafson |
| 2018 | RO-MAN | A Comparison of Visualisation Methods for Disambiguating Verbal Requests in Human-Robot Interaction. | Elena Sibirtseva, Dimosthenis Kontogiorgos, Olov Nykvist, Hakan Karaoguz, Iolanda Leite, Joakim Gustafson, Danica Kragic |
| 2017 | ICMI | Using crowd-sourcing for the design of listening agents: challenges and opportunities. | Catharine Oertel, Patrik Jonell, Kevin El Haddad, va Szkely, Joakim Gustafson |
| 2017 | Interspeech | Controlling Prominence Realisation in Parametric DNN-Based Speech Synthesis. | Zofia Malisz, Harald Berthelsen, Jonas Beskow, Joakim Gustafson |
| 2017 | Interspeech | Crowd-Sourced Design of Artificial Attentive Listeners. | Catharine Oertel, Patrik Jonell, Dimosthenis Kontogiorgos, Joseph Mendelson, Jonas Beskow, Joakim Gustafson |
| 2017 | Interspeech | Synthesising Uncertainty: The Interplay of Vocal Effort and Hesitation Disfluencies. | va Szkely, Joseph Mendelson, Joakim Gustafson |
| 2017 | IVA | Crowd-Powered Design of Virtual Attentive Listeners. | Patrik Jonell, Catharine Oertel, Dimosthenis Kontogiorgos, Jonas Beskow, Joakim Gustafson |
| 2016 | ICMI | On data driven parametric backchannel synthesis for expressing attentiveness in conversational agents. | Catharine Oertel, Joakim Gustafson, Alan W. Black |
| 2016 | ICMI | Towards building an attentive artificial listener: on the perception of attentiveness in audio-visual feedback tokens. | Catharine Oertel, Jos Lopes, Yu Yu, Kenneth Alberto Funes Mora, Joakim Gustafson, Alan W. Black, Jean-Marc Odobez |
| 2016 | Interspeech | Towards Building an Attentive Artificial Listener: On the Perception of Attentiveness in Feedback Utterances. | Catharine Oertel, Joakim Gustafson, Alan W. Black |
| 2016 | LREC | Hidden Resources ― Strategies to Acquire and Exploit Potential Spoken Language Resources in National Archives. | Jens Edlund, Joakim Gustafson |
| 2015 | ICMI | Deciphering the Silent Participant: On the Use of Audio-Visual Cues for the Classification of Listener Categories in Group Discussions. | Catharine Oertel, Kenneth Alberto Funes Mora, Joakim Gustafson, Jean-Marc Odobez |
| 2015 | Interspeech | Detecting repetitions in spoken dialogue systems using phonetic distances. | Jos Lopes, Giampiero Salvi, Gabriel Skantze, Alberto Abad, Joakim Gustafson, Fernando Batista, Raveesh Meena, Isabel Trancoso |
| 2015 | SIGdial | Automatic Detection of Miscommunication in Spoken Dialogue Systems. | Raveesh Meena, Jos Lopes, Gabriel Skantze, Joakim Gustafson |
| 2014 | EACL | Human pause and resume behaviours for unobtrusive humanlike in-car spoken dialogue systems. | Jens Edlund, Fredrik Edelstam, Joakim Gustafson |
| 2014 | HRI | Human-robot collaborative tutoring using multiparty multimodal spoken dialogue. | Samer Al Moubayed, Jonas Beskow, Bajibabu Bollepalli, Joakim Gustafson, Ahmed Hussen Abdelaziz, Martin Johansson, Maria Koutsombogera, Jos David guas Lopes, Jekaterina Novikova, Catharine Oertel, Gabriel Skantze, Kalin Stefanov, Gl Varol |
| 2014 | ICASSP | A comparative evaluation of vocoding techniques for HMM-based laughter synthesis. | Bajibabu Bollepalli, Jrme Urbain, Tuomo Raitio, Joakim Gustafson, Hseyin akmak |
| 2014 | ICMI | Comparison of Human-Human and Human-Robot Turn-Taking Behaviour in Multiparty Situated Interaction. | Martin Johansson, Gabriel Skantze, Joakim Gustafson |
| 2014 | ICMI | Who Will Get the Grant?: A Multimodal Corpus for the Analysis of Conversational Behaviours in Group Interviews. | Catharine Oertel, Kenneth Alberto Funes Mora, Samira Sheikhi, Jean-Marc Odobez, Joakim Gustafson |
| 2014 | SIGdial | Crowdsourcing Street-level Geographic Information Using a Spoken Dialogue System. | Raveesh Meena, Johan Boye, Gabriel Skantze, Joakim Gustafson |
| 2013 | Interspeech | Analysis of gaze and speech patterns in three-party quiz game interaction. | Samer Al Moubayed, Jens Edlund, Joakim Gustafson |
| 2013 | SIGdial | The Map Task Dialogue System: A Test-bed for Modelling Human-Like Dialogue. | Raveesh Meena, Gabriel Skantze, Joakim Gustafson |
| 2013 | SIGdial | A Data-driven Model for Timing Feedback in a Map Task Dialogue System. | Raveesh Meena, Gabriel Skantze, Joakim Gustafson |
| 2012 | ICMI | Multimodal multiparty social interaction with the furhat head. | Samer Al Moubayed, Gabriel Skantze, Jonas Beskow, Kalin Stefanov, Joakim Gustafson |
| 2012 | Interspeech | On the effect of the acoustic environment on the accuracy of perception of speaker orientation from auditory cues alone. | Jens Edlund, Mattias Heldner, Joakim Gustafson |
| 2012 | Interspeech | A Data-driven Approach to Understanding Spoken Route Directions in Human-Robot Dialogue. | Raveesh Meena, Gabriel Skantze, Joakim Gustafson |
| 2012 | Interspeech | Gaze Patterns in Turn-Taking. | Catharine Oertel, Marcin Wlodarczak, Jens Edlund, Petra Wagner, Joakim Gustafson |
| 2011 | Interspeech | Tracking Pitch Contours Using Minimum Jerk Trajectories. | Daniel Neiberg, Gopal Ananthakrishnan, Joakim Gustafson |
| 2011 | Interspeech | Predicting Speaker Changes and Listener Responses with and without Eye-Contact. | Daniel Neiberg, Joakim Gustafson |
| 2011 | Interspeech | A Dual Channel Coupled Decoder for Fillers and Feedback. | Daniel Neiberg, Joakim Gustafson |
| 2011 | IROS | Enhanced visual scene understanding through human-robot dialog. | Matthew Johnson-Roberson, Jeannette Bohg, Gabriel Skantze, Joakim Gustafson, Rolf Carlson, Babak Rasolzadeh, Danica Kragic |
| 2010 | Interspeech | The prosody of Swedish conversational grunts. | Daniel Neiberg, Joakim Gustafson |
| 2009 | Interspeech | The MonAMI reminder: a spoken dialogue system for face-to-face interaction. | Jonas Beskow, Jens Edlund, Bjrn Granstrm, Joakim Gustafson, Gabriel Skantze, Helena Tobiasson |
| 2009 | SIGdial | Eliciting Interactional Phenomena in Human-Human Dialogues. | Joakim Gustafson, Miray Merkes |
| 2009 | SIGdial | Attention and Interaction Control in a Human-Human-Computer Dialogue Setting. | Gabriel Skantze, Joakim Gustafson |
| 2008 | ICMI | Innovative interfaces in MonAMI: the reminder. | Jonas Beskow, Jens Edlund, Teodore Gjermani, Bjrn Granstrm, Joakim Gustafson, Oskar Jonsson, Gabriel Skantze, Helena Tobiasson |
| 2008 | Interspeech | What makes a good speaker? subject ratings, acoustic measurements and perceptual evaluations. | Eva Strangert, Joakim Gustafson |
| 2007 | Interspeech | Children's convergence in referring expressions to graphical objects in a speech-enabled computer game. | Linda Bell, Joakim Gustafson |
| 2005 | Interspeech | The Swedish NICE corpus - spoken dialogues between children and embodied characters in a computer game scenario. | Linda Bell, Johan Boye, Joakim Gustafson, Mattias Heldner, Anders Lindstrm, Mats Wirn |
| 2005 | IVA | Providing Computer Game Characters with Conversational Abilities. | Joakim Gustafson, Johan Boye, Morgan Fredriksson, Lasse Johannesson, Jrgen Knigsmann |
| 2005 | SIGdial | How to do Dialogue in a Fairy-tale World. | Johan Boye, Joakim Gustafson |
| 2004 | SIGdial | The NICE Fairy-tale Game System. | Joakim Gustafson, Linda Bell, Johan Boye, Anders Lindstrm, Mats Wirn |
| 2003 | Interspeech | Child and adult speaker adaptation during error resolution in a publicly available spoken dialogue system. | Linda Bell, Joakim Gustafson |
| 2002 | Interspeech | Voice transformations for improving children²s speech recognition in a publicly available dialogue system. | Joakim Gustafson, Kre Sjlander |
| 2000 | Interspeech | A comparison of disfluency distribution in a unimodal and a multimodal speech interface. | Linda Bell, Robert Eklund, Joakim Gustafson |
| 2000 | Interspeech | Positive and negative user feedback in a spoken dialogue corpus. | Linda Bell, Joakim Gustafson |
| 2000 | Interspeech | Adapt - a multimodal conversational dialogue system in an apartment domain. | Joakim Gustafson, Linda Bell, Jonas Beskow, Johan Boye, Rolf Carlson, Jens Edlund, Bjrn Granstrm, David House, Mats Wirn |
| 1999 | Interspeech | Interaction with an animated agent in a spoken dialogue system. | Linda Bell, Joakim Gustafson |
| 1999 | Interspeech | The august spoken dialogue system. | Joakim Gustafson, Nikolaj Lindberg, Magnus Lundeberg |
| 1998 | Interspeech | An educational dialogue system with a user controllable dialogue manager. | Joakim Gustafson, Patrik Elmberg, Rolf Carlson, Arne Jnsson |
| 1998 | Interspeech | Web-based educational tools for speech technology. | Kre Sjlander, Jonas Beskow, Joakim Gustafson, Erland Lewin, Rolf Carlson, Bjrn Granstrm |
| 1997 | Interspeech | How do system questions influence lexical choices in user answers? | Joakim Gustafson, Anette Larsson, Rolf Carlson, K. Hellman |
| 1997 | Interspeech | An integrated system for teaching spoken dialogue systems technology. | Kre Sjlander, Joakim Gustafson |
| 1995 | Interspeech | The waxholm application database. | J. Bertenstam, Mats Blomberg, Rolf Carlson, Kjell Elenius, Bjrn Granstrm, Joakim Gustafson, Sheri Hunnicutt, Jesper Hgberg, Roger Lindell, Lennart Neovius, Lennart Nord, Antonio de Serpa-Leitao, Nikko Strm |
| 1995 | Interspeech | Using two-level morphology to transcribe Swedish names. | Joakim Gustafson |
| 1993 | Interspeech | An experimental dialogue system: waxholm. | Mats Blomberg, Rolf Carlson, Kjell Elenius, Bjrn Granstrm, Joakim Gustafson, Sheri Hunnicutt, Roger Lindell, Lennart Neovius |