| 2024 | COLING | ConEC: Earnings Call Dataset with Real-world Contexts for Benchmarking Contextual Speech Recognition. | Ruizhe Huang, Mahsa Yarmohammadi, Jan Trmal, Jing Liu, Desh Raj, Leibny Paola Garca, Alexei V. Ivanov, Patrick Ehlen, Mingzhi Yu, Dan Povey, Sanjeev Khudanpur |
| 2023 | ICASSP | Effectiveness of Text, Acoustic, and Lattice-Based Representations in Spoken Language Understanding Tasks. | Esa Villatoro-Tello, Srikanth R. Madikeri, Juan Zuluaga-Gomez, Bidisha Sharma, Seyyed Saeed Sarfjoo, Iuliia Nigmatulina, Petr Motlcek, Alexei V. Ivanov, Aravind Ganapathiraju |
| 2022 | SIGIR | Expanded Lattice Embeddings for Spoken Document Retrieval on Informal Meetings. | Esa Villatoro-Tello, Srikanth R. Madikeri, Petr Motlcek, Aravind Ganapathiraju, Alexei V. Ivanov |
| 2016 | Interspeech | Noise and Metadata Sensitive Bottleneck Features for Improving Speaker Recognition with Non-Native Speech Input. | Yao Qian, Jidong Tao, David Suendermann-Oeft, Keelan Evanini, Alexei V. Ivanov, Vikram Ramanarayanan |
| 2016 | SIGdial | LVCSR System on a Hybrid GPU-CPU Embedded Platform for Real-Time Dialog Applications. | Alexei V. Ivanov, Patrick L. Lange, David Suendermann-Oeft |
| 2015 | Interspeech | Pronunciation accuracy and intelligibility of non-native speech. | Anastassia Loukina, Melissa Lopez, Keelan Evanini, David Suendermann-Oeft, Alexei V. Ivanov, Klaus Zechner |
| 2015 | SIGdial | Automated Speech Recognition Technology for Dialogue Interaction with Non-Native Interlocutors. | Alexei V. Ivanov, Vikram Ramanarayanan, David Suendermann-Oeft, Melissa Lopez, Keelan Evanini, Jidong Tao |
| 2015 | SIGdial | A distributed cloud-based dialog system for conversational application development. | Vikram Ramanarayanan, David Suendermann-Oeft, Alexei V. Ivanov, Keelan Evanini |
| 2013 | ASRU | Phonetic and anthropometric conditioning of MSA-KST cognitive impairment characterization system. | Alexei V. Ivanov, Shahab Jalalvand, Roberto Gretter, Daniele Falavigna |
| 2012 | ICASSP | Kolmogorov-Smirnov test for feature selection in emotion recognition from speech. | Alexei V. Ivanov, Giuseppe Riccardi |
| 2011 | EICS | Tell me your needs: assistance for public transport users. | Bernd Ludwig, Martin Hacker, Richard Schaller, Bjrn Zenker, Alexei V. Ivanov, Giuseppe Riccardi |
| 2011 | ICASSP | Simultaneous dialog act segmentation and classification from human-human spoken conversations. | Silvia Quarteroni, Alexei V. Ivanov, Giuseppe Riccardi |
| 2011 | ICASSP | POMDP concept policies and task structures for hybrid dialog management. | Sebastian Varges, Giuseppe Riccardi, Silvia Quarteroni, Alexei V. Ivanov |
| 2011 | Interspeech | Recognition of Personality Traits from Human Spoken Conversations. | Alexei V. Ivanov, Giuseppe Riccardi, Adam J. Sporka, Jakub Franc |
| 2010 | Interspeech | Automatic turn segmentation in spoken conversations. | Alexei V. Ivanov, Giuseppe Riccardi |
| 2010 | Interspeech | Acoustic correlates of meaning structure in conversational speech. | Alexei V. Ivanov, Giuseppe Riccardi, Sucheta Ghosh, Sara Tonelli, Evgeny A. Stepanov |
| 2010 | SIGdial | Investigating Clarification Strategies in a Hybrid POMDP Dialog Manager. | Sebastian Varges, Silvia Quarteroni, Giuseppe Riccardi, Alexei V. Ivanov |
| 2009 | ACL | Combining POMDPs trained with User Simulations and Rule-based Dialogue Management in a Spoken Dialogue System. | Sebastian Varges, Silvia Quarteroni, Giuseppe Riccardi, Alexei V. Ivanov, Pierluigi Roberti |
| 2009 | ASRU | The exploration/exploitation trade-off in Reinforcement Learning for dialogue management. | Sebastian Varges, Giuseppe Riccardi, Silvia Quarteroni, Alexei V. Ivanov |
| 2009 | SIGdial | Leveraging POMDPs Trained with User Simulations and Rule-based Dialogue Management in a Spoken Dialogue System. | Sebastian Varges, Silvia Quarteroni, Giuseppe Riccardi, Alexei V. Ivanov, Pierluigi Roberti |
| 2006 | IJCNN | Markov Coding Strategy of the Simple Spiking Model of Auditory Neuron. | Alexei V. Ivanov, Alexander A. Petrovsky |
| 2005 | Interspeech | Frequency-domain auditory suppression modelling (FASM) - a WDFT-based anthropomorphic noise-robust feature extraction algorithm for speech recognition. | Alexei V. Ivanov, Marek Parfieniuk, Alexander A. Petrovsky |