| 2025 | ICCV | Zero-Shot Multimodal Compound Expression Recognition Approach Using Off-the-Shelf Large Visual-Language Models. | Elena Ryumina, Maxim Markitantov, Alexandr Axyonov, Dmitry Ryumin, Mikhail Dolgushin, Alexey Karpov |
| 2025 | ICMI | A Multilingual Telegram Chatbot for Mental Health Data Collection. | Danila Mamontov, Alexey Karpov, Wolfgang Minker |
| 2025 | Interspeech | Multi-Modal Multi-Task Affective States Recognition Based on Label Encoder Fusion. | Maxim Markitantov, Elena Ryumina, Heysem Kaya, Alexey Karpov |
| 2024 | CVPR | Multi-modal Arousal and Valence Estimation under Noisy Conditions. | Denis Dresvyanskiy, Maxim Markitantov, Jiawei Yu, Heysem Kaya, Alexey Karpov |
| 2024 | CVPR | Zero-Shot Audio-Visual Compound Expression Recognition Method based on Emotion Probability Fusion. | Elena Ryumina, Maxim Markitantov, Dmitry Ryumin, Heysem Kaya, Alexey Karpov |
| 2024 | ICASSP | Audio-Visual Speech Recognition In-The-Wild: Multi-Angle Vehicle Cabin Corpus and Attention-Based Method. | Alexandr Axyonov, Dmitry Ryumin, Denis Ivanko, Alexey M. Kashevnik, Alexey Karpov |
| 2024 | Interspeech | OCEAN-AI: open multimodal framework for personality traits assessment and HR-processes automatization. | Elena Ryumina, Dmitry Ryumin, Alexey Karpov |
| 2023 | Interspeech | Multimodal Personality Traits Assessment (MuPTA) Corpus: The Impact of Spontaneous and Read Speech. | Elena Ryumina, Dmitry Ryumin, Maxim Markitantov, Heysem Kaya, Alexey Karpov |
| 2022 | ICMI | MIDriveSafely: Multimodal Interaction for Drive Safely. | Denis Ivanko, Alexey M. Kashevnik, Dmitry Ryumin, Andrey Kitenko, Alexandr Axyonov, Igor Lashkov, Alexey Karpov |
| 2022 | Interspeech | DAVIS: Driver's Audio-Visual Speech recognition. | Denis Ivanko, Dmitry Ryumin, Alexey M. Kashevnik, Alexandr Axyonov, Andrey Kitenko, Igor Lashkov, Alexey Karpov |
| 2022 | Interspeech | Biometric Russian Audio-Visual Extended MASKS (BRAVE-MASKS) Corpus: Multimodal Mask Type Recognition Task. | Maxim Markitantov, Elena Ryumina, Dmitry Ryumin, Alexey Karpov |
| 2022 | Interspeech | Complex Paralinguistic Analysis of Speech: Predicting Gender, Emotions and Deception in a Hierarchical Framework. | Alena Velichko, Maxim Markitantov, Heysem Kaya, Alexey Karpov |
| 2022 | LREC | RUSAVIC Corpus: Russian Audio-Visual Speech in Cars. | Denis Ivanko, Alexandr Axyonov, Dmitry Ryumin, Alexey M. Kashevnik, Alexey Karpov |
| 2021 | Interspeech | Annotation Confidence vs. Training Sample Size: Trade-Off Solution for Partially-Continuous Categorical Emotion Recognition. | Elena Ryumina, Oxana Verkholyak, Alexey Karpov |
| 2021 | Interspeech | Ensemble-Within-Ensemble Classification for Escalation Prediction from Speech. | Oxana Verkholyak, Denis Dresvyanskiy, Anastasia Dvoynikova, Denis Kotov, Elena Ryumina, Alena Velichko, Danila Mamontov, Wolfgang Minker, Alexey Karpov |
| 2020 | ICMI | Combining Clustering and Functionals based Acoustic Feature Representations for Classification of Baby Sounds. | Heysem Kaya, Oxana Verkholyak, Maxim Markitantov, Alexey Karpov |
| 2020 | Interspeech | Ensembling End-to-End Deep Models for Computational Paralinguistics Tasks: ComParE 2020 Mask and Breathing Sub-Challenges. | Maxim Markitantov, Denis Dresvyanskiy, Danila Mamontov, Heysem Kaya, Wolfgang Minker, Alexey Karpov |
| 2020 | Interspeech | Is Everything Fine, Grandma? Acoustic and Linguistic Modeling for Robust Elderly Speech Emotion Recognition. | Gizem Sogancioglu, Oxana Verkholyak, Heysem Kaya, Dmitrii Fedotov, Tobias Cade, Albert Ali Salah, Alexey Karpov |
| 2020 | LREC | TheRuSLan: Database of Russian Sign Language. | Ildar Kagirov, Denis Ivanko, Dmitry Ryumin, Alexander A. Petrovsky, Alexey Karpov |
| 2020 | LREC | Class-based LSTM Russian Language Model with Linguistic Information. | Irina S. Kipyatkova, Alexey Karpov |
| 2019 | ICASSP | Hierarchical Two-level Modelling of Emotional States in Spoken Dialog Systems. | Oxana Verkholyak, Dmitrii Fedotov, Heysem Kaya, Yang Zhang, Alexey Karpov |
| 2019 | IDC | Lower Limbs Exoskeleton Control System Based on Intelligent Human-Machine Interface. | Ildar Kagirov, Alexey Karpov, Irina S. Kipyatkova, Konstantin Klyuzhev, Alexander Kudryavcev, Igor Kudryavcev, Dmitry Ryumin |
| 2019 | IDC | Applying Ensemble Learning Techniques and Neural Networks to Deceptive and Truthful Information Detection Task in the Flow of Speech. | Alena Velichko, Viktor Budkov, Ildar Kagirov, Alexey Karpov |
| 2019 | PERCOM | Human-Robot Interaction with Smart Shopping Trolley Using Sign Language: Data Collection. | Dmitry Ryumin, Denis Ivanko, Alexandr Axyonov, Ildar Kagirov, Alexey Karpov, Milos Zelezn |
| 2019 | SIGdial | Cross-Corpus Data Augmentation for Acoustic Addressee Detection. | Oleg Akhtiamov, Ingo Siegert, Alexey Karpov, Wolfgang Minker |
| 2018 | Interspeech | LSTM Based Cross-corpus and Cross-task Acoustic Emotion Recognition. | Heysem Kaya, Dmitrii Fedotov, Ali Yesilkanat, Oxana Verkholyak, Yang Zhang, Alexey Karpov |
| 2016 | HCI | Bimodal Speech Recognition Fusing Audio-Visual Modalities. | Alexey Karpov, Alexander L. Ronzhin, Irina S. Kipyatkova, Andrey Ronzhin, Vasilisa Verkhodanova, Anton I. Saveliev, Milos Zelezn |
| 2016 | HCI | Multimodal Information Coding System for Wearable Devices of Advanced Uniform. | Andrey L. Ronzhin, Oleg Basov, Anna I. Motienko, Alexey Karpov, Yuri V. Mikhailov, Milos Zelezn |
| 2016 | ISNN | Language Models with RNNs for Rescoring Hypotheses of Russian ASR. | Irina S. Kipyatkova, Alexey Karpov |
| 2015 | HCI | Automatic Analysis of Speech and Acoustic Events for Ambient Assisted Living. | Alexey Karpov, Alexander L. Ronzhin, Irina S. Kipyatkova |
| 2015 | Interspeech | Fisher vectors with cascaded normalization for paralinguistic analysis. | Heysem Kaya, Alexey Karpov, Albert Ali Salah |
| 2014 | HCI | A Universal Assistive Technology with Multimodal Input and Multimedia Output Interfaces. | Alexey Karpov, Andrey Ronzhin |
| 2014 | Interspeech | Audio-visual signal processing in a multimodal assisted living environment. | Alexey Karpov, Lale Akarun, Hlya Yalin, Alexander L. Ronzhin, Baris Evrim Demirz, Aysun oban, Milos Zelezn |
| 2013 | HCI | Multimodal Synthesizer for Russian and Czech Sign Languages and Audio-Visual Speech. | Alexey Karpov, Zdenek Krnoul, Milos Zelezn, Andrey Ronzhin |
| 2012 | FedCSIS | Analysis of Long-distance Word Dependencies and Pronunciation Variability at Conversational Russian Speech Recognition. | Irina S. Kipyatkova, Alexey Karpov, Vasilisa Verkhodanova, Milos Zelezn |
| 2011 | HCI | An Assistive Bi-modal User Interface Integrating Multi-channel Speech Recognition and Computer Vision. | Alexey Karpov, Andrey Ronzhin, Irina S. Kipyatkova |
| 2011 | Interspeech | Very Large Vocabulary ASR for Spoken Russian with Syntactic and Morphemic Analysis. | Alexey Karpov, Irina S. Kipyatkova, Andrey Ronzhin |
| 2010 | ICPR | Multimodal Human Computer Interaction with MIDAS Intelligent Infokiosk. | Alexey Karpov, Andrey Ronzhin, Irina S. Kipyatkova, Alexander L. Ronzhin, Lale Akarun |
| 2010 | Interspeech | Viseme-dependent weight optimization for CHMM-based audio-visual speech recognition. | Alexey Karpov, Andrey Ronzhin, Konstantin Markov, Milos Zelezn |
| 2009 | HCI | Designing Cognition-Centric Smart Room Predicting Inhabitant Activities. | Andrey Ronzhin, Alexey Karpov, Irina S. Kipyatkova |
| 2009 | Interspeech | Audio-visual speech asynchrony modeling in a talking head. | Alexey Karpov, Liliya Tsirulnik, Zdenek Krnoul, Andrey Ronzhin, Boris Lobanov, Milos Zelezn |
| 2006 | Interspeech | Multi-modal system ICANDO: intellectual computer assistant for disabled operators. | Alexey Karpov, Andrey Ronzhin, Alexandre Cadiou |