| 2025 | Decoding Affective States without Labels: Bimodal Image-brain Supervision. | Vadym Gryshchuk, Maria Maistro, Christina Lioma, Tuukka Ruotsalo |
| 2025 | Multi-Representation Diagrams for Pain Recognition: Integrating Various Electrodermal Activity Signals into a Single Image. | Stefanos Gkikas, Ioannis Kyprakis, Manolis Tsiknakis |
| 2025 | Tiny-BioMoE: a Lightweight Embedding Model for Biosignal Analysis. | Stefanos Gkikas, Ioannis Kyprakis, Manolis Tsiknakis |
| 2025 | Efficient Pain Recognition via Respiration Signals: A Single Cross-Attention Transformer Multi-Window Fusion Pipeline. | Stefanos Gkikas, Ioannis Kyprakis, Manolis Tsiknakis |
| 2025 | MultiGen: Child-Friendly Multilingual Speech Generator with LLMs. | Xiaoxue Gao, Huayun Zhang, Nancy F. Chen |
| 2025 | Simulated Insight, Real-World Impact: Enhancing Driving Safety with CARLA-Simulated Personalized Lessons and Eye-Tracking Risk Coaching. | Wenbin Gan, Minh-Son Dao, Koji Zettsu |
| 2025 | From Lab to Wrist: Bridging Metabolic Monitoring and Consumer Wearables for Heart Rate and Oxygen Consumption Modeling. | Barak Gahtan, Sanketh Vedula, Gil Samuelly Leichtag, Einat Kodesh, Alex M. Bronstein |
| 2025 | Leveraging Pre-Trained Transformers and Facial Embeddings for Multimodal Hirability Prediction in Job Interviews. | Eric Fithian, Theodora Chaspari |
| 2025 | Unobtrusive Universal Acoustic Adversarial Attacks on Speech Foundation Models in the Wild. | Jayden Fassett, Anjila Budathoki, Jack Morris, Qin Hu, Yi Ding |
| 2025 | When Pose Estimation Fails: Measuring Occlusion for Reliable Multimodal Interaction. | Chengyu Fan, Tahiya Chowdhury |
| 2025 | PoseDoc: An Interactive Tool for Efficient Annotation in Human Pose Estimation. | Chengyu Fan, Tahiya Chowdhury |
| 2025 | Multimodal Task Analysis in Wearable Contexts. | Julien Epps |
| 2025 | Investigation into Unimodal Versus Multimodal Pain Recognition from Physiological Signals. | Anis Elebiary, Shaun J. Canavan |
| 2025 | Beyond Utterance: Understanding Group Problem Solving through Discussion Sequences. | Zhuoxu Duan, Zhengye Yang, Brooke Foucault Welles, Richard J. Radke |
| 2025 | From Speech and PPG to EDA: Stress Detection Based on Cross-Modal Fine-Tuning of Foundation Models. | Alia Ahmed Al Dossary, Mathieu Chollet, Alessandro Vinciarelli |
| 2025 | Disentangling Cross-Modal Interactions for Enhanced Multimodal Emotion Recognition in Conversation. | Jian Ding, Bo Zhang, Dailin Li, Jian Wang, Hongfei Lin |
| 2025 | Multimodal Deepfake Generation and Detection: Challenges, Methods, and Future Directions. | Abhinav Dhall, Zhixi Cai, Shreya Ghosh |
| 2025 | Painthenticate: Feature Engineering on Multimodal Physiological Signals. | Sajeeb Datta, Gourab Datta, Tom Gedeon, Md. Zakir Hossain |
| 2025 | Bridging Video and Symbols: A Hybrid AI for Edge Traffic-Risk Reasoning. | Minh-Son Dao, Phuong Thi Mai Nguyen, Swe Nwe Nwe Htun, Koji Zettsu |
| 2025 | A Platform for Experimenting with Non-Verbal Communication: Inserting facial displays of misunderstanding into live conversations. | Ella Cullen, Karl Clarke, Laman Majid, Andrew Voller, Patrick Healey |
| 2025 | A Multilingual, Multimodal Dataset for Disinformation and Out-of-Context Analysis with Rich Supportive Information. | Shuhan Cui, Hanrui Wang, Ching-Chun Chang, Huy H. Nguyen, Isao Echizen |
| 2025 | Decoding social interaction to understand traumatic behaviours in social dynamics. | Pritesh Nalinbhai Contractor |
| 2025 | SBM: Social Behavior Model for Human-Like Action Generation. | Jouh Yeong Chew, Zhi-Yi Lin, Xucong Zhang |
| 2025 | TGN-PL: Learning to Socialize Using Privileged Information and Temporal Graph Networks. | Jouh Yeong Chew, Joanne Taery Kim, Sehoon Ha |
| 2025 | MERD-360VR: A Multimodal Emotional Response Dataset from 360° VR Videos Across Different Age Groups. | Qiang Chen, Shikun Zhou, Yuming Fang, Dan Luo, Tingsong Lu |