| 2025 | An Agent-Driven Architecture for Harmful Meme Detection through Multimodal Decomposition. | Gian Marco Orlando, Marco Perillo, Diego Russo, Vincenzo Moscato |
| 2025 | A One-Class Structural Similarity-Based Autoencoder for the Detection of Malaria-Infected Cells. | Moses Omondi, Yassine Belkhouche |
| 2025 | LPConv: Laplacian Pyramid Convolutions for Parameter-Efficient Receptive Field Expansion. | Naoki Nishiya, Akira Kubota |
| 2025 | Quality Assessment of Dynamic 3D Model in Virtual Reality: Effects of Level of Detail and Viewing Distance. | Duc V. Nguyen, Nguyen Thi Quynh Ly, Truong Thu Huong |
| 2025 | Analysis of Multimodal LLMs in VQA in the Field of Radiology. | Cristovo Pessoa Cndido Neto, Cludio de Souza Baptista, Andr Luiz Firmino Alves, Vivek Swarnakar, Anselmo Cardoso de Paiva |
| 2025 | DWT Domain Precinct-Wise Scrambling for Encryption-then-Compression with JPEG XS. | Takayuki Nakachi, Park Cheolhwan, Yasuhisa Kato, Mitsuru Maruyama |
| 2025 | Video Classification of Marchantia Polymorpha Using a Video Vision Transformer with Emphasized Channel Information. | Haruhiko Murata, Naoki Minamino, Takashi Ueda, Yohei Kondo, Kazuhiro Hotta |
| 2025 | Graph-Based Evaluation of Visual Brain Decoding from fMRI Data. | Mohammad Moradi, Morteza Moradi, Marco Grassia, Giuseppe Mangioni |
| 2025 | Evaluation of a Floating-Head Communication Prototype for Video-Conferencing. | William Menz, Alexander Zoubarev, David Kutschke, Rakesh Rao Ramachandra Rao, Louay Bassbouss, Sven Bliedung von der Heide, Steve Gring, Alexander Raake |
| 2025 | Smarter Traps: Neural Network-Driven Classification of Small Mammals. | William Menz, Ralf Dittrich, Rakesh Rao Ramachandra Rao, Steve Gring, Alexander Raake |
| 2025 | Personalised Stress Detection: An Exploration of Temporal Multimodal Late Fusion Strategies. | Misha Libman, Gelareh Mohammadi |
| 2025 | AMICO: A Semantic and Multimodal Framework for AI-Assisted Clinical Reporting. | Antonio Laudante, Mariano Barone, Giuseppe Riccio, Antonio Romano, Francesco Di Serio, Antonio Scialdone, Francesco Porciello, Nicola Rainone, Vincenzo Moscato |
| 2025 | Fewer-Shot Self-Supervised Image Recoloring for Deutan Deficiency Based on Laplacian Pyramid. | Onhi Kato, Akira Kubota |
| 2025 | Rendering Compressed Point Clouds with a Voxel-Based Method. | HyungWoo Kang, YeoJun Yoon, Joong-Hwan Baek, Byung Tae Oh |
| 2025 | RAG Chatbots for Educational Virtual Field Trips. | Suryaprakash Reddy Kalvakolu, Heinrich Sbke, Florian Wehking, Mukesh Chandra Kumar Mamidala, Eckhard Kraft |
| 2025 | One Size Doesn't Fit All: Age-Aware Gamification Mechanics for Multimedia Learning Environments. | Sarah Kaier, Markus Kleffmann, Kristina Schaaff |
| 2025 | Dialogue-Pseudo: A Speaker Pseudonymization Framework for Privacy Protection in Dialogue Speech Data. | Aoi Ito, Katunobu Itou |
| 2025 | Adaptation of CDN at the Edge Using Cloud-Native Network Telemetry Across Media Scenarios. | Javier Iglesias, Juan Felipe Mogolln, Iigo Tamayo, Zaloa Fernndez, Olov Danielsson, Ivan Pretel, Asier Lopez |
| 2025 | Diversity-Aware Active Learning for Object Detection Utilizing Time-of-Day Metadata. | Fumiya Higashide, Akira Kubota |
| 2025 | Recognition of Pitching Habits Using Multimodal Data of RGB Video and Skeleton. | Satoki Hidaka, Kazuhiro Hotta |
| 2025 | TARS: Temporal-Spatial Adaptation for Volumetric Video Streaming. | Hadi Heidarirad, Amir Allahveran, Mea Wang |
| 2025 | Opt360: QoE Optimization for 360° Video Streaming. | Reza Hedayati, Mea Wang, Logan Rakai |
| 2025 | 3GPP PDU Set Framework: Release 19 Updates. | Serhan Gl, Igor D. D. Curcio |
| 2025 | Comparative Analysis of Face Recognition Models: Runtime Environments and Compute Units on Edge. | Lukasz Grzymkowski, Tomasz P. Stefanski |
| 2025 | Exploiting LLMs for Metadata-Based Video Quality Prediction. | Steve Gring, Rakesh Rao Ramachandra Rao, Alexander Raake |