| 2025 | QoE Evaluation of BPP Packet Wash Using ROI-Based Scalable Video Coding. | Mohammadreza Ghafari, Thibault Cholez, Olivier Festor |
| 2025 | Application of Computer Vision Research (ISVP.AI) in the Development of the Comprehensive Stellis One Platform for Sports Organizations. | Lukasz Gasiorowski, Jagoda Lazarek, Sebastian Purtak, Pawel Gra |
| 2025 | An Efficient Optimization Criterion for Multi-View Feature Representation Learning. | Lei Gao, Kai Liu, Kevin Tang, Ling Guan |
| 2025 | Comparison of Multimodal Fall Detection Strategies. | Reema Maheshbhai Gadhia, Nasim Hajari |
| 2025 | From 2D to 3D: How Discrete Dependencies Enable Cross-Dimensional Inference in Neural Networks in Defiance of Euclidean Geometry. | Gerald Friedland, Robert Mertens |
| 2025 | Machine Learning Techniques for the Diagnosis and Monitoring of Nevi and Melanomas. | Giulia Di Flamminio, Fabio Persia, Daniela D'Auria, Ciro Esposito, Vincenzo Coppola |
| 2025 | Extrinsic Calibration of RGB-D Cameras Using Depth Refinement. | Peter O. Fasogbon |
| 2025 | High-Fidelity Semantic Video Communication with Controllable Image-To-Video Diffusion Models. | Cem Eteke, Alexander Griessel, Wolfgang Kellerer, Eckehard G. Steinbach |
| 2025 | Personalized Adaptive Magnification in Gaze-Based Interaction. | Florian Eggenkemper, Jana Swerew, Teresa Rehers, Manuel Hanhoff, Constantin A. Rothkopf, Robert Mertens |
| 2025 | On the Suitability of Perceptual Quality Metrics for Learning-Based Screen Content Compression. | H. Burak Dogaroglu, Hongjie You, Atanas Boev, Elena Alshina, Eckehard G. Steinbach |
| 2025 | Layout-Aware Self-Correcting Prompts for Multimodal LLM Parking Lot Monitoring. | Viviana Crescitelli |
| 2025 | MCAD: Multimodal Context-Aware Audio Description Generation for Soccer. | Lipisha Chaudhary, Trisha Mittal, Subhadra Gopalakrishnan, Ifeoma Nwogu, Jaclyn Pytlarz |
| 2025 | AR in HbbTV-Based Hybrid TV Services. | Fernando Boronat, Lluc Sim, Rubn Prieto, Almanzor Sapena |
| 2025 | Empowering Access to Public Services: An Analysis on Multimodal, Retrieval-Augmented Chatbots for Indic Language Support to Farmers. | Mohsina Bilal, G. Gopakumar |
| 2025 | CCAFF: Object Tracking Under Heavy Occlusion. | Abdul Bhutta, Naimul Khan, Ling Guan |
| 2025 | SynthMed: Generating and Detecting Multimodal Deepfakes for Healthcare Communication. | Mariano Barone, Francesco Di Serio, Vincenzo Moscato, Marco Postiglione, Giuseppe Riccio, Antonio Romano |
| 2025 | Foot-Strike Pattern Recognition from Inertial Data with Machine Learning. | Michele Baldassini, Francesco Pistolesi, Beatrice Lazzerini |
| 2025 | A First Look at Open-GoP Streaming with Av1 S-Frames. | Akram Ansari, S. Ali John Naqvi, Mea Wang, Emir Halepovic |
| 2025 | Coding Gaussian Splat Scenes with V3C/V-PCC. | Patrice Rondao Alface, Lauri Ilola, Lukasz Kondrad |
| 2025 | Forecasting "Neg Storms": Time-Aware Modeling of Toxic Situations in Social Media. | Irien Akter, Vivek K. Singh, Pradeep K. Atrey |
| 2025 | Extending Visual Dialog Beyond English: An Analysis of Monolingual and Multilingual Models. | Milena M. Ado, Silvio Jamil Ferzoli Guimares, Zenilton Kleber G. do Patrocnio Jr. |
| 2024 | Flexible And Faithful Data Insights Generation. | Wei Zhang, Victor Soares Bursztyn |
| 2024 | Visual Speech Recognition with Surrounding and Emotional Information. | Pengcheng Zeng, Atsuo Yoshitaka |
| 2024 | AI Maintenance Techniques by Detecting Performance Degradation in Domain Shift Using Model Ensembles. | Keita Yamane, Akira Kitayama, Keigo Hasegawa, Yusuke Obonai, Hiroto Sasao |
| 2024 | Low-latency Software-based Uncompressed Video Transmission. | Takuro Yamaguchi, Yasuhiro Mochida, Hirokazu Takahashi |