| 2023 | Masked Feature Modelling for the unsupervised pre-training of a Graph Attention Network block for bottom-up video event recognition. | Dimitrios Daskalakis, Nikolaos Gkalelis, Vasileios Mezaris |
| 2023 | Progressive Coding for Neural Field Transmission. | Anustup Choudhury, Guan-Ming Su |
| 2023 | Semantic and Lexical Token Based Vectors Improve Precision of Recommendations for TV Programmes. | Taner Cagali, Hadi Wazni, Saba Nazir, Mehrnoosh Sadrzadeh, Chris Newell |
| 2023 | Active Learning for Multi-Class Vehicle Categorization and Traffic Analysis in complex environments. | Gabriel Lugo Bustillo, Joey Quinlan, Lingrui Zhou, Md Nahid Sadik, Irene Cheng |
| 2023 | Identification of Visual Objects in Lecture Videos with Color and Keypoints Analysis. | Dipayan Biswas, Shishir K. Shah, Jaspal Subhlok |
| 2023 | Multi-Scale Image Graph Representation: A Novel GNN Approach for Image Classification through Scale Importance Estimation. | Joo Pedro Oliveira Batisteli, Silvio Jamil Ferzoli Guimares, Zenilton K. G. do Patrocnio |
| 2023 | Illuminating the Bias in Pedestrian Detection. | Afnan Althoupety, Li-Yun Wang, Wu-Chi Feng, Banafsheh Rekabdar |
| 2023 | Weakly Labeled Sound Event Detection using Attention Mechanism with Teacher-Student Model. | Yesim Akar, Mustafa Sert |
| 2023 | NN-VVC: Versatile Video Coding boosted by self-supervisedly learned image coding for machines. | Jukka I. Ahonen, Nam Le, Honglei Zhang, Antti Hallapuro, Francesco Cricri, Hamed Rezazadegan Tavakoli, Miska M. Hannuksela, Esa Rahtu |
| 2023 | Spectrogram-Based Deep Learning for Flute Audition Assessment and Intelligent Feedback. | Manu Agarwal, Ross Greer |
| 2023 | Care3D: An Active 3D Object Detection Dataset of Real Robotic-Care Environments. | Michael G. Adam, Sebastian Eger, Martin Piccolrovazzi, Maged Iskandar, Joern Vogel, Alexander Dietrich, Seongjin Bien, Jon Skerlj, Abdeldjallil Naceri, Eckehard G. Steinbach, Alin Albu-Schffer, Sami Haddadin, Wolfram Burgard |
| 2022 | The Lottery Ticket Adaptation for Neural Video Coding. | Nannan Zou, Francesco Cricri, Honglei Zhang, Hamed R. Tavakoli, Miska M. Hannuksela, Esa Rahtu |
| 2022 | Complete Cross-triplet Loss in Label Space for Audio-visual Cross-modal Retrieval. | Donghuo Zeng, Yanan Wang, Jianming Wu, Kazushi Ikeda |
| 2022 | Low-precision post-filtering in video coding. | Ruiying Yang, Mara Santamara, Francesco Cricri, Honglei Zhang, Jani Lainema, Ramin Ghaznavi Youvalari, Miska M. Hannuksela |
| 2022 | Robust Depth Estimation in Foggy Environments Combining RGB Images and mmWave Radar. | Mengchen Xiong, Xiao Xu, Dong Yang, Eckehard G. Steinbach |
| 2022 | Retaining Semantics in Image to Music Conversion. | Zeyu Xiong, Pei-Chun Lin, Amin Farjudian |
| 2022 | Roundwood Tracking from the Forest to the Sawmill using filter approaches to highlight the annual ring pattern. | Georg Wimmer, Rudolf Schraml, Andreas Uhl, Alexander Petutschnigg |
| 2022 | Towards Accurate Positioning in Multiuser Augmented Reality on Mobile Devices. | Na Wang, Haoliang Wang, Stefano Petrangeli, Viswanathan Swaminathan, Fei Li, Songqing Chen |
| 2022 | Semantic-Aware View Prediction for 360-Degree Videos at the 5G Edge. | Shivi Vats, Jounsup Park, Klara Nahrstedt, Michael Zink, Ramesh K. Sitaraman, Hermann Hellwagner |
| 2022 | An artificial neural network-based system for detecting machine failures using a tiny sound dataset: A case study. | Thanh Tran, Sebastian Bader, Jan Lundgren |
| 2022 | Segmentation Consistency Training: Out-of-Distribution Generalization for Medical Image Segmentation. | Birk Torpmann-Hagen, Vajira Thambawita, Michael A. Riegler, Pl Halvorsen, Kyrre Glette |
| 2022 | Optimizing storage and delivery of Omnidirectional Videos in Viewport-dependent streaming. | Kashyap Kammachi Sreedhar, Miska M. Hannuksela, Emre B. Aksu, Lauri Ilola, Lukasz Condrad |
| 2022 | Emotionally Expressive Motion Controller for Virtual Character Locomotion Animations. | Diogo Gonalves Silva, Pedro Alexandre Simes dos Santos, Joo Dias |
| 2022 | Impact of Conventional and Deep Learning-based Point Cloud Geometry Coding on Deep Learning-based Classification Performance. | Abdelrahman Seleem, Andr F. R. Guarda, Nuno M. M. Rodrigues, Fernando Pereira |
| 2022 | Teardrop Magnification: A Hybrid Linear-Fisheye Magnifier for the Border and Corner of the Screen. | Florian Schniederjann, Darius Rausch, Jens Wiggenbrock, Robert Mertens |