| 2024 | Real-time Multi-modal Highlight Prediction for Simultaneous Viewing of Multiple Live Streams. | Yusuke Maeda, Takahiro Hayashi |
| 2024 | A Power-Law Transformation Approach for Template-Based Cross-Component Prediction. | Zhikai Liu, Kun Zhang, Xin-Yi Cui, Wei Sun, Fan Liang |
| 2024 | FrameCorr: Adaptive, Autoencoder-based Neural Compression for Video Reconstruction in Resource and Timing Constrained Network Settings. | John Li, Deepak Nair, Klara Nahrstedt, Indranil Gupta, Shehab Sarar Ahmed |
| 2024 | Investigation of Feature Distribution and Network Weight Updates in the Machine Unlearning Process. | Wen-Hung Liao, Yang-Jing Lin |
| 2024 | Unveiling the Potential of SSL-Generated Audio Embeddings for Cross-Lingual Speaker Recognition. | Wen-Hung Liao, Po-Han Chen, Yi-Chieh Wu |
| 2024 | Watch your back! Dynamic thumbnails for a 360-degree video player to enhance viewing experience on 2D displays. | Jakub Kovc, Wolfgang Hrst |
| 2024 | Generating Bass Phrases from Guitar Chord Backing with NMF. | Tomoo Kouzai, Junya Koguchi, Tetsuro Kitahara |
| 2024 | Prevention of Unexpected Object Generation in Diffusion Model-Based Inpainting. | Takumi Komori, Takahiro Hayashi |
| 2024 | SpotiView: Partial Face Display Method for Smooth Communication While Protecting Privacy. | Ryota Kishimoto, Shuhei Tsuchida, Tsutomu Terada, Masahiko Tsukamoto |
| 2024 | Evaluation Framework for Novel View Synthesis. | Kolja Kieslich, Louay Bassbouss, Stephan Steglich, Stefan Arbanowski |
| 2024 | Investigating the Impact of High Frame Rate on Video Quality: A SAMVIQ Approach. | Dominik Keller, Paul Rudi Frank, Steve Gring, Alexander Raake |
| 2024 | Disparity Correction Method of the Monocular Omnidirectional Stereo Camera. | Hisayoshi Kaneda, Ryota Kawamata, Kazuyoshi Yamazaki, Kazuya Shimizu |
| 2024 | Speaker Pseudonymization for Japanese Speech Using Duration Embeddings. | Aoi Ito, Katunobu Itou |
| 2024 | Two-stage instrument timbre transfer method using RAVE. | Di Hu, Katunobu Ito |
| 2024 | Appeal prediction for AI up-scaled Images. | Steve Gring, Rasmus Merten, Alexander Raake |
| 2024 | S2MGen: A synthetic skin mask generator for improving segmentation. | Subhadra Gopalakrishnan, Trisha Mittal, Jaclyn Pytlarz, Yuheng Zhao |
| 2024 | SoccerNet-Echoes: A Soccer Game Audio Commentary Dataset. | Sushant Gautam, Mehdi Houshmand Sarkhoosh, Jan Held, Cise Midoglu, Anthony Cioppa, Silvio Giancola, Vajira Thambawita, Michael A. Riegler, Pl Halvorsen, Mubarak Shah |
| 2024 | Evaluating Interactive Concept Maps Produced from E-Portfolios. | Alexander Gantikow, Andreas Isking, Wolfgang Mller, Paul Libbrecht, Sandra Rebholz |
| 2024 | LiveSkeleton: High-Quality Real-Time Human Tracking and Pose Estimation. | Hannes Fassold |
| 2024 | Ensuring Color Consistency in RGB-D Multi-Camera Setup. | Peter O. Fasogbon |
| 2024 | Multi-View Gesture Recognition in Conflict Situations. | Karam Tomotaki-Dawoud, Birgit Nierula, Farelle Toumaleu Siewe, Thomas Koch, Daniel Johannes Meyer, Andreas Bock, Marianne Heinze, Daniela Knuth, Denis Martin, Julia Schander, Anna Hilsmann, Peter Eisert, Sebastian Bosse |
| 2024 | Fusion-Based Human Pose Estimation Using RGB and IR Images with Transformer-Based Decoding. | Viviana Crescitelli, Takashi Oshima |
| 2024 | Perceptual Quality Driven Point Cloud Compression for 6DoF 3D Point Cloud Streaming. | Yumeka Chujo, Yusuke Tagashira, Yukiko Harada, Kenji Kanai, Jiro Katto |
| 2024 | Characterizing students behavior in multi-user multi-computer testing environments. | Rajini Chittimalla, Sujung Choi, Madhu Sai Vineel Reka, Yassine Belkhouche |
| 2024 | Occlusion-Aware Real-Time Tiny Facial Alignment Model for Makeup Virtual Try-On. | Kin Ching Lydia Chau, Zhi Yu, Ruowei Jiang |