| 2026 | ICPR | ClipTBP: Clip-Pair Based Temporal Boundary Prediction with Boundary-Aware Learning for Moment Retrieval. | Ji-Hyeon Kim, Ho-Joong Kim, Seong-Whan Lee |
| 2025 | ACL | MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval. | Yeong-Joon Ju, Ho-Joong Kim, Seong-Whan Lee |
| 2025 | CVPR | Comprehensive Information Bottleneck for Unveiling Universal Attribution to Interpret Vision Transformers. | Jung-Ho Hong, Ho-Joong Kim, Kyu-Sung Jeon, Seong-Whan Lee |
| 2025 | CVPR | DiGIT: Multi-Dilated Gated Encoder and Central-Adjacent Region Integrated Decoder for Temporal Action Detection Transformer. | Ho-Joong Kim, Yearang Lee, Jung-Ho Hong, Seong-Whan Lee |
| 2025 | SMC | FIQ: Fundamental Question Generation with the Integration of Question Embeddings for Video Question Answering. | Juyoung Oh, Ho-Joong Kim, Seong-Whan Lee |
| 2024 | AAAI | Unknown-Aware Graph Regularization for Robust Semi-supervised Learning from Uncurated Data. | Heejo Kong, Suneung Kim, Ho-Joong Kim, Seong-Whan Lee |
| 2024 | CVPR | TE-TAD: Towards Full End-to-End Temporal Action Detection via Time-Aligned Coordinate Expression. | Ho-Joong Kim, Jung-Ho Hong, Heejo Kong, Seong-Whan Lee |
| 2023 | IJCNN | Enhancing Discriminative Ability among Similar Classes with Guidance of Text-Image Correlation for Unsupervised Domain Adaptation. | Yu-Won Lee, Myeong-Seok Oh, Ho-Joong Kim, Seong-Whan Lee |
| 2022 | AVSS | Temporal-Invariant Video Representation Learning with Dynamic Temporal Resolutions. | Seong-Yun Jeong, Ho-Joong Kim, Myeong-Seok Oh, Gun-Hee Lee, Seong-Whan Lee |