| 2025 | ACL | Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates. | Jaewoo Ahn, Heeseung Yun, Dayoon Ko, Gunhee Kim |
| 2025 | EMNLP | FlashAdventure: A Benchmark for GUI Agents Solving Full Story Arcs in Diverse Adventure Games. | Jaewoo Ahn, Junseo Kim, Heeseung Yun, Jaehyeon Son, Dongmin Park, Jaewoong Cho, Gunhee Kim |
| 2025 | ICCV | ChartCap: Mitigating Hallucination of Dense Chart Captioning. | Junyoung Lim, Jaewoo Ahn, Gunhee Kim |
| 2025 | NAACL | Is a Peeled Apple Still Red? Evaluating LLMs' Ability for Conceptual Combination with Property Type. | Seokwon Song, Taehyun Lee, Jaewoo Ahn, Jae Hyuk Sung, Gunhee Kim |
| 2024 | ACL | TimeChara: Evaluating Point-in-Time Character Hallucination of Role-Playing Large Language Models. | Jaewoo Ahn, Taehyun Lee, Junyoung Lim, Jin-Hwa Kim, Sangdoo Yun, Hwaran Lee, Gunhee Kim |
| 2024 | ACL | Who Wrote this Code? Watermarking for Code Generation. | Taehyun Lee, Seokhee Hong, Jaewoo Ahn, Ilgee Hong, Hwaran Lee, Sangdoo Yun, Jamin Shin, Gunhee Kim |
| 2023 | ACL | MPCHAT: Towards Multimodal Persona-Grounded Conversation. | Jaewoo Ahn, Yeda Song, Sangdoo Yun, Gunhee Kim |
| 2023 | EMNLP | mRedditSum: A Multimodal Abstractive Summarization Dataset of Reddit Threads with Images. | Keighley Overbay, Jaewoo Ahn, Fatemeh Pesaran Zadeh, Joonsuk Park, Gunhee Kim |
| 2020 | ICLR | Sequential Latent Knowledge Selection for Knowledge-Grounded Dialogue. | Byeongchang Kim, Jaewoo Ahn, Gunhee Kim |
| 2006 | ISVC | Untitled record | Jaewoo Ahn, Kyungha Min |
| 2005 | ICCSA | General-Purpose Text Entry Rules for Devices with 4x3 Configurations of Buttons. | Jaewoo Ahn, Myung Ho Kim |
| 2003 | ICCSA | Approximating 3D General Sweep Boundary Using Depth-Buffer. | Jaewoo Ahn, Sung Je Hong |