| 2026 | ACL | NeedleChain: Measuring Intact Context Comprehension Capability of Large Language Models. | Hyeonseok Moon, Heuiseok Lim |
| 2026 | ACL | Towards Privacy-Preserving Large Language Model: Text-free Inference Through Alignment and Adaptation. | Jeongho Yoon, Chanhee Park, Yongchan Chun, Hyeonseok Moon, Heuiseok Lim |
| 2026 | SIGIR | Beyond Hard Negatives: The Importance of Score Distribution in Knowledge Distillation. | Youngjoon Jang, Seongtae Hong, Hyeonseok Moon, Heuiseok Lim |
| 2025 | ACL | Cross-Lingual Optimization for Language Transfer in Large Language Models. | Jungseob Lee, Seongtae Hong, Hyeonseok Moon, Heuiseok Lim |
| 2025 | ACL | Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer. | Seungyoon Lee, Seongtae Hong, Hyeonseok Moon, Heuiseok Lim |
| 2025 | ACL | Call for Rigor in Reporting Quality of Instruction Tuning Data. | Hyeonseok Moon, Jaehyung Seo, Heuiseok Lim |
| 2025 | COLING | MIGRATE: Cross-Lingual Adaptation of Domain-Specific LLMs through Code-Switching and Embedding Transfer. | Seongtae Hong, Seungyoon Lee, Hyeonseok Moon, Heuiseok Lim |
| 2025 | EMNLP | Metric Calculating Benchmark: Code-Verifiable Complicate Instruction Following Benchmark for Large Language Models. | Hyeonseok Moon, Seongtae Hong, Jaehyung Seo, Heuiseok Lim |
| 2025 | EMNLP | LimaCost: Data Valuation for Instruction Tuning of Large Language Models. | Hyeonseok Moon, Jaehyung Seo, Seonmin Koo, Jinsung Kim, Young-kyoung Ham, Jiwon Moon, Heuiseok Lim |
| 2025 | EMNLP | The Impact of Negated Text on Hallucination with Large Language Models. | Jaehyung Seo, Hyeonseok Moon, Heuiseok Lim |
| 2025 | NAACL | FLEX: A Benchmark for Evaluating Robustness of Fairness in Large Language Models. | Dahyun Jung, Seungyoon Lee, Hyeonseok Moon, Chanjun Park, Heuiseok Lim |
| 2025 | NAACL | Find the Intention of Instruction: Comprehensive Evaluation of Instruction Understanding for Large Language Models. | Hyeonseok Moon, Jaehyung Seo, Seungyoon Lee, Chanjun Park, Heuiseok Lim |
| 2025 | NAACL | MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation. | Chanhee Park, Hyeonseok Moon, Chanjun Park, Heuiseok Lim |
| 2024 | ACL | Length-aware Byte Pair Encoding for Mitigating Over-segmentation in Korean Machine Translation. | Jungseob Lee, Hyeonseok Moon, Seungjun Lee, Chanjun Park, Sugyeong Eo, Hyunwoong Ko, Jaehyung Seo, Seungyoon Lee, Heuiseok Lim |
| 2024 | COLING | Detecting Critical Errors Considering Cross-Cultural Factors in English-Korean Translation. | Sugyeong Eo, Jungwoo Lim, Chanjun Park, Dahyun Jung, Seonmin Koo, Hyeonseok Moon, Jaehyung Seo, Heuiseok Lim |
| 2024 | COLING | Leveraging Pre-existing Resources for Data-Efficient Counter-Narrative Generation in Korean. | Seungyoon Lee, Chanjun Park, Dahyun Jung, Hyeonseok Moon, Jaehyung Seo, Sugyeong Eo, Heuiseok Lim |
| 2024 | EACL | Generative Interpretation: Toward Human-Like Evaluation for Educational Question-Answer Pair Generation. | Hyeonseok Moon, Jaewook Lee, Sugyeong Eo, Chanjun Park, Jaehyung Seo, Heuiseok Lim |
| 2024 | EACL | Hyper-BTS Dataset: Scalability and Enhanced Analysis of Back TranScription (BTS) for ASR Post-Processing. | Chanjun Park, Jaehyung Seo, Seolhwa Lee, Junyoung Son, Hyeonseok Moon, Sugyeong Eo, Chanhee Lee, Heuiseok Lim |
| 2024 | EMNLP | Translation of Multifaceted Data without Re-Training of Machine Translation Systems. | Hyeonseok Moon, Seungyoon Lee, Seongtae Hong, Seungjun Lee, Chanjun Park, Heuiseok Lim |
| 2023 | ACL | Towards Diverse and Effective Question-Answer Pair Generation from Children Storybooks. | Sugyeong Eo, Hyeonseok Moon, Jinsung Kim, Yuna Hur, Jeongwook Kim, Songeun Lee, Changwoo Chun, Sungsoo Park, Heuiseok Lim |
| 2023 | ACL | PEEP-Talk: A Situational Dialogue-based Chatbot for English Education. | Seungjun Lee, Yoonna Jang, Chanjun Park, Jungseob Lee, Jaehyung Seo, Hyeonseok Moon, Sugyeong Eo, Seounghoon Lee, Bernardo Yahya, Heuiseok Lim |
| 2023 | EMNLP | Post-hoc Utterance Refining Method by Entity Mining for Faithful Knowledge Grounded Conversations. | Yoonna Jang, Suhyune Son, Jeongwoo Lee, Junyoung Son, Yuna Hur, Jungwoo Lim, Hyeonseok Moon, Kisu Yang, Heuiseok Lim |
| 2023 | EMNLP | KEBAP: Korean Error Explainable Benchmark Dataset for ASR and Post-processing. | Seonmin Koo, Chanjun Park, Jinsung Kim, Jaehyung Seo, Sugyeong Eo, Hyeonseok Moon, Heuiseok Lim |
| 2023 | EMNLP | CHEF in the Language Kitchen: A Generative Data Augmentation Leveraging Korean Morpheme Ingredients. | Jaehyung Seo, Hyeonseok Moon, Jaewook Lee, Sugyeong Eo, Chanjun Park, Heuiseok Lim |
| 2023 | IJCNLP | Informative Evidence-guided Prompt-based Fine-tuning for English-Korean Critical Error Detection. | Dahyun Jung, Sugyeong Eo, Chanjun Park, Hyeonseok Moon, Jaehyung Seo, Heuiseok Lim |
| 2022 | COLING | QUAK: A Synthetic Quality Estimation Dataset for Korean-English Neural Machine Translation. | Sugyeong Eo, Chanjun Park, Hyeonseok Moon, Jaehyung Seo, Gyeongmin Kim, Jungseob Lee, Heuiseok Lim |
| 2022 | LREC | Empirical Analysis of Noising Scheme based Synthetic Data Generation for Automatic Post-editing. | Hyeonseok Moon, Chanjun Park, Seolhwa Lee, Jaehyung Seo, Jungseob Lee, Sugyeong Eo, Heuiseok Lim |
| 2022 | LREC | Priming Ancient Korean Neural Machine Translation. | Chanjun Park, Seolhwa Lee, Jaehyung Seo, Hyeonseok Moon, Sugyeong Eo, Heuiseok Lim |
| 2022 | NAACL | A Dog Is Passing Over The Jet? A Text-Generation Dataset for Korean Commonsense Reasoning and Evaluation. | Jaehyung Seo, Seounghoon Lee, Chanjun Park, Yoonna Jang, Hyeonseok Moon, Sugyeong Eo, Seonmin Koo, Heuiseok Lim |
| 2021 | NAACL | Should we find another model?: Improving Neural Machine Translation Performance with ONE-Piece Tokenization Method without Model Modification. | Chanjun Park, Sugyeong Eo, Hyeonseok Moon, Heuiseok Lim |