| 2025 | ACL | Anything Goes? A Crosslinguistic Study of (Im)possible Language Learning in LMs. | Xiulin Yang, Tatsuya Aoyama, Yuekun Yao, Ethan Wilcox |
| 2025 | EMNLP | Reason to Rote: Rethinking Memorization in Reasoning. | Yupei Du, Philipp Mondorf, Silvia Casola, Yuekun Yao, Robert Litschko, Barbara Plank |
| 2025 | EMNLP | Language models can learn implicit multi-hop reasoning, but only if they have lots of training data. | Yuekun Yao, Yupei Du, Dawei Zhu, Michael Hahn, Alexander Koller |
| 2024 | EMNLP | Predicting generalization performance with correctness discriminators. | Yuekun Yao, Alexander Koller |
| 2024 | NAACL | Simple and effective data augmentation for compositional generalization. | Yuekun Yao, Alexander Koller |
| 2023 | EMNLP | SLOG: A Structural Generalization Benchmark for Semantic Parsing. | Bingzhi Li, Lucia Donatelli, Alexander Koller, Tal Linzen, Yuekun Yao, Najoung Kim |
| 2022 | EMNLP | Structural generalization is hard for sequence-to-sequence models. | Yuekun Yao, Alexander Koller |
| 2020 | AMTA | Dynamic Masking for Improved Stability in Online Spoken Language Translation. | Yuekun Yao, Barry Haddow |