| 2026 | ACL | AgentCoMa: A Compositional Benchmark Mixing Commonsense and Mathematical Reasoning in Real-World Scenarios. | Lisa Alazraki, Lihu Chen, Ana Brassard, Joe Stacey, Hossein A. Rahmani, Marek Rei |
| 2025 | AAAI | Identifying Query-Relevant Neurons in Large Language Models for Long-Form Texts. | Lihu Chen, Adam Dejl, Francesca Toni |
| 2025 | COLING | Proceedings of the Workshop on Generative AI and Knowledge Graphs (GenAIK). | Genet Asefa Gesese, Harald Sack, Heiko Paulheim, Albert Meroo-Peuela, Lihu Chen |
| 2025 | EMNLP | Evaluating Uncertainty Quantification Methods in Argumentative Large Language Models. | Kevin Zhou, Adam Dejl, Gabriel Freedman, Lihu Chen, Antonio Rago, Francesca Toni |
| 2024 | EACL | Learning High-Quality and General-Purpose Phrase Representations. | Lihu Chen, Gal Varoquaux, Fabian M. Suchanek |
| 2024 | EMNLP | Reconfidencing LLMs from the Grouping Loss Perspective. | Lihu Chen, Alexandre Perez-Lebel, Fabian M. Suchanek, Gal Varoquaux |
| 2024 | SIGIR | YAGO 4.5: A Large and Clean Knowledge Base with a Rich Taxonomy. | Fabian M. Suchanek, Mehwish Alam, Thomas Bonald, Lihu Chen, Pierre-Henri Paris, Jules Soria |
| 2023 | EACL | GLADIS: A General and Large Acronym Disambiguation Benchmark. | Lihu Chen, Gal Varoquaux, Fabian M. Suchanek |
| 2023 | EMNLP | The Locality and Symmetry of Positional Encodings. | Lihu Chen, Gal Varoquaux, Fabian M. Suchanek |
| 2022 | ACL | Imputing Out-of-Vocabulary Embeddings with LOVE Makes LanguageModels Robust with Little Cost. | Lihu Chen, Gal Varoquaux, Fabian M. Suchanek |
| 2021 | AAAI | A Lightweight Neural Model for Biomedical Entity Linking. | Lihu Chen, Gal Varoquaux, Fabian M. Suchanek |