| 2026 | Dynamically Acquiring Text Content to Enable the Classification of Lesser-known Entities for Real-world Tasks. | Fahmida Alam, Ellen Riloff |
| 2026 | Morphemes without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs. | Yara Alakeel, Chatrine Qwaider, Hanan Aldarmaki, Sawsan Alqahtani |
| 2026 | Ramsa: A Large Sociolinguistically Rich Emirati Arabic Speech Corpus for ASR and TTS. | Rania Al-Sabbagh |
| 2026 | NOVELSUM: Evaluating Long-Form Summary Generation for Historical Scandinavian Novels. | Ali Al-Laith, Alexander Conroy, Kirstine Nielsen Degn, Jens Bjerring-Hansen, Daniel Hershcovich |
| 2026 | ADAB: Arabic Dataset for Automated Politeness Benchmarking - a Large-Scale Resource for Computational Sociopragmatics. | Hend Al-Khalifa, Nadia Ghezaiel, Maria Bounnit, Hend Hamed Alhazmi, Noof Abdullah Alfear, Reem Fahad Alqifari, Ameera Masoud Almasoud, Sharefah Ahmed Al-Ghamdi |
| 2026 | Cohesion-6K: An Arabic Dataset for Analyzing Social Cohesion and Conflict in Online Discourse. | Aisha Ali Al-Athba, Wajdi Zaghouani |
| 2026 | FinER-ABSA: A Benchmark for Implicit and Explicit Entity Recognition and Aspect-Based Sentiment Analysis in Financial News. | Pachara Akkanwanich, Pavorn Thongyoo, Mahannop Thabua, Konlakorn Wongpatikaseree, Natthawut Kertkeidkachorn |
| 2026 | Building Effective Japanese Medical LLMs with an Open Recipe for Domain Adaptation through Continued Pre-training. | Akiko Aizawa, Yuki Arase, Fei Cheng, Jiahao Huang, Zhiyi Huang, Junfeng Jiang, Teruhito Kanazawa, Daisuke Kawahara, Kazuma Kobayashi, Takashi Kodama, Sadao Kurohashi, Yusuke Oda, Tsuta Yuma, Zhen Wan, Zhishen Yang, Rio Yokota |
| 2026 | Exploration of How Hate Is Framed on Social Media. | Rakshitha Rao Ailneni, Sanda M. Harabagiu |
| 2026 | The MISOMEM-Val Dataset for Identifying Human Values in Misogynistic Memes. | Rakshitha Rao Ailneni, Sanda M. Harabagiu |
| 2026 | Assessing the Difficulty of Inference Types in Natural Language Inference for Clinical Trials. | Mathilde Aguiar, Pierre Zweigenbaum, Nona Naderi |
| 2026 | Code-Switching in End-to-End Automatic Speech Recognition: A Systematic Literature Review. | Maha Tufail Agro, Atharva Kulkarni, Karima Kadaoui, Zeerak Talat, Hanan Aldarmaki |
| 2026 | Real-Time Generation of Game Video Commentary with Multimodal LLMs: Pause-Aware Decoding Approaches. | Anum Afzal, Yuki Saito, Hiroya Takamura, Katsuhito Sudoh, Shinnosuke Takamichi, Graham Neubig, Florian Matthes, Tatsuya Ishigaki |
| 2026 | Fine-grained Narrative Classification in Biased News Articles. | Zeba Afroz, Harsh Vardhan, Pawan Bhakuni, Aanchal Punia, Rajdeep Kumar, Md. Shad Akhtar |
| 2026 | GeneFRDebate: Generated French Debates from News Articles with Industrial-Expert Summaries. | Rim Abrougui, Guillaume Lechien, Elisabeth Savatier, Benot Laurent |
| 2026 | Learning Long-Document Embeddings via Chunk-Context Entailment. | Waheed Ahmed Abro, Nam Es-Sebbani, Zied Bouraoui |
| 2026 | Improving Neural Argumentative Stance Classification in Controversial Topics with Emotion-Lexicon Features. | Mohammad Yeghaneh Abkenar, Weixing Wang, Manfred Stede, Mark A. Finlayson, Davide Picca, Panagiotis Ioannidis |
| 2026 | Counter-Hypothesis Generation: Towards Evaluating How LLMs Reason about Alternatives. | Marzieh Abdolmaleki, Aaron Maladry, Vronique Hoste, Els Lefever |
| 2026 | GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge. | Ayomide Abayomi, Jos G. Moreno, Karim Radouane, Lynda Tamine |
| 2026 | HiFi-KPI: A Dataset for Hierarchical KPI Extraction from Earnings Filings. | Rasmus T. Aavang, Giovanni Rizzi, Rasmus Tjalk-Bggild, Alexandre Iolov, Mike Zhang, Johannes Bjerva |
| 2022 | Automatic Correction of Syntactic Dependency Annotation Differences. | Andrew Zupon, Andrew Carnie, Michael Hammond, Mihai Surdeanu |
| 2022 | ClinIDMap: Towards a Clinical IDs Mapping for Data Interoperability. | Elena Zotova, Montse Cuadros, German Rigau |
| 2022 | CoFiF Plus: A French Financial Narrative Summarisation Corpus. | Nadhem Zmandar, Tobias Daudert, Sina Ahmadi, Mahmoud El-Haj, Paul Rayson |
| 2022 | A Whole-Person Function Dictionary for the Mobility, Self-Care and Domestic Life Domains: a Seedset Expansion Approach. | Ayah Zirikly, Bart Desmet, Julia Porcino, Jonathan Camacho Maldonado, Pei-Shu Ho, Rafael Jimnez Silva, Maryanne Sacco |
| 2022 | PLOD: An Abbreviation Detection Dataset for Scientific Documents. | Leonardo Zilio, Hadeel Saadany, Prashant Sharma, Diptesh Kanojia, Constantin Orasan |