| 2026 | English to Central Kurdish Speech Translation: Corpus Creation, Evaluation, and Orthographic Standardization. | Mohammad MohammadAmini, Daban Q. Jaff, Josep Crego, Marie Tahon, Antoine Laurent |
| 2026 | From Trial by Fire to Sleep like a Baby: A Lexicon of Anxiety Associations for 20K English Multi-Word Expressions. | Saif M. Mohammad |
| 2026 | Are the LLMs Capable of Maintaining at Least the Language Genus? | Sandra Mitrovic, David Kletz, Ljiljana Dolamic, Fabio Rinaldi |
| 2026 | AMR Parsing beyond English: An Experiment on Bulgarian, French, Hungarian and Ukrainian. | Ivaylo Mitov, Tadzhat Marharian, Zsofia F. Hauk, Samba Fall, Maxime Amblard, Bruno Guillaume |
| 2026 | The Romanian Corpus Annotated with Multiword Expressions. PARSEME-Ro Version 2.0. | Verginica Barbu Mititelu, Mihaela Cristescu, Elena Irimia, Carmen Mrzea Vasile |
| 2026 | DReUD: Discourse Relations in Universal Dependencies. | Jir Mrovsk, Pavlna Synkov |
| 2026 | Presenting the Prague Discourse Treebank 4.0. | Jir Mrovsk, Pavlna Synkov |
| 2026 | SouDeC: Source Detection and Classification in Czech. | Jir Mrovsk, Barbora Hladk |
| 2026 | Introducing PerMet 1.0: A Metaphor-Annotated Corpus for Persian. | Mohammad Saeid Miri |
| 2026 | Multilingual Target-Stance Extraction. | Ethan Leigh Mines, Bonnie J. Dorr |
| 2026 | A Dataset of Psychiatric Hospital Notes with Temporal Information Annotations. | Timothy A. Miller, Gaby Dinh, David Harris, Wonjin Yoon, Spencer Thomas, Boyu Ren, Mei-Hua Hall, Guergana Savova |
| 2026 | Best-Worst Scaling of Hype in Biomedical Research: Building an Intensity Lexicon of Promotional Adjectives. | Neil Millar, Dipesh Satav, Bojan Batalo, Erica K. Shimomoto, Ryosuke L. Ohniwa |
| 2026 | What Are LLMs Doing to Scientific Communication? Measuring Changes in Writing Practices and Reading Experience. | Filip Miletic, Neele Falk |
| 2026 | Meet UD_Czech-PDTC: A Large and Genre-Rich Treebank in Universal Dependencies. | Marie Mikulov, Barbora Stepnkov, Daniel Zeman, Jan Stepnek, Milan Straka, Jan Hajic |
| 2026 | Prague Dependency Treebank - Consolidated 2.0: Enriching a Complex Annotation Scheme. | Marie Mikulov, Jir Mrovsk, Milan Straka, Pavlna Synkov, Jan Stepnek, Barbora Stepnkov, Jan Hajic |
| 2026 | A Recipe for Adapting Multilingual Embedders to OCR-Error Robustness and Historical Texts. | Andrianos Michail, Stylianos Psychias, Juri Opitz, Simon Clematide |
| 2026 | GerVLPro: A CEFR-Graded Vocabulary List of L2 Learners' Productive Vocabulary in German. | Noah-Manuel Michael, Anna Hlsing, Andrea Horbach |
| 2026 | Breaking the Benchmark: Revealing LLM Bias via Minimal Contextual Augmentation. | Kaveh Eskandari Miandoab, Mahammed Kamruzzaman, Arshia Gharooni, Gene Louis Kim, Vasanth Sarathy, Ninareh Mehrabi |
| 2026 | SALOMO: An Annotation Tool for Complex Annotation Tasks with a Large Number of Labels. | Tim Menzner |
| 2026 | Towards a Diagnostic and Predictive Evaluation Methodology for Sequence Labeling Tasks. | Elena lvarez Mellado, Julio Gonzalo |
| 2026 | Automating FAIRness: A FAIRification Tool within the Language Resources Infrastructure. | Daniele Melaccio, Monica Monachini |
| 2026 | Efficient Topic Extraction via Graph-Based Labeling: A Lightweight Alternative to Deep Models. | Salma Mekaoui, Hiba Sofyan, Imane Benchrif, Imane Amaaz, Ilham Chaker, Arsalane Zarghili, Nikola S. Nikolov |
| 2026 | Can LLMs Faithfully Explain Themselves in Low-Resource Languages? A Case Study on Emotion Detection in Persian. | Mobina Mehrazar, Mohammad Amin Yousefi, Parisa Abolfath Beygi, Behnam Bahrak |
| 2026 | Transcription Accuracy in the Icelandic Gigaword Corpus: Evaluating Automatic and Manual Annotation. | Johanna Mechler, Lilja Bjrk Stefnsdttir, Anton Karl Ingason |
| 2026 | Synthetic Function Demonstrations Improve Generation in Low-Resource Programming Languages. | Nick McKenna, Xinnuo Xu, Jack Williams, Nicholas C. Wilson, Benjamin Van Durme, Christian Plitz |