| 2026 | Integrating TEI, NER/NEL, Textometry, and Linked Data for a Semantically Enriched Interview Corpus. | Ranka Stankovic, Tamara Vucenovic, Biljana Rujevic, Milica Ikonic Nesic, Mihailo Skoric |
| 2026 | Glossed Data in Northern Interior Salish. | Anna Stacey |
| 2026 | Knowledge-Infused Hierarchy-Aware Emotion Recognition in Code-mixed Mental Health Counseling Conversations. | Aseem Srivastava, Kushagra Mittal, Anusha Tiwari, Md. Shad Akhtar |
| 2026 | Neural Network-assisted Analysis of Tube Vocal Tract Models. | Runhui Song, Johan Sjons, Axel G. Ekstrm |
| 2026 | Open Korean Historical Corpus: A Millennia-Scale Diachronic Collection of Public Domain Texts. | Seyoung Song, Nawon Kim, Songeun Chae, Kiwoong Park, Jiho Jin, Haneul Yoo, Kyunghyun Cho, Alice Oh |
| 2026 | Scare Quotes as Markers of "Questionable" Word Usages and Misalignment in Conversation: An Annotation Study. | Aina Gar Soler, Juan Carlos Zevallos Huaco, Matthieu Labeau, Chlo Clavel |
| 2026 | Prompting Instruction-tuned LLMs for Semantic Similarity Values. | Xander Akiko Snelder, Yunchong Huang, Jelke Bloem |
| 2026 | Extending Czech Aspect-Based Sentiment Analysis with Opinion Terms: Dataset and LLM Benchmarks. | Jakub Smd, Pavel Pribn, Pavel Krl |
| 2026 | MultiWikiQA: A Reading Comprehension Benchmark in 300+ Languages. | Dan Saattrup Smart |
| 2026 | The Swedish Benchmark of Linguistic Minimal Pairs. | Johan Sjons, Fredrik Heinat, Murathan Kurfali |
| 2026 | Question and Response Dynamics in Public Service Encounters. | Wassiliki Siskou, Ingrid Espinoza, Laurin Friedrich, Steffen Eckhard, Annette Hautli-Janisz |
| 2026 | MUSCAT: MUltilingual, SCientific ConversATion Benchmark. | Supriti Sinhamahapatra, Thai-Binh Nguyen, Yigit Oguz, Enes Yavuz Ugan, Jan Niehues, Alexander Waibel |
| 2026 | AssamLegalTrans: A Parallel Corpus, Benchmark and Analysis for English-Assamese Machine Translation of Legal Judgments. | Telem Joyson Singh, Hemanta Baruah, Sanasam Ranbir Singh, Anindita Talukdar, Nasrin Shahnaz, Okram Jimmy Singh, Priyankoo Sarmah, Pallav Kumar Dutta, Sukumar Nandi, Pranab Duara |
| 2026 | BRAGD: Constrained Multi-Label POS Tagging for Faroese. | Annika Simonsen, Barbara Scalvini, Uni Johannesen, Iben Nyholm Debess, Hafsteinn Einarsson, Vsteinn Snbjarnarson |
| 2026 | Reformulate and Create, Don't Translate: Creating Natural Prompts for Underserved Languages. | Annika Simonsen, Mathias Stenlund, Lars Bungum, Marc Danel Skipsta Volhardt, Hafsteinn Einarsson |
| 2026 | Is Semi-Automatic Transcription Useful in Corpus Creation? Preliminary Considerations on the KIParla Corpus. | Martina Simonotti, Ludovica Pannitto, Eleonora Zucchini, Silvia Ballar, Caterina Mauri |
| 2026 | Benchmarking Portuguese Open Information Extraction. | Gabriel Silva, Mrio Rodrigues, Antnio J. S. Teixeira, Marlene Amorim |
| 2026 | Fables-DTR: A Corpus of Fables Annotated for Discourse and Temporal Relations. | Purificao Silvano, Antnio Leal, Maciej Ogrodniczuk, Aleksandra Tomaszewska, Joana Gomes, Lus Filipe Cunha, Evelin Amorim, Martyna Lewandowska, Anna Sliwicka, Alpio Jorge |
| 2026 | Are Language Models Borrowing-Blind? A Multilingual Evaluation of Loanword Identification across 10 Languages. | Mrilin Sousa Silva, Sina Ahmadi |
| 2026 | LombardoGraphia: Automatic Classification of Lombard Orthography Variants. | Edoardo Signoroni, Pavel Rychl |
| 2026 | Voices across Decades: A Multimodal Diachronic Corpus of German Bundestag Debates (GerParlDia-MM). | Ingo Siegert |
| 2026 | MekongPhon: A Large-Scale Parallel IPA Corpus for Lao and Khmer. | Ammon Shurtz, Christian Richardson, Stephen D. Richardson |
| 2026 | JFC-Recipe: A Dataset for Nutrient Estimation from Japanese User-Generated Cooking Recipes. | Keisuke Shirai, Yoko Yamakata, Hirotaka Kameko, Akiko Sunto, Jun Harashima, Shinsuke Mori |
| 2026 | Simple Additions, Substantial Gains: Expanding Scripts, Languages, and Lineage Coverage in URIEL+. | Mason Shipton, York Hay Ng, Aditya Khan, Phuong Hanh Hoang, Xiang Lu, A. Seza Dogruz, Annie En-Shiun Lee |
| 2026 | Open-access Dataset on Acceptability Ratings of Korean Clausal Constructions by Humans and GPT Models. | Gyu-Ho Shin, Soo-Hwan Lee, Chanyoung Lee |