| 2026 | Multi-dimensional Evaluation of Character-Authentic Dialogue Models Learned from Question-Answer Data. | Atsushi Otsuka, Kazuya Matsuo, Kenta Hama, Masahiro Mizukami, Tsunehiro Arimoto, Hiroaki Sugiyama, Makoto Nakatsuji, Narichika Nomoto |
| 2026 | The Impact of Tokenization Algorithms on Hungarian Language Model Performance. | Mtys Osvth, Mt Norbert Molnr, Roland Gunics, Nomi Ligeti-Nagy |
| 2026 | A Corpus of Joint EEG and Self-Paced Reading of Natural Dutch Texts. | Sara Mller stergaard, Lenneke Doris Lichtenberg, Laura Boon, Bruno Nicenboim |
| 2026 | PETra: A Multilingual Corpus of Pragmatic Explicitation in Translation. | Doreen Osmelak, Koel Dutta Chowdhury, Uliana Sentsova, Cristina Espaa-Bonet, Josef van Genabith |
| 2026 | GPT-NL Public Corpus: A Permissively Licensed, Dutch-First Dataset for LLM Pre-training. | Jesse van Oort, Frank Brinkkemper, Erik de Graaf, Bram Vanroy, Saskia Lensink |
| 2026 | Multimodal LLMs Do Not Compose Skills Optimally across Modalities. | Paula Ontalvilla, Aitor Ormazabal, Gorka Azkune |
| 2026 | Prompt-Based Stance Control in German: An Evaluation of LLMs for Experimental Research on Attitude Change. | Florian Omiecienski, Cornelia Sindermann, Agnieszka Falenska |
| 2026 | MUC-4 Revisited: Document-level Event Analysis beyond Span-based Arguments. | Helene Bsei Olsen, Erik Velldal, Lilja vrelid |
| 2026 | Amulwe Kimn: A Community-Grounded Demo, Resource, and ASR Baseline for Mapuzugun. | Cristian Eduardo Ahumada Oliva, Fatiha Sadat |
| 2026 | JamC-QA: A Multiple-Choice Question Answering Benchmark for Japan-Specific Knowledge. | Teruaki Oka, Tomohide Shibata, Nao Yoshida |
| 2026 | Estonian WinoGrande Dataset: Comparative Analysis of LLM Performance on Human and Machine Translation. | Marii Ojastu, Hele-Andra Kuulmets, Aleksei Dorkin, Marika Borovikova, Dage Srg, Kairit Sirts |
| 2026 | Multi-modal, Multi-task, Multi-criteria Automatic Evaluation with Vision Language Models. | Masanari Oi, Masahiro Kaneko, Naoaki Okazaki, Nakamasa Inoue |
| 2026 | A Single Model Ensemble Framework for Neural Machine Translation Using Pivot Translation. | Seokjin Oh, Keonwoong Noh, Woohwan Jung |
| 2026 | Cross-Dataset Inconsistencies in Morphological Annotation: Evidence from Universal Dependencies. | Vlasta Ohldalov |
| 2026 | Can Video LLMs See Through Illusions? Video-Illusion QA Benchmark Dataset. | Souto Ohira, Tosho Hirasawa, Mamoru Komachi |
| 2026 | Designing LLM Agents for User-Centered Language Service Selection. | Ryoichiro Ogawa, Donghui Lin, Fumito Uwano |
| 2026 | HPLT 3.0: Very Large-Scale Multilingual Resources for LLMs and MT. Mono- and Bi-lingual Data, Multilingual Evaluation, and Pre-Trained Models. | Stephan Oepen, Nikolay Arefyev, Mikko Aulamo, Marta Ban, Maja Buljan, Laurie Burchell, Lucas Georges Gabriel Charpentier, Pinzhen Chen, Mariia Fedorova, Ona de Gibert, Barry Haddow, Jan Hajic, Jindrich Helcl, Andrey Kutuzov, Veronika Laippala, Zihao Li, Bhavitvya Malik, Vladislav Mikhailov, Amanda Myntti, Dayyn O'Brien, Lucie Polkov, Gema Ramrez-Snchez, Janine Siewert, Pavel Stepachev, Jrg Tiedemann, Teemu Vahtola, Dusan Varis, Fedor Vitiugin, Jaume Zaragoza |
| 2026 | Dynamic Layer Selection for Efficient Tone Recognition in Self-Supervised Speech Models. | Saint Germes B. Bengono Obiang, Norbert Tsopz, Paulin Melatagia Yonta |
| 2026 | Preserving Endangered Linguistic Heritage: Developing a Corpus for the Study of Contact-induced Changes in Corfioto. | Giorgio Maria Di Nunzio, Georgios Vardakis |
| 2026 | Sentiment Analysis of German Sign Language Fairy Tales. | Fabrizio Nunnari, Siddhant Jain, Patrick Gebhard |
| 2026 | Investigating How LLMs Propagate Female Stereotypes: Comparing What Models Say via Prompts with What They Represent in Their Embeddings. | Andrea Valderrey Nuez, Jelke Bloem |
| 2026 | From Bones to Rocks: A Systematic Evaluation of Specialized Definition Generation for Portuguese. | Rafael Oleques Nunes, Dennis Giovani Balreira, Joel Lus Carbonera |
| 2026 | A Novel Synthetic Dataset for Few-Shot Legal Relation Extraction in German. | Shiva Banasaz Nouri, Elena Leitner, Julin Moreno Schneider, Georg Rehm |
| 2026 | Push and Pull: Training Sentence Encoders with Contrastive Losses for Distance-Based Multi-Label Text Classification. | Jens Van Nooten, Andriy Kosar |
| 2026 | AYN: A Tiny Yet Competitive Indian Legal Language Model Pretrained from Scratch. | Mitodru Niyogi, ric Gaussier, Arnab Bhattacharya |