| 2026 | Transformer-Enabled Diachronic Analysis of Vedic Sanskrit: Neural Methods for Quantifying Types of Language Change. | Ananth Hariharan, David R. Mortensen |
| 2026 | A Comprehensive Full-Form Lexicon for Arabic NLP and Speech Technology. | Yannis Haralambous, Jack Halpern |
| 2026 | Investigating the Role of Synthetic Data Augmentation and Training Strategies on Improving Low-Resource Language ASR. | Yun Hao, Reihaneh Amooie, Wietse de Vries, Rik van Noord, Martijn Wieling |
| 2026 | TCMPHal: A Large-scale Dataset for Hallucination Detection in Traditional Chinese Medicine Pharmacy. | Nijia Han, Zimu Wang, Ziwen Xie, Wei Wang, Jia Meng, John Moraros, Shuihua Wang |
| 2026 | A Japanese Dataset for Aspect-based Sentiment Polarity Classification and Emotion Intensity Estimation. | Kentaro Hanafusa, Kota Manabe, Yuki Maeda, Daisuke Maekawa, Tomoyuki Kajiwara, Hideaki Hayashi, Yuta Nakashima, Hajime Nagahara |
| 2026 | Using Multimodal and Language-Agnostic Sentence Embeddings for Abstractive Summarization. | Chaimae Chellaf El Hammoud, Salima Mdhaffar, Yannick Estve, Stphane Huet |
| 2026 | AraREQ: A Dataset and End-to-End System for Conflict Detection and Resolution in Software Requirements. | Tymaa Hasanain Hammouda, Alaa Aljabari, Nagham Fahim Hamad, Mustafa Jarrar |
| 2026 | Benchmarking Arabic Authorship Attribution and Style Transfer with Large Language Models. | Injy Hamed, Bashar Alhafni, Nizar Habash, Thamar Solorio |
| 2026 | Annotating Conversational Phases and Communication Techniques: A Corpus of German Teacher-Parent Counseling Conversations. | Tobias Hallmen, Kathrin Gietl, Karoline Hillesheim, Annemarie Friedrich, Elisabeth Andr |
| 2026 | What Triggers My Model? Contrastive Explanations Inform Gender Choices by Translation Models. | Jania Hackenbuchner |
| 2026 | Nawatl Context-Free Grammars for Natural Language Processing. | Juan Jos Guzmn-Landa, Juan-Manuel Torres-Moreno, Graham Ranger, Miguel Figueroa-Saavedra, Ligia Quintana-Torres, Carlos-Emiliano Gonzlez-Gallardo, Luis-Gil Moreno-Jimnez, Martha Lorena Avendao-Garrido |
| 2026 | Unsupervised Labelling of Mutation Triggers in Welsh. | Nicols Gutirrez-Roln, Fernando Alva-Manchego |
| 2026 | Proffiliadur: Welsh Language Text Profiling Toolkit. | Nicols Gutirrez-Roln, Jonathan Davies, Tomos Williams, Dawn Knight, Fernando Alva-Manchego |
| 2026 | Ragability Benchmark: A Dataset and Library to Test LLMs on Inter-context Conflicts. | Stephanie Gross, Johann Petrak, Brigitte Krenn |
| 2026 | Automatic Suggestions of Supplements in the Herculaneum Papyri: Language Models and RESTful API. | Angelo Mario Del Grosso, Gabriele Giannessi, Simone Zenzaro, Federico Boschetti |
| 2026 | Prerequisites for Advancing Automatic Speech Recognition in Breton. | Morgan Grobol, Alice Millour, Wassim Zemouri, Yuna Drapier, Mlanie Jouitteau |
| 2026 | Enhancing and Evaluating Tabular Models on the Fly via Synthetic Question-Answer Generation. | Jorge Oss Grijalba, Eugenio Martnez Cmara, Luis Alfonso Ure Lpez, Jos Camacho-Collados |
| 2026 | Trust Me, I Can Convince You: The Contextualized Argument Appraisal Framework and the ContArgA Corpus. | Lynn Greschner, Sabine Weber, Roman Klinger |
| 2026 | Categorical Emotions or Appraisals - Which Emotion Model Explains Argument Convincingness Better? | Lynn Greschner, Meike Bauer, Sabine Weber, Roman Klinger |
| 2026 | Identifying Contexts of Distress in College Students' Reddit Posts: A Comparative Study of Classical NLP and Large Language Models. | Carine Graff, Nikhil Krishnaswamy |
| 2026 | IMaSC: A Malayalam Speech Corpus for High-Quality Text-to-Speech Synthesis. | Deepa P. Gopinath, Thennal D. K, Vrinda V. Nair, Swaraj K. S, Sachin G |
| 2026 | Temporal Expression Recognition in Legal Transcripts. | Elizabeth J. Goldstein, Maria Berger |
| 2026 | Adja-French Parallel Corpus: A New Resource for Machine Translation of a West African Under-Resourced Language. | Josue Frejus Godeme, Rolando Coto-Solano |
| 2026 | Fruitcakes and Cupcakes Emerging from Noise: The ComposiGen Dataset of Compounds and Their Compositionality. | Jule Godbersen, Sinan Cem Kurtyigit, Emma Raimundo Schulz, Tonmoy Rakshit, Diego Frassinelli, Sabine Schulte im Walde, Carina Silberer |
| 2026 | MELD: Melding Diverse Multilingual and Multi-Domain Datasets for Named Entity Recognition Evaluation. | Kevin Glocker, Marco Kuhlmann |