| 2026 | Contextualizing Toxicity: An Annotation Framework for Unveiling Pragmatics in Conversations of Online Discussion Forums. | Yingxue Fu, Anas Ollagnier |
| 2026 | "Emphasizing the Commendable": A Study of Homogenized Transitive Verb Constructions in Machine Generated Peer Reviews. | Hing-Yuet Fung, Chi-kiu Lo, Samuel Larkin |
| 2026 | LegalRikai: Open Benchmark - a Benchmark for Complex Japanese Corporate Legal Tasks. | Shogo Fujita, Yuji Naraki, Yiqing Zhu, Shinsuke Mori |
| 2026 | Every Word Presented in Context: Syntactic Coverage as Objective for Low-Resource Machine Translation with Large Language Models. | Samuel Frontull, Thomas Strhle |
| 2026 | SciLaD: A Large-Scale, Transparent, Reproducible Dataset for Natural Scientific Language Processing. | Luca Foppiano, Sotaro Takeshita, Pedro Ortiz Suarez, Ekaterina Borisova, Raia Abu Ahmad, Malte Ostendorff, Fabio Barth, Julin Moreno Schneider, Georg Rehm |
| 2026 | Systematic Multi-Aspect Evaluation of Time Series-Based Report Generation: The Case of Financial Analysis from Stock Data. | Elizabeth Fons, Elena Kochkina, Rachneet Kaur, Zhen Zeng, Berowne Hlavaty, Charese Smiley, Svitlana Vyetrenko, Manuela Veloso |
| 2026 | CoSt-BR: A Language Resource for Conversational Stance Detection. | Felipe Penhorate Carvalho da Fonseca, Ivandr Paraboni, Luciano Antnio Digiampietri |
| 2026 | The GELATO Dataset for Legislative NER. | Matthew Flynn, Timothy Obiso, Sam Newman |
| 2026 | Human-in-the-Loop Mass Transcription and Ground Truth Annotation for Challenging Historical Documents. | Norbert Fischer, Frank Puppe |
| 2026 | Procrustes Analysis for Improving Language Model Merging. | Olivier Ferret |
| 2026 | Small LLMs for Medical NLP: A Systematic Analysis of Few-Shot, Constraint Decoding, Fine-Tuning and Continual Pre-Training in Italian. | Pietro Ferrazzi, Mattia Franzin, Alberto Lavelli, Bernardo Magnini |
| 2026 | Deep Learning-Based Multi-Aspect Pronunciation Assessment for Individuals with Down Syndrome. | David Fernndez-Garca, Csar Gonzlez Ferreras, Valentn Cardeoso-Payo, Mario Corrales-Astorgano |
| 2026 | Widespread Gender and Pronoun Bias in Moral Judgments across LLMs. | Gustavo Lcius Fernandes, Jeiverson C. V. M. Santos, Pedro O. S. Vaz-de-Melo |
| 2026 | Referenceless Evaluation of Machine Translation Models by Ranking Performance in Romanian to English Translate-train Settings. | Mihail Feraru, Alexandra Diaconu, Alexe Dumitru-Bogdan |
| 2026 | A Typologically Grounded Evaluation Framework for Word Order and Morphology Sensitivity in Multilingual Masked LMs. | Anna Feldman, Libby Barak, Jing Peng |
| 2026 | Zero-Shot to Full-Resource: Cross-lingual Transfer Strategies for Aspect-Based Sentiment Analysis. | Jakob Fehle, Nils Constantin Hellwig, Udo Kruschwitz, Christian Wolff |
| 2026 | Leveraging Linguistic Similarity for Low-Resource Speech Transcription. | Valentina Fedchenko, Eric Jordan |
| 2026 | PBBQ: A Persian Bias Benchmark Dataset Curated with Human-AI Collaboration for Large Language Models. | Farhan Farsi, Shayan Bali, Fatemeh Valeh, Parsa Ghofrani, Alireza Pakniat, Seyedkian Kashfipour, Amir H. Payberah |
| 2026 | PREMOVE in LiLa: Integrating Latin Preverbed Motion Verbs with WordNet and VerbNet. | Andrea Farina, Marco Passarotti, Francesco Mambrini, Matteo Pellegrini, Eleonora Litta, Giovanni Moretti |
| 2026 | MedPT: A Massive Medical Question Answering Dataset for Brazilian-Portuguese Speakers. | Fernanda Bufon Frber, Iago Alves Brito, Julia Soares Dollis, Pedro Schindler Freire Brasil Ribeiro, Rafael Teixeira Sousa, Arlindo R. Galvo Filho |
| 2026 | Semantic Capacity in Language Learners and LLMs: A Case Study of Quantifier Scope. | Shaohua Fang, Yue Li, Yan Cong |
| 2026 | YoNER: A New Yorb Multi-domain Named Entity Recognition Dataset. | Peace Busola Falola, Jesujoba Alabi, Solomon O. Akinola, Folashade T. Ogunajo, Emmanuel Oluwadunsin Alabi, David Ifeoluwa Adelani |
| 2026 | Building a Dataset for French Accent Classification Evaluation: Are We There Yet? | Diandra Fabre, Mathieu Avanzi, Franois Portet |
| 2026 | StoryCCDial: Collecting and Analyzing Human-Human Co-Creation Dialogues for Personalized Creative Support. | Natsumi Ezure, Michimasa Inaba |
| 2026 | LoveHate: Stance Detection and Generation for Multiple Topics in User-generated Comments in Russian and English. | Natalia Evgrafova, Vronique Hoste, Els Lefever |