| 2026 | Chronocept: Instilling a Sense of Time in Machines. | Krish Goel, Sanskar Pandey, KS Mahadevan, Harsh Kumar, Vishesh Khadaria |
| 2026 | Exploring the Semantic Space of Second Language Learners. | Trisha Godara, Rui He, Wolfram Hinzen, Yan Cong |
| 2026 | VIGiA: Instructional Video Guidance via Dialogue Reasoning and Retrieval. | Diogo Glria-Silva, David Semedo, Joo Magalhes |
| 2026 | Foundations of LLM Knowledge Materialization: Termination, Reproducibility, Robustness. | Luca Giordano, Simon Razniewski |
| 2026 | LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation. | Himel Ghosh, Nick Elias Werner |
| 2026 | A Computational Approach to Visual Metonymy. | Saptarshi Ghosh, Linfeng Liu, Tianyu Jiang |
| 2026 | When Speed Meets Intelligence: Scalable Conversational NER in an Ever-evolving World. | Karim Ghonim, Antonio Roberto, Davide Bernardi |
| 2026 | MEENA (PersianMMMU): Multimodal-Multilingual Educational Exams for N-level Assessment. | Omid Ghahroodi, Arshia Hemmat, Marzia Nouri, Seyed Mohammad Hadi Hosseini, Doratossadat Dastgheib, Mohammad V. Sanian, Alireza Sahebi, Reihaneh Zohrabi, Mohammad Hossein Rohban, Ehsaneddin Asgari, Mahdieh Soleymani Baghshah |
| 2026 | Sycophancy Hides Linearly in the Attention Heads. | Rifo Ahmad Genadi, Munachiso Nwadike, Nurdaulet Mukhituly, Tatsuya Hiraoka, Hilal AlQuabeh, Kentaro Inui |
| 2026 | VaseVQA: Multimodal Agent and Benchmark for Ancient Greek Pottery. | Jinchao Ge, Tengfei Cheng, Biao Wu, Zeyu Zhang, Shiya Huang, Judith Bishop, Gillian Shepherd, Meng Fang, Ling Chen, Yang Zhao |
| 2026 | Is Information Density Uniform when Utterances are Grounded on Perception and Discourse? | Matteo Gay, Coleman Haley, Mario Giulianelli, Edoardo M. Ponti |
| 2026 | Training in Step-by-Step Formal Reasoning Improves Pronominal Reasoning in Language Models. | Vagrant Gautam |
| 2026 | Negative-Aware Diffusion Process for Temporal Knowledge Graph Extrapolation. | Yanglei Gan, Peng He, Yuxiang Cai, Run Lin, Guanyu Zhou, Qiao Liu |
| 2026 | Rethinking Hallucinations: Correctness, Consistency, and Prompt Multiplicity. | Prakhar Ganesh, Reza Shokri, Golnoosh Farnadi |
| 2026 | Feature Drift: How Fine-Tuning Repurposes Representations in LLMs. | Andrey V. Galichin, Anton Korznikov, Alexey Dontsov, Oleg Rogov, Elena Tutubalina, Ivan V. Oseledets |
| 2026 | A Benchmark and Evaluation of Automated Language of Study Extraction from Computational Linguistics Publications. | Henry Gagnier, Ashwin Kirubakaran |
| 2026 | Beyond Blind Following: Evaluating Robustness of LLM Agents under Imperfect Guidance. | Yao Fu, Ran Qiu, Xinhe Wang, Jacob Sansom, Sathvika Ayyappa Prabhu, Huijie Tang, Jaekyeom Kim, Sungryull Sohn, Honglak Lee |
| 2026 | Ensemble Privacy Defense for Knowledge-Intensive LLMs against Membership Inference Attacks. | Haowei Fu, Bo Ni, Han Xu, Kunpeng Liu, Dan Lin, Tyler Derr |
| 2026 | From Delegates to Trustees: How Optimizing for Long-Term Interests Shapes Bias and Alignment in LLMs. | Suyash Fulay, Jocelyn Zhu, Michiel A. Bakker |
| 2026 | Toward Automatic Delegation Extraction in Japanese Law. | Tsuyoshi Fujita, Yuya Sawada, Yusuke Sakai, Taro Watanabe |
| 2026 | TimeMachine-bench: A Benchmark for Evaluating Model Capabilities in Repository-Level Migration Tasks. | Ryo Fujii, Makoto Morishita, Kazuki Yano, Jun Suzuki |
| 2026 | AlignFix: A Tool for Parallel Corpora Augmentation and Refinement. | Samuel Frontull, Simon Haller-Seeber |
| 2026 | Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes. | Johann Frei, Nils Feldhus, Lisa Raithel, Roland Roller, Alexander Meyer, Frank Kramer |
| 2026 | PTEB: Towards Robust Text Embedding Evaluation via Stochastic Paraphrasing at Evaluation Time with LLMs. | Manuel Frank, Haithem Afli |
| 2026 | ConLID: Supervised Contrastive Learning for Low-Resource Language Identification. | Negar Foroutan, Jakhongir Saydaliev, Grace Kim, Antoine Bosselut |