| 2026 | Improving Chain-of-Thought for Logical Reasoning via Attention-Aware Intervention. | Phuong Minh Nguyen, Dang Huu-Tien, Naoya Inoue |
| 2026 | Beyond Coherence: Improving Temporal Consistency and Interpretability in Dynamic Topic Models. | Thanh Vinh Nguyen, Ngo Van Dong, Minh Chu Xuan, Tung Nguyen, Linh Ngo Van, Dinh Viet Sang, Trung Le |
| 2026 | Investigating Gender Stereotypes in Large Language Models via Social Determinants of Health. | Trung Hieu Ngo, Adrien Bazoge, Solen Quiniou, Pierre-Antoine Gourraud, Emmanuel Morin |
| 2026 | Modality Matching Matters: Calibrating Language Distances for Cross-Lingual Transfer in URIEL+. | York Hay Ng, Aditya Armaan Khan, Xiang Lu, Matteo Salloum, Michael Zhou, Phuong Hanh Hoang, A. Seza Dogruz, En-Shiun Annie Lee |
| 2026 | SymCode: A Neurosymbolic Approach to Mathematical Reasoning via Verifiable Code Generation. | Sina Bagheri Nezhad, Yao Li, Ameeta Agrawal |
| 2026 | Compact Multimodal Language Models as Robust OCR Alternatives for Noisy Textual Clinical Reports. | Nikita Neveditsin, Pawan Lingras, Salil Patil, Swarup Patil, Vijay Kumar Mago |
| 2026 | Do Large Language Models Reflect Demographic Pluralism in Safety? | Usman Naseem, Gautam Siddharth Kashyap, Sushant Kumar Ray, Rafiq Ali, Ebad Shabbir, Abdullah Mohammad |
| 2026 | Scaling Intent Understanding: A Framework for Classification with Clarification using Lightweight LLMs. | Subhadip Nandi, Tanishka Agarwal, Anshika Singh, Priyanka Bhatt |
| 2026 | IDEAlign: Comparing Ideas of Large Language Models to Domain Experts. | Hyunji Nam, Luca Langlois, James Malamut, Mei Tan, Dorottya Demszky |
| 2026 | When Does Auxiliary Modality Matter in Solving Geometric Problems? A Comprehensive Study of Textual, Formal, and Visual Modalities. | Hyuk Namgoong, Jeesu Jung, Yerim Han, Sangkeun Jung |
| 2026 | BitBypass: A New Direction in Jailbreaking Aligned Large Language Models with Bitstream Camouflage. | Kalyan Nakka, Nitesh Saxena |
| 2026 | Demystifying Mixed Outcomes of Self-Training: Pre-training Analyses on Non-Toy LLMs. | Yusuke Nakamura, Hirokazu Kiyomaru, Chaoran Liu, Shuhei Kurita, Daisuke Kawahara |
| 2026 | Beyond Divergent Creativity: A Human-Based Evaluation of Creativity in Large Language Models. | Kumiko Nakajima, Jan Zuiderveld, Sandro Pezzelle |
| 2026 | Offline Preference Optimization via Maximum Marginal Likelihood Estimation. | Saeed Najafi, Alona Fyshe |
| 2026 | Rethinking Schema Linking: A Context-Aware Bidirectional Retrieval Approach for Text-to-SQL. | Md Mahadi Hasan Nahid, Davood Rafiei, Weiwei Zhang, Yong Zhang |
| 2026 | Cross-lingual and Word-Independent Methods for Quantifying Degree of Grammaticalization. | Ryo Nagata, Daichi Mochihashi, Misato Ido, Yusuke Kubota, Naoki Otani, Yoshifumi Kawasaki, Hiroya Takamura |
| 2026 | AITutor-EvalKit: Exploring the Capabilities of AI Tutors. | Numaan Naeem, Kaushal Kumar Maurya, Kseniia Petukhova, Ekaterina Kochmar |
| 2026 | TechING: Towards Real World Technical Image Understanding via VLMs. | Tafazzul Nadeem, Bhavik Shangari, Manish Rai, Gagan Raj Gupta, Ashutosh Modi |
| 2026 | Seeing All Sides: Multi-Perspective In-Context Learning for Subjective NLP. | Benedetta Muscato, Yue Li, Gizem Gezici, Zhixue Zhao, Fosca Giannotti |
| 2026 | Mind Your Special Tokens! On the Importance of Dedicated Sequence-End Tokens in Vision-Language Embedding Models. | Elio Musacchio, Giovanni Semeraro, Goran Glavas |
| 2026 | RAGVUE: A Diagnostic View for Explainable and Automated Evaluation of Retrieval-Augmented Generation. | Keerthana Murugaraj, Salima Lamsiyah, Martin Theobald |
| 2026 | Don't Trust Generative Agents to Mimic Communication on Social Networks Unless You Benchmarked their Empirical Realism. | Simon Mnker, Nils Schwager, Achim Rettinger |
| 2026 | RefusalBench: Generative Evaluation of Selective Refusal in Grounded Language Models. | Aashiq Muhamed, Leonardo F. R. Ribeiro, Markus Dreyer, Virginia Smith, Mona T. Diab |
| 2026 | Nahw: A Comprehensive Benchmark of Arabic Grammar Understanding, Error Detection, Correction, and Explanation. | Hamdy Mubarak, Majd Hawasly, Abubakr Mohamed |
| 2026 | Garbage In, Reasoning Out? Why Benchmark Scores are Unreliable and What to Do About It. | Seyed Mahed Mousavi, Edoardo Cecchinato, Lucia Hornikova, Giuseppe Riccardi |