| 2026 | DIRECT: Directional Relevance in Conversational Trajectories. | Anshuman Mourya, Rajdeep Mukherjee, Prerna Jolly, Vinayak S. Puranik, Sivaramakrishnan R. Kaveri |
| 2026 | Argument-Based Consistency in Toxicity Explanations of LLMs. | Ramaravind Kommiya Mothilal, Joanna Roy, Syed Ishtiaque Ahmed, Shion Guha |
| 2026 | Bring the Apple, Not the Sofa: Impact of Irrelevant Context in Embodied AI Commands on VLA Models. | Andrey Moskalenko, Daria Pugacheva, Denis Shepelev, Andrey Kuznetsov, Vlad Shakhuro, Elena Tutubalina |
| 2026 | FactAppeal: Identifying Epistemic Factual Appeals in News Media. | Guy Mor-Lan, Tamir Sheafer, Shaul R. Shenhav |
| 2026 | Don't Judge Code by Its Cover: Exploring Biases in LLM Judges for Code Evaluation. | Jiwon Moon, Yerin Hwang, Dongryeol Lee, Taegwan Kang, Yongil Kim, Kyomin Jung |
| 2026 | Event Detection with a Context-Aware Encoder and LoRA for Improved Performance on Long-Tailed Classes. | Abdullah Al Monsur, Nitesh Vamshi Bommisetty, Gene Louis Kim |
| 2026 | SMART-Editor: A Multi-Agent Framework for Human-Like Design Editing with Structural Integrity. | Ishani Mondal, Meera Bharadwaj, Ayush Roy, Aparna Garimella, Jordan Lee Boyd-Graber |
| 2026 | Surprisal and Metaphor Novelty Judgments: Moderate Correlations and Divergent Scaling Effects Revealed by Corpus-Based and Synthetic Datasets. | Omar Momen, Emilie Sitter, J. Berenike Herrmann, Sina Zarrie |
| 2026 | Navigating the Infinite Dynamic Web Space: Effective In-Context Exploration via Cognitive Multi-Agent Collaboration. | Guozhao Mo, Yanjiang Liu, Yafei Shi, Jiawei Chen, Yang Li, Yaojie Lu, Hongyu Lin, Ben He, Le Sun, Bo Zheng, Xianpei Han |
| 2026 | Exploring Fine-Tuning for In-Context Retrieval and Efficient KV-Caching in Long-Context Language Models. | Francesco Maria Molfese, Momchil Hardalov, Rexhina Blloshmi, Bill Byrne, Adri de Gispert |
| 2026 | Unlocking Latent Discourse Translation in LLMs Through Quality-Aware Decoding. | Wafaa Mohammed, Vlad Niculae, Chrysoula Zerva |
| 2026 | LingVarBench: Benchmarking LLMs on Entity Recognitions and Linguistic Verbalization Patterns in Phone-Call Transcripts. | Seyedali Mohammadi, Manas Paldhe, Amit Chhabra, Youngseo Son, Vishal Seshagiri |
| 2026 | MALicious INTent Dataset and Inoculating LLMs for Enhanced Disinformation Detection. | Arkadiusz Modzelewski, Witold Sosnowski, Eleni Papadopulos, Elisa Sartori, Tiziano Labruna, Giovanni Da San Martino, Adam Wierzbicki |
| 2026 | RECAP: REwriting Conversations for Intent Understanding in Agentic Planning. | Kushan Mitra, Dan Zhang, Hannah Kim, Estevam Hruschka |
| 2026 | SD-E2: Semantic Exploration for Reasoning Under Token Budgets. | Kshitij Mishra, Nils Lukas, Salem Lahlou |
| 2026 | Router-Suggest: Dynamic Routing for Multimodal Auto-Completion in Visually-Grounded Dialogs. | Sandeep Mishra, Devichand Budagam, Anubhab Mandal, Bishal Santra, Pawan Goyal, Manish Gupta |
| 2026 | The Model's Language Matters: A Comparative Privacy Analysis of LLMs. | Abhishek K. Mishra, Antoine Boutet, Lucas Magnana |
| 2026 | Thesis Proposal: A Multi-Agent System for Ontology-Based Perspective-Aware Knowledge Extraction. | Luiz do Valle Miranda, Grzegorz J. Nalepa |
| 2026 | A Browser-based Open Source Assistant for Multimodal Content Verification. | Rosanna Milner, Michael Foster, Olesya Razuvayevskaya, Valentin Porcellini, Denis Teyssou, Ian Roberts, Kalina Bontcheva |
| 2026 | Thinking Long, but Short: Stable Sequential Test-Time Scaling for Large Reasoning Models. | Michael R. Metel, Yufei Cui, Boxing Chen, Prasanna Parthasarathi |
| 2026 | ConvApparel: A Benchmark Dataset and Validation Framework for User Simulators in Conversational Recommenders. | Ofer Meshi, Krisztian Balog, Sally Goldman, Avi Caciularu, Guy Tennenholtz, Jihwan Jeong, Amir Globerson, Craig Boutilier |
| 2026 | Detecting Primary Progressive Aphasia (PPA) from Text: A Benchmarking Study. | Ghofrane Merhbene, Fabian Lecron, Philippe Fortemps, Bradford C. Dickerson, Mascha Kurpicz-Briki, Neguine Rezaii |
| 2026 | Can LLMs Reason Like Doctors? Exploring the Limits of Large Language Models in Complex Medical Reasoning. | Flavio Merenda, Jos Manul Gmez-Prez, German Rigau |
| 2026 | ParsTranslit: Truly Versatile Tajik-Farsi Transliteration. | Rayyan Merchant, Kevin Tang |
| 2026 | MEDAL: A Framework for Benchmarking LLMs as Multilingual Open-Domain Dialogue Evaluators. | John Mendona, Alon Lavie, Isabel Trancoso |