| 2026 | Sycophants in the Courtroom: Are LLMs Fragile to Juridical Authority and Evolving Legal Standards? | Lorenzo Molfetta, Alessio Cocchieri, Luca Ragazzi, Ilaria Bartolini, Marco Patella, Gianluca Moro |
| 2026 | ReTraceQA: Evaluating Reasoning Traces of Small Language Models in Commonsense Question Answering. | Francesco Maria Molfese, Luca Moroni, Ciro Porcaro, Simone Conia, Roberto Navigli |
| 2026 | Frame of Reference: Addressing the Challenges of Common Ground Representation in Situational Dialogs. | Biswesh Mohapatra, Tho Charlot, Giovanni Duca, Mayank Palan, Laurent Romary, Justine Cassell |
| 2026 | Are Large Language Models Economically Viable for Industry Deployment? | Abdullah Mohammad, Sushant Kumar Ray, Pushkar Arora, Rafiq Ali, Ebad Shabbir, Gautam Siddharth Kashyap, Jiechao Gao, Usman Naseem |
| 2026 | Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models. | Seyedali Mohammadi, Manas Gaur, Francis Ferraro |
| 2026 | Fast-Decoding Diffusion Language Models via Progress-Aware Confidence Schedules. | Amr Mohamed, Yang Zhang, Michalis Vazirgiannis, Guokan Shang |
| 2026 | Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text. | Amr Mohamed, Yang Zhang, Michalis Vazirgiannis, Guokan Shang |
| 2026 | Sem-DPO: Mitigating Semantic Inconsistency in Preference Optimization for Prompt Engineering. | Anas Mohamed, Azal Ahmad Khan, Xinran Wang, Ahmad Faraz Khan, Shuwen Ge, Saman Bahzad Khan, Ayaan Ahmad, Ali Anwar |
| 2026 | Agentic Conversational Search with Contextualized Reasoning via Reinforcement Learning. | Fengran Mo, Yifan Gao, Sha Li, Hansi Zeng, Xin Liu, Zhaoxuan Tan, Xian Li, Jianshu Chen, Dakuo Wang, Meng Jiang |
| 2026 | RePrompT: Recurrent Prompt Tuning for Integrating Structured EHR Encoders with Large Language Models. | Arya Hadizadeh Moghaddam, Drew Ross, Mohsen Nayebi Kerdabadi, Dongjie Wang, Zijun Yao |
| 2026 | SciText2Eq: Assessing LLMs for Explainable Equation Generation for Scientific Creativity. | Yifan Mo, Xiao Fu, Yue Su, Qingyu Meng, Koen V. Hindriks, Qingzhi Liu, Jiahuan Pei |
| 2026 | Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences. | Arkadiusz Modzelewski, Pawel Golik, Anna Kolos, Giovanni Da San Martino |
| 2026 | Disentangling the Effects of Unlearning in Measuring Parametric Faithfulness of Chain-of-Thought. | Ryo Mitsuhashi, Gaku Morio, Ayana Niwa, Masahiro Kaneko, Kentaro Inui, Terufumi Morishita, Yuta Koreeda, Yasuhiro Sogawa |
| 2026 | Language Acquisition Device in Large Language Models. | Masato Mita, Taiga Someya, Ryo Yoshida, Yohei Oseki |
| 2026 | Logical Consistency as a Bridge: Improving LLM Hallucination Detection via Label Constraint Modeling between Responses and Self-Judgments. | Hao Mi, Qiang Sheng, Shaofei Wang, Beizhe Hu, Yifan Sun, Zhengjia Wang, Hengqi Zeng, Yang Li, Danding Wang, Juan Cao |
| 2026 | GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities. | Diganta Misra, Nizar Islah, Victor May, Brice Rauby, Zihan Wang, Justine Gehring, Antonio Orvieto, Muawiz Sajjad Chaudhary, Eilif B. Muller, Irina Rish, Samira Ebrahimi Kahou, Massimo Caccia |
| 2026 | CEBC: Conformal Evidence-Bounded Control for Low-Hallucination Vision-Language Generation. | Ashish Mishra, Tarun Kumar, Arpit Shah, Suparna Bhattacharya, Martin Foltin |
| 2026 | PE-QAT: Parameter-Efficient Quantization-Aware Training for Large Language Models. | Shresth Mishra |
| 2026 | Social Story Frames: Contextual Reasoning about Narrative Intent and Reception. | Joel Mire, Maria Antoniak, Steven R. Wilson, Zexin Ma, Achyutarama R. Ganti, Andrew Piper, Maarten Sap |
| 2026 | QuCo-RAG: Quantifying Uncertainty from the Pre-training Corpus for Dynamic Retrieval-Augmented Generation. | Dehai Min, Kailin Zhang, Tongtong Wu, Lu Cheng |
| 2026 | Orchestrating Tokens and Sequences: Dynamic Hybrid Policy Optimization for RLVR. | Zijun Min, Bingshuai Liu, Ante Wang, Long Zhang, Anxiang Zeng, Haibo Zhang, Jinsong Su |
| 2026 | GOAT: A Training Framework for Goal-Oriented Agent with Tools. | Hyunji Min, Sangwon Jung, Junyoung Sung, Dosung Lee, Leekyeung Han, Paul Hongsuck Seo |
| 2026 | Progressive Re-ranking for Multimodal Retrieval-Augmented Generation via Curriculum Learning. | Zhu Min, Yanchao Hao, Jian Liu, Shizhu He, Xi Chen |
| 2026 | MonCulture-Eval: A Hierarchical Benchmark for Evaluating Mongolian Cultural Capabilities of Large Language Models across Scripts and Regions. | Quulgan Minggad, Zinan Xiao, Yuan Sun |
| 2026 | Understanding Emergent Misalignment via Feature Superposition Geometry. | Gouki Minegishi, Hiroki Furuta, Takeshi Kojima, Yusuke Iwasawa, Yutaka Matsuo |