Skip to content

Empirical Methods in Natural Language Processing

EMNLP

A*

CORE rank

CORE rank (raw)

A*

Acceptance rate

23.5% (2023 main conference)

Fields of research

Artificial Intelligence

Papers indexed

16,441

1996–2025

Papers per year

19963,494 peak2025

EMNLP papers

16,441 records sourced from DBLP. Search titles, filter by year, sort by recency.

YearTitleAuthors
2025CAVE : Detecting and Explaining Commonsense Anomalies in Visual Environments.Rishika Bhagwatkar, Syrielle Montariol, Angelika Romanou, Beatriz Borges, Irina Rish, Antoine Bosselut
2025Evaluating Compound AI Systems through Behaviors, Not Benchmarks.Pranav Bhagat, K. N. Ajay Shastry, Pranoy Panda, Chaitanya Devaguptapu
2025For a Fistful of Puns: Evaluating a Puns in Multiword Expressions Identification Algorithm Without Dedicated Dataset.Julien Bezanon, Gal Lejeune
2025GEAR: A Scalable and Interpretable Evaluation Framework for RAG-Based Car Assistant Systems.Niloufar Beyranvand, Hamidreza Dastmalchi, Aijun An, Heidar Davoudi, Winston Chan, Ron DiCarlantonio
2025The Validation Gap: A Mechanistic Analysis of How Language Models Compute Arithmetic but Fail to Validate It.Leonardo Bertolazzi, Philipp Mondorf, Barbara Plank, Raffaella Bernardi
2025LLM-Based Web Data Collection for Research Dataset Creation.Thomas Berkane, Marie-Laure Charpignon, Maimuna S. Majumder
2025Learning to Translate Ambiguous Terminology by Preference Optimization on Post-Edits.Nathaniel Berger, Johannes Eschbach-Dymanus, Miriam Exel, Matthias Huck, Stefan Riezler
2025Improving Informally Romanized Language Identification.Adrian Benton, Alexander Gutkin, Christo Kirov, Brian Roark
2025UnityAI Guard: Pioneering Toxicity Detection Across Low-Resource Indian Languages.Himanshu Beniwal, Reddybathuni Venkat, Rohit Kumar, Birudugadda Srivibhav, Daksh Jain, Pavan Doddi, Eshwar Dhande, Adithya Ananth, Kuldeep, Mayank Singh
2025Beyond Averages: Learning with Annotator Disagreement in STS.Alejandro Benito-Santos, Adrin Ghajari
2025Type-Less yet Type-Aware Inductive Link Prediction with Pretrained Language Models.Alessandro De Bellis, Salvatore Bufi, Giovanni Servedio, Vito Walter Anelli, Tommaso Di Noia, Eugenio Di Sciascio
2025Overcoming Black-box Attack Inefficiency with Hybrid and Dynamic Select Algorithms.Abhinay Shankar Belde, Rohit Ramkumar, Jonathan Rusert
2025AfroXLMR-Social: Adapting Pre-trained Language Models for African Languages Social Media Text.Tadesse Destaw Belay, Israel Abebe Azime, Ibrahim Said Ahmad, David Ifeoluwa Adelani, Idris Abdulmumin, Abinew Ali Ayele, Shamsuddeen Hassan Muhammad, Seid Muhie Yimam
2025Sycophancy Mitigation Through Reinforcement Learning with Uncertainty-Aware Adaptive Reasoning Trajectories.Mohammad Beigi, Ying Shen, Parshin Shojaee, Qifan Wang, Zichao Wang, Chandan K. Reddy, Ming Jin, Lifu Huang
2025Scaling Down, Serving Fast: Compressing and Deploying Efficient LLMs for Recommendation Systems.Kayhan Behdin, Ata Fatahi Baarzi, Qingquan Song, Yun Dai, Aman Gupta, Zhipeng Wang, Hejian Sang, Shao Tang, Gregory Dexter, Sirou Zhu, Siyu Zhu, Tejas Dharamsi, Vignesh Kothapalli, Zhoutong Fu, Yihan Cao, Pin-Lun Hsu, Fedor Borisyuk, Natesh S. Pillai, Luke Simon, Rahul Mazumder
2025MALLM: Multi-Agent Large Language Models Framework.Jonas Becker, Lars Benedikt Kaesberg, Niklas Bauer, Jan Philip Wahle, Terry Ruas, Bela Gipp
2025QFrCoLA: a Quebec-French Corpus of Linguistic Acceptability Judgments.David Beauchemin, Richard Khoury
2025JUDGEBERT: Assessing Legal Meaning Preservation Between Sentences.David Beauchemin, Michelle Albert-Rochette, Richard Khoury, Pierre-Luc Dziel
2025TR-MTEB: A Comprehensive Benchmark and Embedding Model Suite for Turkish Sentence Representations.Mehmet Selman Baysan, Tunga Gungor
2025NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls.Kinjal Basu, Ibrahim Abdelaziz, Kiran Kate, Mayank Agarwal, Maxwell Crouse, Yara Rizk, Kelsey Bradford, Asim Munawar, Sadhana Kumaravel, Saurabh Goyal, Xin Wang, Luis A. Lastras, Pavan Kapanipathi
2025On Guardrail Models' Robustness to Mutations and Adversarial Attacks.Elias Bassani, Ignacio Sanchez
2025TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs.Ezgi Basar, Francesca Padovani, Jaap Jumelet, Arianna Bisazza
2025StandUp4AI: A New Multilingual Dataset for Humor Detection in Stand-up Comedy Videos.Valentin Barrire, Nahuel Gomez, Lo Hemamou, Sofa Callejas, Brian Ravenet
2025Can LLMs Find a Needle in a Haystack? A Look at Anomaly Detection Language Modeling.Leslie Barrett, Vikram Sunil Bajaj, Robert J. Kingan
2025This is not a Disimprovement: Improving Negation Reasoning in Large Language Models via Prompt Engineering.Joshua Jose Dias Barreto, Abhik Jana
3,3013,325 of 16,441← PreviousNext →

Comparable venues

Other A*/A conferences filed under the same field of research.