Empirical Methods in Natural Language Processing
EMNLP
A*
CORE rank
CORE rank (raw)
A*
Acceptance rate
23.5% (2023 main conference)
Fields of research
Artificial Intelligence
Papers indexed
16,441
1996–2025
Papers per year
19963,494 peak2025
Most published authors
EMNLP papers
16,441 records sourced from DBLP. Search titles, filter by year, sort by recency.
| Year | Title | Authors |
|---|---|---|
| 2025 | CAVE : Detecting and Explaining Commonsense Anomalies in Visual Environments. | Rishika Bhagwatkar, Syrielle Montariol, Angelika Romanou, Beatriz Borges, Irina Rish, Antoine Bosselut |
| 2025 | Evaluating Compound AI Systems through Behaviors, Not Benchmarks. | Pranav Bhagat, K. N. Ajay Shastry, Pranoy Panda, Chaitanya Devaguptapu |
| 2025 | For a Fistful of Puns: Evaluating a Puns in Multiword Expressions Identification Algorithm Without Dedicated Dataset. | Julien Bezanon, Gal Lejeune |
| 2025 | GEAR: A Scalable and Interpretable Evaluation Framework for RAG-Based Car Assistant Systems. | Niloufar Beyranvand, Hamidreza Dastmalchi, Aijun An, Heidar Davoudi, Winston Chan, Ron DiCarlantonio |
| 2025 | The Validation Gap: A Mechanistic Analysis of How Language Models Compute Arithmetic but Fail to Validate It. | Leonardo Bertolazzi, Philipp Mondorf, Barbara Plank, Raffaella Bernardi |
| 2025 | LLM-Based Web Data Collection for Research Dataset Creation. | Thomas Berkane, Marie-Laure Charpignon, Maimuna S. Majumder |
| 2025 | Learning to Translate Ambiguous Terminology by Preference Optimization on Post-Edits. | Nathaniel Berger, Johannes Eschbach-Dymanus, Miriam Exel, Matthias Huck, Stefan Riezler |
| 2025 | Improving Informally Romanized Language Identification. | Adrian Benton, Alexander Gutkin, Christo Kirov, Brian Roark |
| 2025 | UnityAI Guard: Pioneering Toxicity Detection Across Low-Resource Indian Languages. | Himanshu Beniwal, Reddybathuni Venkat, Rohit Kumar, Birudugadda Srivibhav, Daksh Jain, Pavan Doddi, Eshwar Dhande, Adithya Ananth, Kuldeep, Mayank Singh |
| 2025 | Beyond Averages: Learning with Annotator Disagreement in STS. | Alejandro Benito-Santos, Adrin Ghajari |
| 2025 | Type-Less yet Type-Aware Inductive Link Prediction with Pretrained Language Models. | Alessandro De Bellis, Salvatore Bufi, Giovanni Servedio, Vito Walter Anelli, Tommaso Di Noia, Eugenio Di Sciascio |
| 2025 | Overcoming Black-box Attack Inefficiency with Hybrid and Dynamic Select Algorithms. | Abhinay Shankar Belde, Rohit Ramkumar, Jonathan Rusert |
| 2025 | AfroXLMR-Social: Adapting Pre-trained Language Models for African Languages Social Media Text. | Tadesse Destaw Belay, Israel Abebe Azime, Ibrahim Said Ahmad, David Ifeoluwa Adelani, Idris Abdulmumin, Abinew Ali Ayele, Shamsuddeen Hassan Muhammad, Seid Muhie Yimam |
| 2025 | Sycophancy Mitigation Through Reinforcement Learning with Uncertainty-Aware Adaptive Reasoning Trajectories. | Mohammad Beigi, Ying Shen, Parshin Shojaee, Qifan Wang, Zichao Wang, Chandan K. Reddy, Ming Jin, Lifu Huang |
| 2025 | Scaling Down, Serving Fast: Compressing and Deploying Efficient LLMs for Recommendation Systems. | Kayhan Behdin, Ata Fatahi Baarzi, Qingquan Song, Yun Dai, Aman Gupta, Zhipeng Wang, Hejian Sang, Shao Tang, Gregory Dexter, Sirou Zhu, Siyu Zhu, Tejas Dharamsi, Vignesh Kothapalli, Zhoutong Fu, Yihan Cao, Pin-Lun Hsu, Fedor Borisyuk, Natesh S. Pillai, Luke Simon, Rahul Mazumder |
| 2025 | MALLM: Multi-Agent Large Language Models Framework. | Jonas Becker, Lars Benedikt Kaesberg, Niklas Bauer, Jan Philip Wahle, Terry Ruas, Bela Gipp |
| 2025 | QFrCoLA: a Quebec-French Corpus of Linguistic Acceptability Judgments. | David Beauchemin, Richard Khoury |
| 2025 | JUDGEBERT: Assessing Legal Meaning Preservation Between Sentences. | David Beauchemin, Michelle Albert-Rochette, Richard Khoury, Pierre-Luc Dziel |
| 2025 | TR-MTEB: A Comprehensive Benchmark and Embedding Model Suite for Turkish Sentence Representations. | Mehmet Selman Baysan, Tunga Gungor |
| 2025 | NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls. | Kinjal Basu, Ibrahim Abdelaziz, Kiran Kate, Mayank Agarwal, Maxwell Crouse, Yara Rizk, Kelsey Bradford, Asim Munawar, Sadhana Kumaravel, Saurabh Goyal, Xin Wang, Luis A. Lastras, Pavan Kapanipathi |
| 2025 | On Guardrail Models' Robustness to Mutations and Adversarial Attacks. | Elias Bassani, Ignacio Sanchez |
| 2025 | TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs. | Ezgi Basar, Francesca Padovani, Jaap Jumelet, Arianna Bisazza |
| 2025 | StandUp4AI: A New Multilingual Dataset for Humor Detection in Stand-up Comedy Videos. | Valentin Barrire, Nahuel Gomez, Lo Hemamou, Sofa Callejas, Brian Ravenet |
| 2025 | Can LLMs Find a Needle in a Haystack? A Look at Anomaly Detection Language Modeling. | Leslie Barrett, Vikram Sunil Bajaj, Robert J. Kingan |
| 2025 | This is not a Disimprovement: Improving Negation Reasoning in Large Language Models via Prompt Engineering. | Joshua Jose Dias Barreto, Abhik Jana |
3,301–3,325 of 16,441← PreviousNext →
Comparable venues
Other A*/A conferences filed under the same field of research.
- A*AAAINational Conference of the American Association for Artificial Intelligence
- A*ICRAIEEE International Conference on Robotics and Automation
- AInterspeechInterspeech (combined EuroSpeech and ICSLP in 2000)
- AIROSIEEE/RSJ International Conference on Intelligent Robots and Systems
- A*ACLAssociation for Computational Linguistics
- A*IJCAIInternational Joint Conference on Artificial Intelligence
- AGECCOGenetic and Evolutionary Computations
- ANAACLNorth American Association for Computational Linguistics