| 2026 | Adaptive Token-Aware Query Reformulation for Text-to-Image Retrieval. | Seonah Kim, Minkeon Kim, Youjin Lee, Jaekwang Kim, Jinyoung Han, Eunil Park |
| 2026 | Rank, Don't Generate: Statement-level Ranking for Explainable Recommendation. | Ben Kabongo, Arthur Satouf, Vincent Guigue |
| 2026 | A Replicability Study of XTR. | Rohan Jha, Reno Kriz, Benjamin Van Durme |
| 2026 | ARIC: A Cognitive Framework for Explanatory Narrative Evaluation in Conversational Information Seeking Systems. | Vahid Sadiri Javadi, Sadia Naseer, Ali Ather, Lucie Flek, Johanne R. Trippas |
| 2026 | To Believe or Not To Believe: Comparing Supporting Information Tools to Aid Human Judgments of AI Veracity. | Jessica Irons, Patrick S. Cooper, Necva Blc, Andreas Duenser, Roelien C. Timmer, Huichen Yang, Changhyun Lee, Brian Jin, Stephen Wan |
| 2026 | Does Form Affect Function? An Extended Study of LLM Re-Ranking Behavior. | Reyhaneh Goli, Alistair Moffat |
| 2026 | Entity Labels Are Not Entity Signals: A Framework for Observable Relevance in Document Re-Ranking. | Utshab Kumar Ghosh, Shubham Chatterjee |
| 2026 | Adaptive Re-Ranking. | Ata Cinar Genc, Emir Kaan Korukluoglu, James Allan |
| 2026 | Reasoning with Large Language Models for Relevance Judgements. | Louis Geiger, Danula Hettiachchi, Falk Scholer, Johanne R. Trippas |
| 2026 | A Theoretical Framework for Risk Analysis of Stochastic Rankers. | Debasis Ganguly |
| 2026 | Towards Adaptive and Retriever-friendly Retrieval-augmented Generation via Reinforcement Learning. | Yubo Fang, Hai-Tao Yu, Hideo Joho, Sumio Fujita |
| 2026 | From Noise to Order: Learning to Rank via Denoising Diffusion. | Sajad Ebrahimi, Bhaskar Mitra, Negar Arabzadeh, Ye Yuan, Haolun Wu, Fattane Zarrinkalam, Ebrahim Bagheri |
| 2026 | Bridging the Gap between Subsampled and Full-Corpus Evaluation. | Michael Dinzinger, Kanishka Ghosh Dastidar, Laura Caspari, Jelena Mitrovic, Michael Granitzer |
| 2026 | LLM-Driven Usefulness Judgment for Web Search Evaluation. | Mouly Dewan, Jiqun Liu, Aditya Gautam, Chirag Shah |
| 2026 | DeepResearchGym: A Free, Transparent, and Reproducible Sandbox for Deep Research. | Joo Coelho, Jingjie Ning, Jingyuan He, Kangrui Mao, Abhijay Sai Paladugu, Pranav Setlur, Jiahe Jin, Jamie Callan, Joo Magalhes, Bruno Martins, Chenyan Xiong |
| 2026 | Uncertainty Quantification for Multimodal Retrieval Augmented Generation. | Simon Binz, Heydar Soudani, Faegheh Hasibi |
| 2026 | Preventing Content Leakage in LLM-Based Medical RAG: Structure-Only Retrieval for Faithful Clinical Summarization. | Aleka Melese Ayalew, Tapio Seppnen, Mourad Oussalah |
| 2026 | Ranking Passages in Relevant Documents Using LLMs. | Eyal El Ani, Eilon Sheetrit, Oren Kurland |
| 2026 | TopicTune: A Topic Alignment Approach to Improve LLMs Reasoning on Social Questions. | Maryam Amirizaniani, Baktash Ansari, Chirag Shah, Afra Mashhadi |
| 2025 | Generalized Personalized PageRank with Graph Convolutional Networks in Recommender Systems. | Hiroshi Wayama, Kazunari Sugiyama |
| 2025 | Correctness is not Faithfulness in Retrieval Augmented Generation Attributions. | Jonas Wallat, Maria Heuss, Maarten de Rijke, Avishek Anand |
| 2025 | A Large-Scale Study of Relevance Assessments with Large Language Models Using UMBRELA. | Shivani Upadhyay, Ronak Pradeep, Nandan Thakur, Daniel Campos, Nick Craswell, Ian Soboroff, Jimmy Lin |
| 2025 | Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications. | Leila Tavakoli, Hamed Zamani |
| 2025 | AIDME: A Scalable, Interpretable Framework for AI-Aided Scoping Reviews. | Michael Soprano, Sandip Modha, Kevin Roitero, Eddy Maddalena, Marco Viviani, Gabriella Pasi, Stefano Mizzaro |
| 2025 | A Substring Extraction-Based RAG Method for Minimising Hallucinations in Aircraft Maintenance Question Answering. | Quentin Sign, Mohand Boughanem, Jos G. Moreno, Thiziri Belkacem |