| 2026 | ACL | BracketRank: Large Language Model Document Ranking via Reasoning-based Competitive Elimination. | Abdelrahman Abdallah, Mohammed Ali, Bhawna Piryani, Adam Jatowt |
| 2026 | ACL | Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation. | Abdelrahman Abdallah, Bhawna Piryani, Jamshid Mozafari, Andreas Herzinger, Jamie Holdcroft, Adam Jatowt |
| 2026 | ACL | RECOR: Reasoning-focused Multi-turn Conversational Retrieval Benchmark. | Mohammed Ali, Abdelrahman Abdallah, Amit Agarwal, Hitesh Laxmichand Patel, Adam Jatowt |
| 2026 | ACL | It's High Time: A Survey of Temporal Question Answering. | Bhawna Piryani, Abdelrahman Abdallah, Jamshid Mozafari, Avishek Anand, Adam Jatowt |
| 2026 | EACL | Negative Sampling Techniques in Dense Retrieval: A Survey. | Laurin Wischounig, Abdelrahman Abdallah, Adam Jatowt |
| 2026 | SIGIR | Are LLM-Based Retrievers Worth Their Cost? An Empirical Study of Efficiency, Robustness, and Reasoning Overhead. | Abdelrahman Abdallah, Jamie Holdcroft, Mohammed Ali, Adam Jatowt |
| 2026 | WSDM | TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions. | Abdelrahman Abdallah, Bhawna Piryani, Jonas Wallat, Avishek Anand, Adam Jatowt |
| 2025 | ACL | A Study into Investigating Temporal Robustness of LLMs. | Jonas Wallat, Abdelrahman Abdallah, Adam Jatowt, Avishek Anand |
| 2025 | CIKM | RerankArena: A Unified Platform for Evaluating Retrieval, Reranking and RAG with Human and LLM Feedback. | Abdelrahman Abdallah, Mahmoud Abdalla, Bhawna Piryani, Jamshid Mozafari, Mohammed Ali, Adam Jatowt |
| 2025 | CIKM | Evaluating Robustness of LLMs in Question Answering on Multilingual Noisy OCR Data. | Bhawna Piryani, Jamshid Mozafari, Abdelrahman Abdallah, Antoine Doucet, Adam Jatowt |
| 2025 | COLING | DynRank: Improve Passage Retrieval with Dynamic Zero-Shot Prompting Based on Question Classification. | Abdelrahman Abdallah, Jamshid Mozafari, Bhawna Piryani, Mohammed M. Abdelgwad, Adam Jatowt |
| 2025 | EMNLP | DeAR: Dual-Stage Document Reranking with Reasoning Agents via LLM Distillation. | Abdelrahman Abdallah, Jamshid Mozafari, Bhawna Piryani, Adam Jatowt |
| 2025 | EMNLP | How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models. | Abdelrahman Abdallah, Bhawna Piryani, Jamshid Mozafari, Mohammed Ali, Adam Jatowt |
| 2025 | EMNLP | ComplexTempQA: A 100m Dataset for Complex Temporal Question Answering. | Raphael Gruber, Abdelrahman Abdallah, Michael Frber, Adam Jatowt |
| 2025 | NAACL | ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval. | Abdelrahman Abdallah, Jamshid Mozafari, Bhawna Piryani, Adam Jatowt |
| 2025 | SIGIR | Wrong Answers Can Also Be Useful: PlausibleQA - A Large-Scale QA Dataset with Answer Plausibility Scores. | Jamshid Mozafari, Abdelrahman Abdallah, Bhawna Piryani, Adam Jatowt |
| 2024 | EMNLP | Exploring Hint Generation Approaches for Open-Domain Question Answering. | Jamshid Mozafari, Abdelrahman Abdallah, Bhawna Piryani, Adam Jatowt |
| 2024 | EMNLP | Detecting Temporal Ambiguity in Questions. | Bhawna Piryani, Abdelrahman Abdallah, Jamshid Mozafari, Adam Jatowt |
| 2024 | MICCAI | IHRRB-DINO: Identifying High-Risk Regions of Breast Masses in Mammogram Images Using Data-Driven Instance Noise (DINO). | Mahmoud SalahEldin Kasem, Abdelrahman Abdallah, Ibrahim Abdelhalim, Norah Saleh Alghamdi, Sohail Contractor, Ayman El-Baz |
| 2024 | SIGIR | ArabicaQA: A Comprehensive Dataset for Arabic Question Answering. | Abdelrahman Abdallah, Mahmoud SalahEldin Kasem, Mahmoud Abdalla, Mohamed Mahmoud, Mohamed Elkasaby, Yasser Elbendary, Adam Jatowt |