| 2026 | ICTIR | Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage. | Saron Samuel, Alexander Martin, Eugene Yang, Andrew Yates, Dawn Lawrie, Ian Soboroff, Laura Dietz, Benjamin Van Durme |
| 2026 | SIGIR | CoverageBench: Evaluating Information Coverage across Tasks and Domains. | Saron Samuel, Andrew Yates, Dawn J. Lawrie, Ian Soboroff, Trevor Adriaanse, Benjamin Van Durme, Eugene Yang |
| 2025 | ECIR | The Second Search Futures Workshop at ECIR'25. | Charles Clarke, Paul B. Kantor, Adam Roegiest, Ian Soboroff, Johanne R. Trippas, Zhaochun Ren |
| 2025 | ICTIR | A Large-Scale Study of Relevance Assessments with Large Language Models Using UMBRELA. | Shivani Upadhyay, Ronak Pradeep, Nandan Thakur, Daniel Campos, Nick Craswell, Ian Soboroff, Jimmy Lin |
| 2025 | NAACL | Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance. | Somnath Banerjee, Avik Halder, Rajarshi Mandal, Sayan Layek, Ian Soboroff, Rima Hazra, Animesh Mukherjee |
| 2025 | SIGIR | The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models. | Ronak Pradeep, Nandan Thakur, Shivani Upadhyay, Daniel Campos, Nick Craswell, Ian Soboroff, Hoa Trang Dang, Jimmy Lin |
| 2025 | SIGIR | LLM-Assisted Relevance Assessments: When Should We Ask LLMs for Help? | Rikiya Takehi, Ellen M. Voorhees, Tetsuya Sakai, Ian Soboroff |
| 2025 | SIGIR | Assessing Support for the TREC 2024 RAG Track: A Large-Scale Comparative Study of LLM and Human Evaluations. | Nandan Thakur, Ronak Pradeep, Shivani Upadhyay, Daniel Campos, Nick Craswell, Ian Soboroff, Hoa Trang Dang, Jimmy Lin |
| 2025 | SIGIR | Nugget-based Annotation Protocol and Tool For Evaluating Long-form Retrieval-Augmented Generation. | Eugene Yang, Dawn J. Lawrie, Hoa Dang, Ian Soboroff, James Mayfield |
| 2024 | SIGIR | Browsing and Searching Metadata of TREC. | Timo Breuer, Ellen M. Voorhees, Ian Soboroff |
| 2024 | SIGIR | On the Evaluation of Machine-Generated Reports. | James Mayfield, Eugene Yang, Dawn J. Lawrie, Sean MacAvaney, Paul McNamee, Douglas W. Oard, Luca Soldaini, Ian Soboroff, Orion Weller, Efsun Selin Kayi, Kate Sanders, Marc Mason, Noah Hibbler |
| 2023 | SIGIR | The BETTER Cross-Language Datasets. | Ian Soboroff |
| 2022 | CHIIR | CHIIR Workshop on Audio Collection Human Interaction (AudioCHI 2022): http: //speechretrievalworkshop.github.io. | Gareth J. F. Jones, Maria Eskevich, Ben Carterette, Joana Correia, Rosie Jones, Jussi Karlgren, Ian Soboroff |
| 2022 | ICMI | Second International Workshop on Deep Video Understanding. | Keith Curtis, George Awad, Shahzad Rajput, Ian Soboroff |
| 2022 | SIGIR | What Makes a Good Podcast Summary? | Rezvaneh Rezapour, Sravana Reddy, Rosie Jones, Ian Soboroff |
| 2021 | SIGIR | Podcast Metadata and Content: Episode Relevance and Attractiveness in Ad Hoc Search. | Ben Carterette, Rosie Jones, Gareth J. F. Jones, Maria Eskevich, Sravana Reddy, Ann Clifton, Yongze Yu, Jussi Karlgren, Ian Soboroff |
| 2021 | SIGIR | TREC Deep Learning Track: Reusable Test Collections in the Large Data Regime. | Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, Ellen M. Voorhees, Ian Soboroff |
| 2020 | ICMI | International Workshop on Deep Video Understanding. | Keith Curtis, George Awad, Shahzad Rajput, Ian Soboroff |
| 2020 | SIGIR | How to Measure the Reproducibility of System-oriented IR Experiments. | Timo Breuer, Nicola Ferro, Norbert Fuhr, Maria Maistro, Tetsuya Sakai, Philipp Schaer, Ian Soboroff |
| 2019 | ECIR | CENTRE@CLEF 2019. | Nicola Ferro, Norbert Fuhr, Maria Maistro, Tetsuya Sakai, Ian Soboroff |
| 2019 | SIGMOD | Financial Entity Identification and Information Integration (FEIII) 2019 Challenge: The Report of the Organizing Committee. | Louiqa Raschid, Douglas Burdick, Cesar de Pablo, Mark D. Flood, John Grant, Joe Langsam, Jay Pujara, Elena Tomas, Ian Soboroff |
| 2018 | CIKM | Meta-Analysis for Retrieval Experiments Involving Multiple Test Collections. | Ian Soboroff |
| 2018 | ECIR | TREC 2018 News Track. | Shudong Huang, Ian Soboroff, Donna Harman |
| 2018 | SIGMOD | Financial Entity Identification and Information Integration (FEIII) 2018 Challenge: The Report of the Organizing Committee. | Louiqa Raschid, Douglas Burdick, John Grant, Joe Langsam, Jay Pujara, Elizabeth Roman, Ian Soboroff, Mohammed J. Zaki, Elena Zotkina |
| 2017 | SIGIR | Building Test Collections: An Interactive Guide for Students and Others Without Their Own Evaluation Conference Series. | Ian Soboroff |
| 2017 | SIGMOD | Financial Entity Identification and Information Integration (FEIII) 2017 Challenge: The Report of the Organizing Committee. | Louiqa Raschid, Douglas Burdick, Mark D. Flood, John Grant, Joe Langsam, Ian Soboroff |
| 2016 | CIKM | Financial Entity Identification and Information Integration (FEIII) Challenge: The Report of the Organizing Committee. | Mark D. Flood, John Grant, Haiping Luo, Louiqa Raschid, Ian Soboroff, Kyungjin Yoo |
| 2016 | SIGIR | The BOLT IR Test Collections of Multilingual Passage Retrieval from Discussion Forums. | Ian Soboroff, Kira Griffitt, Stephanie M. Strassel |
| 2016 | SIGIR | Privacy-Preserving IR 2016: Differential Privacy, Search, and Social Media. | Grace Hui Yang, Ian Soboroff, Li Xiong, Charles L. A. Clarke, Simson L. Garfinkel |
| 2015 | SIGIR | Privacy-Preserving IR 2015: When Information Retrieval Meets Privacy and Security. | Hui Yang, Ian Soboroff |
| 2013 | SIGIR | Building test collections: an interactive tutorial for students and others without their own evaluation conference series. | Ian Soboroff |
| 2012 | SIGIR | On building a reusable Twitter corpus. | Richard McCreadie, Ian Soboroff, Jimmy Lin, Craig Macdonald, Iadh Ounis, Dean McCullough |
| 2011 | WSDM | A comparative analysis of cascade measures for novelty and diversity. | Charles L. A. Clarke, Nick Craswell, Ian Soboroff, Azin Ashkan |
| 2010 | SIGIR | The effect of assessor error on IR system evaluation. | Ben Carterette, Ian Soboroff |
| 2009 | SIGIR | Is spam an issue for opinionated blog post search? | Craig Macdonald, Iadh Ounis, Ian Soboroff |
| 2008 | ICWSM | On the TREC Blog Track. | Iadh Ounis, Craig Macdonald, Ian Soboroff |
| 2008 | SIGIR | Relevance assessment: are judges exchangeable and does it matter. | Peter Bailey, Nick Craswell, Ian Soboroff, Paul Thomas, Arjen P. de Vries, Emine Yilmaz |
| 2008 | SIGIR | Limits of opinion-finding baseline systems. | Craig Macdonald, Ben He, Iadh Ounis, Ian Soboroff |
| 2007 | SIGIR | Reliable information retrieval evaluation with incomplete and biased judgements. | Stefan Bttcher, Charles L. A. Clarke, Peter C. K. Yeung, Ian Soboroff |
| 2007 | SIGIR | Problems with Kendall's tau. | Mark Sanderson, Ian Soboroff |
| 2007 | SIGIR | A comparison of pooled and sampled relevance judgments. | Ian Soboroff |
| 2006 | SIGIR | Bias and the limits of pooling. | Chris Buckley, Darrin Dimmick, Ian Soboroff, Ellen M. Voorhees |
| 2006 | SIGIR | Dynamic test collections: measuring search effectiveness on the live web. | Ian Soboroff |
| 2005 | NAACL | Novelty Detection: The TREC Experience. | Ian Soboroff, Donna Harman |
| 2004 | SIGIR | On evaluating web search with very few relevant documents. | Ian Soboroff |
| 2003 | SIGIR | Building a filtering test collection for TREC 2002. | Ian Soboroff, Stephen E. Robertson |
| 2002 | SIGIR | Automatic evaluation of world wide web search services. | Abdur Chowdhury, Ian Soboroff |
| 2002 | SIGIR | Does WT10g look like the web?. | Ian Soboroff |
| 2001 | SIGIR | Ranking Retrieval Systems without Relevance Judgments. | Ian Soboroff, Charles K. Nicholas, Patrick Cahan |
| 2000 | SIGIR | Collaborative filtering and the generalized vector space model. | Ian Soboroff, Charles K. Nicholas |
| 1997 | CIKM | Visualizing Document Authorship Using n-grams and Latent Semantic Indexing. | Ian Soboroff, Charles K. Nicholas, James M. Kukla, David S. Ebert |