Kyle Lo
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
48
Venues
12
Active years
2018–2026
Best venue rank
A*
Where they publish
Papers
48 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | The olmOCR Project: Building Fully Open OCR using VLMs. | Jake Poznanski, Kyle Lo, Luca Soldaini |
| 2025 | AAAI | RouterRetriever: Routing over a Mixture of Expert Embedding Models. | Hyunji Lee, Luca Soldaini, Arman Cohan, Minjoon Seo, Kyle Lo |
| 2025 | ACL | Human-AI Collaboration: How AIs Augment Human Teammates. | Sherry Wu, Diyi Yang, Joseph Chang, Marti A. Hearst, Kyle Lo |
| 2025 | CVPR | Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models. | Matt Deitke, Christopher Clark, Sangho Lee, Rohun Tripathi, Yue Yang, Jae Sung Park, Mohammadreza Salehi, Niklas Muennighoff, Kyle Lo, Luca Soldaini, Jiasen Lu, Taira Anderson, Erin Bransom, Kiana Ehsani, Huong Ngo, Yen-Sung Chen, Ajay Patel, Mark Yatskar, Chris Callison-Burch, Andrew Head, Rose Hendrix, Favyen Bastani, Eli VanderBilt, Nathan Lambert, Yvonne Chou, Arnavi Chheda, Jenna Sparks, Sam Skjonsberg, Michael Schmitz, Aaron Sarnat, Byron Bischoff, Pete Walsh, Chris Newell, Piper Wolters, Tanmay Gupta, Kuo-Hao Zeng, Jon Borchardt, Dirk Groeneveld, Crystal Nam, Sophie Lebrecht, Caitlin Wittlif, Carissa Schoenick, Oscar Michel, Ranjay Krishna, Luca Weihs, Noah A. Smith, Hannaneh Hajishirzi, Ross B. Girshick, Ali Farhadi, Aniruddha Kembhavi |
| 2025 | EMNLP | Intent-aware Schema Generation and Refinement for Literature Review Tables. | Vishakh Padmakumar, Joseph Chee Chang, Kyle Lo, Doug Downey, Aakanksha Naik |
| 2025 | EMNLP | SciRIFF: A Resource to Enhance Language Model Instruction-Following over Scientific Literature. | David Wadden, Kejian Shi, Jacob Morrison, Alan Li, Aakanksha Naik, Shruti Singh, Nitzan Barzilay, Kyle Lo, Tom Hope, Luca Soldaini, Shannon Zejiang Shen, Doug Downey, Hannaneh Hajishirzi, Arman Cohan |
| 2025 | ICLR | OLMoE: Open Mixture-of-Experts Language Models. | Niklas Muennighoff, Luca Soldaini, Dirk Groeneveld, Kyle Lo, Jacob Morrison, Sewon Min, Weijia Shi, Evan Pete Walsh, Oyvind Tafjord, Nathan Lambert, Yuling Gu, Shane Arora, Akshita Bhagia, Dustin Schwenk, David Wadden, Alexander Wettig, Binyuan Hui, Tim Dettmers, Douwe Kiela, Ali Farhadi, et al. |
| 2025 | ICML | Organize the Web: Constructing Domains Enhances Pre-Training Data Curation. | Alexander Wettig, Kyle Lo, Sewon Min, Hannaneh Hajishirzi, Danqi Chen, Luca Soldaini |
| 2025 | NAACL | DrawEduMath: Evaluating Vision Language Models with Expert-Annotated Students' Hand-Drawn Math Images. | Sami Baral, Li Lucy, Ryan Knight, Alice Ng, Luca Soldaini, Neil T. Heffernan, Kyle Lo |
| 2025 | NAACL | FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions. | Orion Weller, Benjamin Chang, Sean MacAvaney, Kyle Lo, Arman Cohan, Benjamin Van Durme, Dawn J. Lawrie, Luca Soldaini |
| 2024 | ACL | OLMo: Accelerating the Science of Language Models. | Dirk Groeneveld, Iz Beltagy, Evan Pete Walsh, Akshita Bhagia, Rodney Kinney, Oyvind Tafjord, Ananya Harsh Jha, Hamish Ivison, Ian Magnusson, Yizhong Wang, Shane Arora, David Atkinson, Russell Authur, Khyathi Raghavi Chandu, Arman Cohan, Jennifer Dumas, Yanai Elazar, Yuling Gu, Jack Hessel, Tushar Khot, William Merrill, Jacob Morrison, Niklas Muennighoff, Aakanksha Naik, Crystal Nam, Matthew E. Peters, Valentina Pyatkin, Abhilasha Ravichander, Dustin Schwenk, Saurabh Shah, Will Smith, Emma Strubell, Nishant Subramani, Mitchell Wortsman, Pradeep Dasigi, Nathan Lambert, Kyle Richardson, Luke Zettlemoyer, Jesse Dodge, Kyle Lo, Luca Soldaini, Noah A. Smith, Hannaneh Hajishirzi |
| 2024 | ACL | Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research. | Luca Soldaini, Rodney Kinney, Akshita Bhagia, Dustin Schwenk, David Atkinson, Russell Authur, Ben Bogin, Khyathi Raghavi Chandu, Jennifer Dumas, Yanai Elazar, Valentin Hofmann, Ananya Harsh Jha, Sachin Kumar, Li Lucy, Xinxi Lyu, Nathan Lambert, Ian Magnusson, Jacob Morrison, Niklas Muennighoff, Aakanksha Naik, Crystal Nam, Matthew E. Peters, Abhilasha Ravichander, Kyle Richardson, Zejiang Shen, Emma Strubell, Nishant Subramani, Oyvind Tafjord, Pete Walsh, Luke Zettlemoyer, Noah A. Smith, Hannaneh Hajishirzi, Iz Beltagy, Dirk Groeneveld, Jesse Dodge, Kyle Lo |
| 2024 | ACL | InfoLossQA: Characterizing and Recovering Information Loss in Text Simplification. | Jan Trienes, Sebastian Joseph, Jrg Schltterer, Christin Seifert, Kyle Lo, Wei Xu, Byron C. Wallace, Junyi Jessy Li |
| 2024 | ACL | KIWI: A Dataset of Knowledge-Intensive Writing Instructions for Answering Research Questions. | Fangyuan Xu, Kyle Lo, Luca Soldaini, Bailey Kuehl, Eunsol Choi, David Wadden |
| 2024 | CHI | Know Your Audience: The benefits and pitfalls of generating plain language summaries beyond the "general" audience. | Tal August, Kyle Lo, Noah A. Smith, Katharina Reinecke |
| 2024 | EACL | When do Generative Query and Document Expansions Fail? A Comprehensive Study Across Methods, Retrievers, and Datasets. | Orion Weller, Kyle Lo, David Wadden, Dawn J. Lawrie, Benjamin Van Durme, Arman Cohan, Luca Soldaini |
| 2024 | EMNLP | One Thousand and One Pairs: A "novel" challenge for long-context language models. | Marzena Karpinska, Katherine Thai, Kyle Lo, Tanya Goyal, Mohit Iyyer |
| 2024 | EMNLP | MathFish: Evaluating Language Model Math Reasoning via Grounding in Educational Curricula. | Li Lucy, Tal August, Rose E. Wang, Luca Soldaini, Courtney Allison, Kyle Lo |
| 2024 | EMNLP | ArxivDIGESTables: Synthesizing Scientific Literature into Tables using Language Models. | Benjamin Newman, Yoonjoo Lee, Aakanksha Naik, Pao Siangliulue, Raymond Fok, Juho Kim, Daniel S. Weld, Joseph Chee Chang, Kyle Lo |
| 2024 | ICLR | BooookScore: A systematic exploration of book-length summarization in the era of LLMs. | Yapei Chang, Kyle Lo, Tanya Goyal, Mohit Iyyer |
| 2023 | ACL | Are Layout-Infused Language Models Robust to Layout Distribution Shifts? A Case Study with Scientific Documents. | Catherine Chen, Zejiang Shen, Dan Klein, Gabriel Stanovsky, Doug Downey, Kyle Lo |
| 2023 | CHI | CiteSee: Augmenting Citations in Scientific Papers with Persistent and Personalized Historical Context. | Joseph Chee Chang, Amy X. Zhang, Jonathan Bragg, Andrew Head, Kyle Lo, Doug Downey, Daniel S. Weld |
| 2023 | EACL | LongEval: Guidelines for Human Evaluation of Faithfulness in Long-form Summarization. | Kalpesh Krishna, Erin Bransom, Bailey Kuehl, Mohit Iyyer, Pradeep Dasigi, Arman Cohan, Kyle Lo |
| 2023 | EMNLP | Open Domain Multi-document Summarization: A Comprehensive Study of Model Brittleness under Retrieval. | John M. Giorgi, Luca Soldaini, Bo Wang, Gary D. Bader, Kyle Lo, Lucy Lu Wang, Arman Cohan |
| 2023 | EMNLP | Decomposing Complex Queries for Tip-of-the-tongue Retrieval. | Kevin Lin, Kyle Lo, Joseph Gonzalez, Dan Klein |
| 2023 | EMNLP | PaperMage: A Unified Toolkit for Processing, Representing, and Manipulating Visually-Rich Scientific Documents. | Kyle Lo, Zejiang Shen, Benjamin Newman, Joseph Chee Chang, Russell Authur, Erin Bransom, Stefan Candra, Yoganand Chandrasekhar, Regan Huff, Bailey Kuehl, Amanpreet Singh, Chris Wilhelm, Angele Zamarron, Marti A. Hearst, Daniel S. Weld, Doug Downey, Luca Soldaini |
| 2023 | EMNLP | A Question Answering Framework for Decontextualizing User-facing Snippets from Scientific Documents. | Benjamin Newman, Luca Soldaini, Raymond Fok, Arman Cohan, Kyle Lo |
| 2023 | IUI | Scim: Intelligent Skimming Support for Scientific Papers. | Raymond Fok, Hita Kambhamettu, Luca Soldaini, Jonathan Bragg, Kyle Lo, Marti A. Hearst, Andrew Head, Daniel S. Weld |
| 2022 | ACL | Generating Scientific Claims for Zero-Shot Scientific Fact Checking. | Dustin Wright, David Wadden, Kyle Lo, Bailey Kuehl, Arman Cohan, Isabelle Augenstein, Lucy Lu Wang |
| 2022 | CHI | Exploring the Role of Local and Global Explanations in Recommender Systems. | Marissa Radensky, Doug Downey, Kyle Lo, Zoran Popovic, Daniel S. Weld |
| 2022 | COLING | Overview of the Third Workshop on Scholarly Document Processing. | Arman Cohan, Guy Feigenblat, Dayne Freitag, Tirthankar Ghosal, Drahomira Herrmannova, Petr Knoth, Kyle Lo, Philipp Mayr, Michal Shmueli-Scheuer, Anita de Waard, Lucy Lu Wang |
| 2022 | EMNLP | ACCoRD: A Multi-Document Approach to Generating Diverse Descriptions of Scientific Concepts. | Sonia K. Murthy, Kyle Lo, Daniel King, Chandra Bhagavatula, Bailey Kuehl, Sophie Johnson, Jonathan Borchardt, Daniel S. Weld, Tom Hope, Doug Downey |
| 2022 | EMNLP | SciFact-Open: Towards open-domain scientific claim verification. | David Wadden, Kyle Lo, Bailey Kuehl, Arman Cohan, Iz Beltagy, Lucy Lu Wang, Hannaneh Hajishirzi |
| 2022 | NAACL | MultiCite: Modeling realistic citations requires moving beyond the single-sentence single-label setting. | Anne Lauscher, Brandon Ko, Bailey Kuehl, Sophie Johnson, Arman Cohan, David Jurgens, Kyle Lo |
| 2022 | NAACL | MultiVerS: Improving scientific claim verification with weak supervision and full-document context. | David Wadden, Kyle Lo, Lucy Lu Wang, Arman Cohan, Iz Beltagy, Hannaneh Hajishirzi |
| 2021 | ACL | Explaining Relationships Between Scientific Documents. | Kelvin Luu, Xinyi Wu, Rik Koncel-Kedziorski, Kyle Lo, Isabel Cachola, Noah A. Smith |
| 2021 | CHI | Augmenting Scientific Papers with Just-in-Time, Position-Sensitive Definitions of Terms and Symbols. | Andrew Head, Kyle Lo, Dongyeop Kang, Raymond Fok, Sam Skjonsberg, Daniel S. Weld, Marti A. Hearst |
| 2021 | EACL | Discourse Understanding and Factual Consistency in Abstractive Summarization. | Saadia Gabriel, Antoine Bosselut, Jeff Da, Ari Holtzman, Jan Buys, Kyle Lo, Asli Celikyilmaz, Yejin Choi |
| 2021 | NAACL | A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers. | Pradeep Dasigi, Kyle Lo, Iz Beltagy, Arman Cohan, Noah A. Smith, Matt Gardner |
| 2020 | ACL | Don't Stop Pretraining: Adapt Language Models to Domains and Tasks. | Suchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, Noah A. Smith |
| 2020 | ACL | S2ORC: The Semantic Scholar Open Research Corpus. | Kyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Kinney, Daniel S. Weld |
| 2020 | ECIR | The COVID-19 Open Research Dataset - Abstract. | Lucy Lu Wang, Kyle Lo |
| 2020 | EMNLP | TLDR: Extreme Summarization of Scientific Documents. | Isabel Cachola, Kyle Lo, Arman Cohan, Daniel S. Weld |
| 2020 | EMNLP | Document-Level Definition Detection in Scholarly Documents: Existing Models, Error Analyses, and Future Directions. | Dongyeop Kang, Andrew Head, Risham Sidhu, Kyle Lo, Daniel S. Weld, Marti A. Hearst |
| 2020 | EMNLP | Fact or Fiction: Verifying Scientific Claims. | David Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang, Madeleine van Zuylen, Arman Cohan, Hannaneh Hajishirzi |
| 2019 | EMNLP | SciBERT: A Pretrained Language Model for Scientific Text. | Iz Beltagy, Kyle Lo, Arman Cohan |
| 2019 | NAACL | Combining Distant and Direct Supervision for Neural Relation Extraction. | Iz Beltagy, Kyle Lo, Waleed Ammar |
| 2018 | NAACL | Construction of the Literature Graph in Semantic Scholar. | Waleed Ammar, Dirk Groeneveld, Chandra Bhagavatula, Iz Beltagy, Miles Crawford, Doug Downey, Jason Dunkelberger, Ahmed Elgohary, Sergey Feldman, Vu Ha, Rodney Kinney, Sebastian Kohlmeier, Kyle Lo, Tyler Murray, Hsu-Han Ooi, Matthew E. Peters, Joanna Power, Sam Skjonsberg, Lucy Lu Wang, Chris Wilhelm, Zheng Yuan, Madeleine van Zuylen, Oren Etzioni |