| 2026 | ACL | Don't Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination. | Prafulla Kumar Choubey, Kung-Hsiang Huang, Pranav Narayanan Venkit, Jiaxin Zhang, Vaibhav Vats, Yu Li, Xiangyu Peng, Chien-Sheng Wu |
| 2026 | ACL | GTA: Generating Long-horizon Tasks for Web Agents at Scale. | Tenghao Huang, Kung-Hsiang Huang, Prafulla Kumar Choubey, Yilun Zhou, Muhao Chen, Jonathan May, Chien-Sheng Wu |
| 2025 | ACL | Turning Conversations into Workflows: A Framework to Extract and Evaluate Dialog Workflows for Service AI Agents. | Prafulla Kumar Choubey, Xiangyu Peng, Shilpa Bhagavath, Caiming Xiong, Shiva Kumar Pentyala, Chien-Sheng Wu |
| 2025 | ACL | Unanswerability Evaluation for Retrieval Augmented Generation. | Xiangyu Peng, Prafulla Kumar Choubey, Caiming Xiong, Chien-Sheng Wu |
| 2025 | EMNLP | Benchmarking Deep Search over Heterogeneous Enterprise Data. | Prafulla Kumar Choubey, Xiangyu Peng, Shilpa Bhagavath, Kung-Hsiang Huang, Caiming Xiong, Chien-Sheng Wu |
| 2025 | ICLR | SiReRAG: Indexing Similar and Related Information for Multihop Reasoning. | Nan Zhang, Prafulla Kumar Choubey, Alexander R. Fabbri, Gabriel Bernadett-Shapiro, Rui Zhang, Prasenjit Mitra, Caiming Xiong, Chien-Sheng Wu |
| 2025 | NAACL | Do RAG Systems Cover What Matters? Evaluating and Optimizing Responses with Sub-Question Coverage. | Kaige Xie, Philippe Laban, Prafulla Kumar Choubey, Caiming Xiong, Chien-Sheng Wu |
| 2024 | NAACL | Embrace Divergence for Richer Insights: A Multi-document Summarization Benchmark and a Case Study on Summarizing Diverse Information from News Articles. | Kung-Hsiang Huang, Philippe Laban, Alexander R. Fabbri, Prafulla Kumar Choubey, Shafiq Joty, Caiming Xiong, Chien-Sheng Wu |
| 2023 | ACL | CaPE: Contrastive Parameter Ensembling for Reducing Hallucination in Abstractive Summarization. | Prafulla Kumar Choubey, Alexander R. Fabbri, Jesse Vig, Chien-Sheng Wu, Wenhao Liu, Nazneen Rajani |
| 2023 | EMNLP | Lexical Repetitions Lead to Rote Learning: Unveiling the Impact of Lexical Overlap in Train and Test Reference Summaries. | Prafulla Kumar Choubey, Alexander R. Fabbri, Caiming Xiong, Chien-Sheng Wu |
| 2023 | ICLR | Model ensemble instead of prompt fusion: a sample-specific knowledge transfer method for few-shot prompt tuning. | Xiangyu Peng, Chen Xing, Prafulla Kumar Choubey, Chien-Sheng Wu, Caiming Xiong |
| 2022 | ACL | Predicting Sentence Deletions for Text Simplification Using a Functional Discourse Structure. | Bohan Zhang, Prafulla Kumar Choubey, Ruihong Huang |
| 2022 | EMNLP | Conformal Predictor for Improving Zero-Shot Text Classification Efficiency. | Prafulla Kumar Choubey, Yu Bai, Chien-Sheng Wu, Wenhao Liu, Nazneen Rajani |
| 2022 | EMNLP | Improving Factual Consistency in Summarization with Compression-Based Post-Editing. | Alexander R. Fabbri, Prafulla Kumar Choubey, Jesse Vig, Chien-Sheng Wu, Caiming Xiong |
| 2022 | ICLR | P-Adapters: Robustly Extracting Factual Information from Language Models with Diverse Prompts. | Benjamin Newman, Prafulla Kumar Choubey, Nazneen Rajani |
| 2022 | IJCNLP | Modeling Document-level Temporal Structures for Building Temporal Dependency Graphs. | Prafulla Kumar Choubey, Ruihong Huang |
| 2021 | EACL | Automatic Data Acquisition for Event Coreference Resolution. | Prafulla Kumar Choubey, Ruihong Huang |
| 2021 | EMNLP | GFST: Gender-Filtered Self-Training for More Accurate Gender in Translation. | Prafulla Kumar Choubey, Anna Currey, Prashant Mathur, Georgiana Dinu |
| 2021 | EMNLP | Profiling News Discourse Structure Using Explicit Subtopic Structures Guided Critics. | Prafulla Kumar Choubey, Ruihong Huang |
| 2020 | ACL | Discourse as a Function of Event: Profiling Discourse Structure in News Articles around the Main Event. | Prafulla Kumar Choubey, Aaron Lee, Ruihong Huang, Lu Wang |
| 2020 | LREC | One Classifier for All Ambiguous Words: Overcoming Data Sparsity by Utilizing Sense Correlations Across Words. | Prafulla Kumar Choubey, Ruihong Huang |
| 2019 | EMNLP | In Plain Sight: Media Bias Through the Lens of Factual Reporting. | Lisa Fan, Marshall White, Eva Sharma, Ruisi Su, Prafulla Kumar Choubey, Ruihong Huang, Lu Wang |
| 2019 | NAACL | Modeling Document-level Causal Structures for Event Causal Relation Identification. | Lei Gao, Prafulla Kumar Choubey, Ruihong Huang |
| 2019 | NAACL | Improving Dialogue State Tracking by Discerning the Relevant Context. | Sanuj Sharma, Prafulla Kumar Choubey, Ruihong Huang |
| 2018 | ACL | Improving Event Coreference Resolution by Modeling Correlations between Event Coreference Chains and Document Topic Structures. | Prafulla Kumar Choubey, Ruihong Huang |
| 2018 | NAACL | Identifying the Most Dominant Event in a News Article by Mining Event Coreference Relations. | Prafulla Kumar Choubey, Kaushik Raju, Ruihong Huang |
| 2017 | EMNLP | A Sequential Model for Classifying Temporal Relations between Intra-Sentence Events. | Prafulla Kumar Choubey, Ruihong Huang |
| 2017 | EMNLP | Event Coreference Resolution by Iteratively Unfolding Inter-dependencies among Events. | Prafulla Kumar Choubey, Ruihong Huang |