| 2026 | AAAI | TFRank: Think-Free Reasoning Enables Practical Pointwise LLM Ranking. | Yongqi Fan, Xiaoyang Chen, Dezhi Ye, Jie Liu, Haijin Liang, Jin Ma, Ben He, Yingfei Sun, Tong Ruan |
| 2026 | ACL | CoCoNUTS: Concentrating on Content while Neglecting Uninformative Textual Styles for AI-Generated Peer Review Detection. | Yihan Chen, Jiawei Chen, Guozhao Mo, Xuanang Chen, Ben He, Xianpei Han, Le Sun |
| 2026 | ACL | SURE or Not? Investigating Semantic Understanding in Dense Retrieval Models. | Lingdi Kong, Xuanang Chen, Ben He, Le Sun |
| 2026 | ACL | All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG. | Dan Wang, Guozhao Mo, Yafei Shi, Cheng Zhang, Bo Zheng, Boxi Cao, Xuanang Chen, Yaojie Lu, Hongyu Lin, Ben He, Xianpei Han, Le Sun |
| 2026 | ACL | CodeRise: Bootstrapping LLMs for Ultra Low-Resource Programming Languages via Progressive Self-Refinement Curriculum. | Tengfei Wen, Xuanang Chen, Ben He, Xiaoliang Cong, Le Sun |
| 2026 | EACL | Expanding the Boundaries of Vision Prior Knowledge in Multi-modal Large Language Models. | Qiao Liang, Yanjiang Liu, Weixiang Zhou, Ben He, Yaojie Lu, Hongyu Lin, Jia Zheng, Xianpei Han, Le Sun, Yingfei Sun |
| 2026 | EACL | Navigating the Infinite Dynamic Web Space: Effective In-Context Exploration via Cognitive Multi-Agent Collaboration. | Guozhao Mo, Yanjiang Liu, Yafei Shi, Jiawei Chen, Yang Li, Yaojie Lu, Hongyu Lin, Ben He, Le Sun, Bo Zheng, Xianpei Han |
| 2025 | ACL | Not All Terms Matter: Recall-Oriented Adaptive Learning for PLM-aided Query Expansion in Open-Domain Question Answering. | Xinran Chen, Ben He, Xuanang Chen, Le Sun |
| 2025 | ACL | Code-SPA: Style Preference Alignment to Large Language Models for Effective and Robust Code Debugging. | Tengfei Wen, Xuanang Chen, Ben He, Le Sun |
| 2025 | ACL | Cheems: A Practical Guidance for Building and Evaluating Chinese Reward Models from Scratch. | Xueru Wen, Jie Lou, Zichao Li, Yaojie Lu, XingYu, Yuqiu Ji, Guohai Xu, Hongyu Lin, Ben He, Xianpei Han, Le Sun, Debing Zhang |
| 2025 | ACL | On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation. | Xueru Wen, Jie Lou, Xinyu Lu, Yuqiu Ji, Xinyan Guan, Yaojie Lu, Hongyu Lin, Ben He, Xianpei Han, Debing Zhang, Le Sun |
| 2025 | ACL | Self-Steering Optimization: Autonomous Preference Optimization for Large Language Models. | Hao Xiang, Bowen Yu, Hongyu Lin, Keming Lu, Yaojie Lu, Xianpei Han, Ben He, Le Sun, Jingren Zhou, Junyang Lin |
| 2025 | ACL | CRUXEVAL-X: A Benchmark for Multilingual Code Reasoning, Understanding and Execution. | Ruiyang Xu, Jialun Cao, Yaojie Lu, Ming Wen, Hongyu Lin, Xianpei Han, Ben He, Shing-Chi Cheung, Le Sun |
| 2025 | ACL | Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning. | Ruoxi Xu, Yunjie Ji, Boxi Cao, Yaojie Lu, Hongyu Lin, Xianpei Han, Ben He, Yingfei Sun, Xiangang Li, Le Sun |
| 2025 | ACL | DiffLM: Controllable Synthetic Data Generation via Diffusion Language Models. | Ying Zhou, Xinyao Wang, Yulei Niu, Yaojie Shen, Lexin Tang, Fan Chen, Ben He, Le Sun, Longyin Wen |
| 2025 | COLING | Can LLMs Clarify? Investigation and Enhancement of Large Language Models on Argument Claim Optimization. | Yiran Wang, Ben He, Xuanang Chen, Le Sun |
| 2025 | EMNLP | ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch. | Jiawei Chen, Xinyan Guan, Qianhao Yuan, Guozhao Mo, Weixiang Zhou, Yaojie Lu, Hongyu Lin, Ben He, Le Sun, Xianpei Han |
| 2025 | ICLR | Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree? | Xueru Wen, Jie Lou, Yaojie Lu, Hongyu Lin, XingYu, Xinyu Lu, Ben He, Xianpei Han, Debing Zhang, Le Sun |
| 2025 | KDD | Multi-Agent Proactive Information Seeking with Adaptive LLM Orchestration for Non-Factoid Question Answering. | Xinran Chen, Yuchen Li, Hengyi Cai, Zhuoran Ma, Xuanang Chen, Haoyi Xiong, Shuaiqiang Wang, Ben He, Le Sun, Dawei Yin |
| 2024 | AAAI | Mitigating Large Language Model Hallucinations via Autonomous Knowledge Graph-Based Retrofitting. | Xinyan Guan, Yanjiang Liu, Hongyu Lin, Yaojie Lu, Ben He, Xianpei Han, Le Sun |
| 2024 | ACL | Rule or Story, Which is a Better Commonsense Expression for Talking with Large Language Models? | Ning Bian, Xianpei Han, Hongyu Lin, Yaojie Lu, Ben He, Le Sun |
| 2024 | ACL | Analyze, Generate and Refine: Query Expansion with LLMs for Zero-Shot Open-Domain QA. | Xinran Chen, Xuanang Chen, Ben He, Tengfei Wen, Le Sun |
| 2024 | ACL | Spiral of Silence: How is Large Language Model Killing Information Retrieval? - A Case Study on Open Domain Question Answering. | Xiaoyang Chen, Ben He, Hongyu Lin, Xianpei Han, Tianshu Wang, Boxi Cao, Le Sun, Yingfei Sun |
| 2024 | ACL | XMC-Agent : Dynamic Navigation over Scalable Hierarchical Index for Incremental Extreme Multi-label Classification. | Yanjiang Liu, Tianyun Zhong, Yaojie Lu, Hongyu Lin, Ben He, Shuheng Zhou, Huijia Zhu, Weiqiang Wang, Zhongyi Liu, Xianpei Han, Le Sun |
| 2024 | ACL | PRP-Graph: Pairwise Ranking Prompting to LLMs with Graph Aggregation for Effective Text Re-ranking. | Jian Luo, Xuanang Chen, Ben He, Le Sun |
| 2024 | ACL | Navigating the Shadows: Unveiling Effective Disturbances for Modern AI Content Detectors. | Ying Zhou, Ben He, Le Sun |
| 2024 | COLING | ChatGPT Is a Knowledgeable but Inexperienced Solver: An Investigation of Commonsense Problem in Large Language Models. | Ning Bian, Xianpei Han, Le Sun, Hongyu Lin, Yaojie Lu, Ben He, Shanshan Jiang, Bin Dong |
| 2024 | COLING | Humanizing Machine-Generated Content: Evading AI-Text Detection through Adversarial Attack. | Ying Zhou, Ben He, Le Sun |
| 2023 | ACL | Towards Imperceptible Document Manipulations against Neural Ranking Models. | Xuanang Chen, Ben He, Zheng Ye, Le Sun, Yingfei Sun |
| 2023 | ACL | Understanding Differential Search Index for Text Retrieval. | Xiaoyang Chen, Yanjiang Liu, Ben He, Le Sun, Yingfei Sun |
| 2023 | EMNLP | Contrastive Distant Supervision for Debiased and Denoised Machine Reading Comprehension. | Ning Bian, Hongyu Lin, Xianpei Han, Ben He, Le Sun |
| 2023 | EMNLP | Hidding the Ghostwriters: An Adversarial Evaluation of AI-Generated Student Essay Detection. | Xinlin Peng, Ying Zhou, Ben He, Le Sun, Yingfei Sun |
| 2023 | EMNLP | Contextual Interaction for Argument Post Quality Assessment. | Yiran Wang, Xuanang Chen, Ben He, Le Sun |
| 2023 | SIGIR | Offline Pseudo Relevance Feedback for Efficient and Effective Single-pass Dense Retrieval. | Xueru Wen, Xiaoyang Chen, Xuanang Chen, Ben He, Le Sun |
| 2022 | ECIR | Incorporating Ranking Context for End-to-End BERT Re-ranking. | Xiaoyang Chen, Kai Hui, Ben He, Xianpei Han, Le Sun, Zheng Ye |
| 2022 | ECIR | Groupwise Query Performance Prediction with BERT. | Xiaoyang Chen, Ben He, Le Sun |
| 2022 | IJCAI | Towards Robust Dense Retrieval via Local Ranking Alignment. | Xuanang Chen, Jian Luo, Ben He, Le Sun, Yingfei Sun |
| 2022 | SIGIR | Re-thinking Knowledge Graph Completion Evaluation from an Information Retrieval Perspective. | Ying Zhou, Xuanang Chen, Ben He, Zheng Ye, Le Sun |
| 2021 | ECIR | Simplified TinyBERT: Knowledge Distillation for Document Retrieval. | Xuanang Chen, Ben He, Kai Hui, Le Sun, Yingfei Sun |
| 2021 | SIGIR | Contextualized Offline Relevance Weighting for Efficient and Effective Neural Retrieval. | Xuanang Chen, Ben He, Kai Hui, Yiran Wang, Le Sun, Yingfei Sun |
| 2020 | AAAI | Learning to Map Frequent Phrases to Sub-Structures of Meaning Representation for Neural Semantic Parsing. | Bo Chen, Xianpei Han, Ben He, Le Sun |
| 2020 | AAAI | End-to-End Bootstrapping Neural Network for Entity Set Expansion. | Lingyong Yan, Xianpei Han, Ben He, Le Sun |
| 2020 | EMNLP | Global Bootstrapping Neural Network for Entity Set Expansion. | Lingyong Yan, Xianpei Han, Ben He, Le Sun |
| 2020 | EMNLP | BERT-QE: Contextualized Query Expansion for Document Re-ranking. | Zhi Zheng, Kai Hui, Ben He, Xianpei Han, Le Sun, Andrew Yates |
| 2020 | KSEM | End-to-End Multi-task Learning for Allusion Detection in Ancient Chinese Poems. | Lei Liu, Xiaoyang Chen, Ben He |
| 2019 | CIKM | Deep Sequence-to-Sequence Entity Matching for Heterogeneous Entity Resolution. | Hao Nie, Xianpei Han, Ben He, Le Sun, Bo Chen, Wei Zhang, Suhui Wu, Hao Kong |
| 2019 | EMNLP | Learning to Bootstrap for Entity Set Expansion. | Lingyong Yan, Xianpei Han, Le Sun, Ben He |
| 2018 | ACL | TDNN: A Two-stage Deep Neural Network for Prompt-independent Automated Essay Scoring. | Cancan Jin, Ben He, Kai Hui, Le Sun |
| 2018 | EMNLP | NPRF: A Neural Pseudo Relevance Feedback Framework for Ad-hoc Information Retrieval. | Canjia Li, Yingfei Sun, Ben He, Le Wang, Kai Hui, Andrew Yates, Le Sun, Jungang Xu |
| 2018 | ICTAI | Long-Term Recurrent Merge Network Model for Image Captioning. | Yang Fan, Jungang Xu, Yingfei Sun, Ben He |
| 2018 | KSEM | A Study on Performance Sensitivity to Data Sparsity for Automated Essay Scoring. | Yanhua Ran, Ben He, Jungang Xu |
| 2017 | APWEB | Integrating Feedback-Based Semantic Evidence to Enhance Retrieval Effectiveness for Clinical Decision Support. | Chenhao Yang, Ben He, Jungang Xu |
| 2017 | ICANN | An Improved Convolutional Neural Network for Sentence Classification Based on Term Frequency and Segmentation. | Qi Wang, Jungang Xu, Ben He, Zhengcai Qin |
| 2017 | IJCNN | Two improved continuous bag-of-word models. | Qi Wang, Jungang Xu, Hong Chen, Ben He |
| 2017 | KSEM | A Study of Distributed Semantic Representations for Automated Essay Scoring. | Cancan Jin, Ben He, Jungang Xu |
| 2016 | HPCC | A Novel Method for Tuning Configuration Parameters of Spark Based on Machine Learning. | Guolu Wang, Jungang Xu, Ben He |
| 2016 | KSEM | A Document Modeling Method Based on Deep Generative Model and Spectral Hashing. | Hong Chen, Jungang Xu, Qi Wang, Ben He |
| 2016 | SAC | Direct measurement of training query quality for learning to rank. | Qingli Ma, Ben He, Jungang Xu |
| 2015 | ECIR | Selecting Training Data for Learning-Based Twitter Search. | Dongxing Li, Ben He, Tiejian Luo, Xin Zhang |
| 2015 | UIC | Utilizing Latent Semantic Word Representations for Automated Essay Scoring. | Cancan Jin, Ben He |
| 2014 | SMC | A model of Demand Response scheduling for cement plant. | XinZhang Zhao, Ben He, Fang-Yuan Xu, Loi Lei Lai, Chenhao Yang, Sha Lu, Dongxing Li |
| 2013 | CIKM | Clustering-based transduction for learning a ranking model with limited human labels. | Xin Zhang, Ben He, Tiejian Luo, Dongxing Li, Jungang Xu |
| 2013 | ECIR | Sponsored Search Ad Selection by Keyword Structure Analysis. | Kai Hui, Bin Gao, Ben He, Tiejian Luo |
| 2013 | EMNLP | Automated Essay Scoring by Maximizing Human-Machine Agreement. | Hongbo Chen, Ben He |
| 2013 | PAKDD | Learn to Rank Tweets by Integrating Query-Specific Characteristics. | Xin Zhang, Ben He, Tiejian Luo |
| 2012 | CIKM | Question-answer topic model for question retrieval in community question answering. | Zongcheng Ji, Fei Xu, Bin Wang, Ben He |
| 2012 | CIKM | Query-biased learning to rank for real-time twitter search. | Xin Zhang, Ben He, Tiejian Luo, Baobin Li |
| 2012 | ICWSM | Transductive Learning for Real-Time Twitter Search. | Xin Zhang, Ben He, Tiejian Luo |
| 2011 | CIKM | Relevance weighting using within-document term statistics. | Kai Hui, Ben He, Tiejian Luo, Bin Wang |
| 2011 | CIKM | Exploring categorization property of social annotations for information retrieval. | Peng Li, Bin Wang, Wei Jin, Jian-Yun Nie, Zhiwei Shi, Ben He |
| 2011 | ICTIR | A Comparative Study of Pseudo Relevance Feedback for Ad-hoc Retrieval. | Kai Hui, Ben He, Tiejian Luo, Bin Wang |
| 2011 | SIGIR | CRTER: using cross terms to enhance probabilistic information retrieval. | Jiashu Zhao, Jimmy Xiangji Huang, Ben He |
| 2011 | SIGIR | Enhancing ad-hoc relevance weighting using probability density estimation. | Xiaofeng Zhou, Jimmy Xiangji Huang, Ben He |
| 2009 | CIKM | Finding good feedback documents. | Ben He, Iadh Ounis |
| 2009 | CIKM | A study of selective collection enrichment for enterprise search. | Jie Peng, Craig Macdonald, Ben He, Iadh Ounis |
| 2009 | ECIR | Studying Query Expansion Effectiveness. | Ben He, Iadh Ounis |
| 2009 | ECIR | Integrating Proximity to Subjective Sentences for Blog Opinion Retrieval. | Rodrygo L. T. Santos, Ben He, Craig Macdonald, Iadh Ounis |
| 2009 | ICTIR | Predicting the Usefulness of Collection Enrichment for Enterprise Search. | Jie Peng, Ben He, Iadh Ounis |
| 2009 | SIGIR | Fitting score distribution for blog opinion retrieval. | Ben He, Jie Peng, Iadh Ounis |
| 2008 | CIKM | An effective statistical approach to blog post opinion retrieval. | Ben He, Craig Macdonald, Jiyin He, Iadh Ounis |
| 2008 | SIGIR | Retrieval sensitivity under training using different measures. | Ben He, Craig Macdonald, Iadh Ounis |
| 2008 | SIGIR | Ranking opinionated blog posts using OpinionFinder. | Ben He, Craig Macdonald, Iadh Ounis |
| 2008 | SIGIR | Limits of opinion-finding baseline systems. | Craig Macdonald, Ben He, Iadh Ounis, Ian Soboroff |
| 2007 | CIKM | Parameter sensitivity in the probabilistic model for ad-hoc retrieval. | Ben He, Iadh Ounis |
| 2007 | ECIR | Setting Per-field Normalisation Hyper-parameters for the Named-Page Finding Search Task. | Ben He, Iadh Ounis |
| 2007 | SIGIR | Incorporating term dependency in the dfr framework. | Jie Peng, Craig Macdonald, Ben He, Vassilis Plachouras, Iadh Ounis |
| 2005 | ECIR | Term Frequency Normalisation Tuning for BM25 and DFR Models. | Ben He, Iadh Ounis |
| 2005 | ECIR | Terrier Information Retrieval Platform. | Iadh Ounis, Gianni Amati, Vassilis Plachouras, Ben He, Craig Macdonald, Douglas Johnson |
| 2005 | SIGIR | A study of the dirichlet priors for term frequency normalisation. | Ben He, Iadh Ounis |
| 2004 | SPIRE | Inferring Query Performance Using Pre-retrieval Predictors.. | Ben He, Iadh Ounis |
| 2003 | CIKM | A study of parameter tuning for term frequency normalization. | Ben He, Iadh Ounis |