| 2026 | AAAI | HEV Generative Sandbox: A Framework for Assessing Domain-Specific Social Risks Through Human-LLM Simulation. | Yiran Liu, Zhiyi Hou, Xiaoang Xu, Shuo Wang, Huijia Wu, Kaicheng Yu, Yang Yu, ChengXiang Zhai |
| 2026 | ACL | BiasGRPO: Stabilizing Bias Mitigation in High-Variance Reward Landscapes via Group-Relative Policy Optimization. | Saket Reddy, Ke Yang, ChengXiang Zhai |
| 2025 | ACL | The Law of Knowledge Overshadowing: Towards Understanding, Predicting and Preventing LLM Hallucination. | Yuji Zhang, Sha Li, Cheng Qian, Jiateng Liu, Pengfei Yu, Chi Han, Yi R. Fung, Kathleen McKeown, ChengXiang Zhai, Manling Li, Heng Ji |
| 2025 | ACL | ORBIT: Cost-Effective Dataset Curation for Large Language Model Domain Adaptation with an Astronomy Case Study. | Eric Modesitt, Ke Yang, Spencer Hulsey, Xin Liu, ChengXiang Zhai, Volodymyr V. Kindratenko |
| 2025 | ACL | Beyond Reactive Safety: Risk-Aware LLM Alignment via Long-Horizon Simulation. | Chenkai Sun, Denghui Zhang, ChengXiang Zhai, Heng Ji |
| 2025 | COLING | Persona-DB: Efficient Large Language Model Personalization for Response Prediction with Collaborative Data Refinement. | Chenkai Sun, Ke Yang, Revanth Gangi Reddy, Yi Ren Fung, Hou Pong Chan, Kevin Small, ChengXiang Zhai, Heng Ji |
| 2025 | EMNLP | ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges. | Cheng Qian, Hongyi Du, Hongru Wang, Xiusi Chen, Yuji Zhang, Avirup Sil, ChengXiang Zhai, Kathleen McKeown, Heng Ji |
| 2025 | EMNLP | Cache-of-Thought: Master-Apprentice Framework for Cost-Effective Vision Language Model Reasoning. | Mingyuan Wu, Jize Jiang, Haozhen Zheng, Meitang Li, Zhaoheng Li, Beitong Tian, Bo Chen, Yongjoo Park, Minjia Zhang, ChengXiang Zhai, Klara Nahrstedt |
| 2025 | ICML | What Makes In-context Learning Effective for Mathematical Reasoning. | Jiayu Liu, Zhenya Huang, Chaokun Wang, Xunpeng Huang, ChengXiang Zhai, Enhong Chen |
| 2025 | KDD | Learning to Slice: Self-Supervised Interpretable Hierarchical Representation Learning with Graph Auto-Encoder Tree. | Jinning Li, Ruipeng Han, Jingying Zeng, Dachun Sun, Chenkai Sun, Hanghang Tong, ChengXiang Zhai, Boleslaw K. Szymanski, Tarek F. Abdelzaher |
| 2025 | SIGIR | TINK: Text Information Navigation Kit. | Dean E. Alvarez, ChengXiang Zhai |
| 2025 | SIGIR | Theory and Toolkits for User Simulation in the Era of Generative AI: User Modeling, Synthetic Data Generation, and System Evaluation. | Krisztian Balog, Nolwenn Bernard, Saber Zerhoudi, ChengXiang Zhai |
| 2025 | SIGIR | InstInfo: A Just-in-Time Literature Recommendation System for Presentations. | Kevin Ros, Rahul Suresh, ChengXiang Zhai |
| 2025 | SIGIR | Information Retrieval for Artificial General Intelligence: A New Perspective of Information Retrieval Research. | ChengXiang Zhai |
| 2024 | AAAI | Seed-Guided Fine-Grained Entity Typing in Science and Engineering Domains. | Yu Zhang, Yunyi Zhang, Yanzhen Shen, Yu Deng, Lucian Popa, Larisa Shwartz, ChengXiang Zhai, Jiawei Han |
| 2024 | EACL | AnaDE1.0: A Novel Data Set for Benchmarking Analogy Detection and Extraction. | Bhavya, Shradha Sehgal, Jinjun Xiong, ChengXiang Zhai |
| 2024 | EDM | Exploring AI-powered Multimodal Analogies for Science Education. | Shradha Sehgal, Bhavya, Krishna Phani Datta, Aditi Mallavarapu, ChengXiang Zhai |
| 2024 | EMNLP | UOUO: Uncontextualized Uncommon Objects for Measuring Knowledge Horizons of Vision Language Models. | Xinyu Pi, Mingyuan Wu, Jize Jiang, Haozhen Zheng, Beitong Tian, ChengXiang Zhai, Klara Nahrstedt, Zhiting Hu |
| 2024 | WWW | Tutorial on User Simulation for Evaluating Information Access Systems on the Web. | Krisztian Balog, ChengXiang Zhai |
| 2024 | SIGCSE | Scaling Collaborative Learning: Using the Community Digital Library to Enrich Course Content. | Kevin Ros, ChengXiang Zhai |
| 2024 | SIGIR | TextData: Save What You Know and Find What You Don't. | Kevin Ros, Kedar Takwane, Ashwin Patil, Rakshana Jayaprakash, ChengXiang Zhai |
| 2024 | SIGIR | Large Language Models and Future of Information Retrieval: Opportunities and Challenges. | ChengXiang Zhai |
| 2024 | WSDM | CharmBana: Progressive Responses with Real-Time Internet Search for Knowledge-Powered Conversations. | Revanth Gangi Reddy, Sharath Chandra Etagi Suresh, Hao Bai, Wentao Yao, Mankeerat Sidhu, Karan Aggarwal, Prathamesh Sonawane, ChengXiang Zhai |
| 2023 | AAAI | Learning by Applying: A General Framework for Mathematical Reasoning via Enhancing Explicit Knowledge Learning. | Jiayu Liu, Zhenya Huang, ChengXiang Zhai, Qi Liu |
| 2023 | ACL | When to Use What: An In-Depth Comparative Empirical Analysis of OpenIE Systems for Downstream Applications. | Kevin Pei, Ishan Jindal, Kevin Chen-Chuan Chang, ChengXiang Zhai, Yunyao Li |
| 2023 | ACL | Measuring the Effect of Influential Messages on Varying Personas. | Chenkai Sun, Jinning Li, Hou Pong Chan, ChengXiang Zhai, Heng Ji |
| 2023 | CIKM | Tutorial on User Simulation for Evaluating Information Access Systems. | Krisztian Balog, ChengXiang Zhai |
| 2023 | CSCW | The CDL: An Online Platform for Creating Community-based Digital Libraries. | Kevin Ros, ChengXiang Zhai |
| 2023 | EMNLP | Social Commonsense-Guided Search Query Generation for Open-Domain Knowledge-Powered Conversations. | Revanth Gangi Reddy, Hao Bai, Wentao Yao, Sharath Chandra Etagi Suresh, Heng Ji, ChengXiang Zhai |
| 2023 | EMNLP | Decoding the Silent Majority: Inducing Belief Augmented Social Graph with Large Language Model for Response Forecasting. | Chenkai Sun, Jinning Li, Yi Ren Fung, Hou Pong Chan, Tarek F. Abdelzaher, ChengXiang Zhai, Heng Ji |
| 2023 | ICDM | Exploring Large Language Models for Low-Resource IT Information Extraction. | Bhavya, Paulina Toro Isaza, Yu Deng, Michael Nidd, Amar Prakash Azad, Larisa Shwartz, ChengXiang Zhai |
| 2023 | ICDM | An Exploration of Large Language Models for Verification of News Headlines. | Yifan Li, ChengXiang Zhai |
| 2023 | ICTIR | Retrieving Webpages Using Online Discussions. | Kevin Ros, Matthew Jin, Jacob Levine, ChengXiang Zhai |
| 2022 | ACL | Domain Representative Keywords Selection: A Probabilistic Approach. | Pritom Saha Akash, Jie Huang, Kevin Chen-Chuan Chang, Yunyao Li, Lucian Popa, ChengXiang Zhai |
| 2022 | ACL | Improving Candidate Retrieval with Entity Profile Generation for Wikidata Entity Linking. | Tuan Manh Lai, Heng Ji, ChengXiang Zhai |
| 2022 | COLING | CONCRETE: Improving Cross-lingual Fact-checking with Cross-lingual Retrieval. | Kung-Hsiang Huang, ChengXiang Zhai, Heng Ji |
| 2022 | ECIR | RATE: A Reliability-Aware Tester-Based Evaluation Framework of User Simulators. | Sahiti Labhishetty, ChengXiang Zhai |
| 2022 | EMNLP | Language Model Pre-Training with Sparse Latent Typing. | Liliang Ren, Zixuan Zhang, Han Wang, Clare R. Voss, ChengXiang Zhai, Heng Ji |
| 2022 | HCI | Fine Grained Categorization of Drug Usage Tweets. | Priyanka Dey, ChengXiang Zhai |
| 2022 | ICTIR | PRE: A Precision-Recall-Effort Optimization Framework for Query Simulation. | Sahiti Labhishetty, ChengXiang Zhai |
| 2022 | ICTIR | I3A: An Intelligent Interactive Information Agent Model for Information Retrieval. | ChengXiang Zhai |
| 2022 | ICWSM | Drink Bleach or Do What Now? COVID-HeRA: A Study of Risk-Informed Health Decision Making in the Presence of COVID-19 Misinformation. | Arkin Dharawat, Ismini Lourentzou, Alex Morales, ChengXiang Zhai |
| 2022 | WSDM | Differential Query Semantic Analysis: Discovery of Explicit Interpretable Knowledge from E-Com Search Logs. | Sahiti Labhishetty, ChengXiang Zhai, Min Xie, Lin Gong, Rahul Sharnagat, Satya Chembolu |
| 2021 | ACL | Joint Biomedical Entity and Relation Extraction with Knowledge-Enhanced Collective Inference. | Tuan Manh Lai, Heng Ji, ChengXiang Zhai, Quan Hung Tran |
| 2021 | ECIR | A Study of Distributed Representations for Figures of Research Articles. | Saar Kuzi, ChengXiang Zhai |
| 2021 | ECIR | Towards Dark Jargon Interpretation in Underground Forums. | Dominic Seyler, Wei Liu, XiaoFeng Wang, ChengXiang Zhai |
| 2021 | EMNLP | Text2Mol: Cross-Modal Molecule Retrieval with Natural Language Queries. | Carl Edwards, ChengXiang Zhai, Heng Ji |
| 2021 | EMNLP | BERT might be Overkill: A Tiny but Effective Biomedical Entity Linker based on Residual Convolutional Neural Networks. | Tuan Manh Lai, Heng Ji, ChengXiang Zhai |
| 2021 | KDD | Neural-Answering Logical Queries on Knowledge Graphs. | Lihui Liu, Boxin Du, Heng Ji, ChengXiang Zhai, Hanghang Tong |
| 2021 | SIGIR | DarkJargon.net: A Platform for Understanding Underground Conversation with Latent Meaning. | Dominic Seyler, Wei Liu, Yunan Zhang, XiaoFeng Wang, ChengXiang Zhai |
| 2020 | AAAI | Predicting Opioid Overdose Crude Rates with Text-Based Twitter Features (Student Abstract). | Nupoor Gandhi, Alex Morales, Sally Man-Pui Chan, Dolores Albarracin, ChengXiang Zhai |
| 2020 | AAAI | Transductive Ensemble Learning for Neural Machine Translation. | Yiren Wang, Lijun Wu, Yingce Xia, Tao Qin, ChengXiang Zhai, Tie-Yan Liu |
| 2020 | CIKM | Empirical Analysis of Impact of Query-Specific Customization of nDCG: A Case-Study with Learning-to-Rank Methods. | Shubhra (Santu) K. Karmaker, Parikshit Sondhi, ChengXiang Zhai |
| 2020 | EMNLP | Multi-task Learning for Multilingual Neural Machine Translation. | Yiren Wang, ChengXiang Zhai, Hany Hassan |
| 2020 | ICTIR | Leveraging Personalized Sentiment Lexicons for Sentiment Analysis. | Dominic Seyler, Jiaming Shen, Jinfeng Xiao, Yiren Wang, ChengXiang Zhai |
| 2020 | KDD | Finding Contextually Consistent Information Units in Legal Text. | Dominic Seyler, Paul Bruin, Pavan Bayyapu, ChengXiang Zhai |
| 2020 | SIGCSE | Collective Development of Large Scale Data Science Products via Modularized Assignments: An Experience Report. | Bhavya, Assma Boughoula, Aaron Green, ChengXiang Zhai |
| 2020 | SIGIR | FigExplorer: A System for Retrieval and Exploration of Figures from Collections of Research Articles. | Saar Kuzi, ChengXiang Zhai, Yin Tian, Haichuan Tang |
| 2020 | SIGIR | A Study of Methods for the Generation of Domain-Aware Word Embeddings. | Dominic Seyler, ChengXiang Zhai |
| 2020 | SIGIR | Interactive Information Retrieval: Models, Algorithms, and Evaluation. | ChengXiang Zhai |
| 2019 | AAAI | Non-Autoregressive Machine Translation with Auxiliary Regularization. | Yiren Wang, Fei Tian, Di He, Tao Qin, ChengXiang Zhai, Tie-Yan Liu |
| 2019 | CIKM | Analysis of Adaptive Training for Learning to Rank in Information Retrieval. | Saar Kuzi, Sahiti Labhishetty, Shubhra Kanti Karmaker Santu, Prasad Pradip Joshi, ChengXiang Zhai |
| 2019 | ECIR | Figure Retrieval from Collections of Research Articles. | Saar Kuzi, ChengXiang Zhai |
| 2019 | ICLR | Multi-Agent Dual Learning. | Yiren Wang, Yingce Xia, Tianyu He, Fei Tian, Tao Qin, ChengXiang Zhai, Tie-Yan Liu |
| 2019 | ICWSM | A Generative Model for Discovering Action-Based Roles and Community Role Compositions on Community Question Answering Platforms. | Chase Geigle, Himel Dev, Hari Sundaram, ChengXiang Zhai |
| 2019 | ICWSM | Adapting Sequence to Sequence Models for Text Normalization in Social Media. | Ismini Lourentzou, Kabir Manghnani, ChengXiang Zhai |
| 2019 | SIGIR | Help Me Search: Leveraging User-System Collaboration for Query Construction to Improve Accuracy for Difficult Queries. | Saar Kuzi, Abhishek Narwekar, Anusri Pampari, ChengXiang Zhai |
| 2018 | CIKM | JIM: Joint Influence Modeling for Collective Search Behavior. | Shubhra Kanti Karmaker Santu, Liangda Li, Yi Chang, ChengXiang Zhai |
| 2018 | ITiCSE | CLaDS: a cloud-based virtual lab for the delivery of scalable hands-on assignments for practical data science education. | Chase Geigle, Ismini Lourentzou, Hari Sundaram, ChengXiang Zhai |
| 2018 | PSB | VisAGE: Integrating external knowledge into electronic medical record visualization. | Edward W. Huang, Sheng Wang, ChengXiang Zhai |
| 2018 | SIGIR | Are we on the Right Track?: An Examination of Information Retrieval Methodologies. | Enrique Amig, Hui Fang, Stefano Mizzaro, ChengXiang Zhai |
| 2018 | SIGIR | A Taxonomy of Queries for E-commerce Search. | Parikshit Sondhi, Mohit Sharma, Pranam Kolari, ChengXiang Zhai |
| 2018 | SIGIR | A Tutorial on Probabilistic Topic Models for Text Data Retrieval and Analysis. | ChengXiang Zhai, Chase Geigle |
| 2017 | AMIA | Framing Electronic Medical Records as Polylingual Documents in Query Expansion. | Edward W. Huang, Sheng Wang, Doris J. Lee, Runshun Zhang, Baoyan Liu, Xuezhong Zhou, ChengXiang Zhai |
| 2017 | CIKM | A Study of Feature Construction for Text-based Forecasting of Time Series Variables. | Yiren Wang, Dominic Seyler, Shubhra Kanti Karmaker Santu, ChengXiang Zhai |
| 2017 | IJCAI | ContextCare: Incorporating Contextual Information Networks to Representation Learning on Medical Forum Data. | Stan Zhao, Meng Jiang, Quan Yuan, Bing Qin, Ting Liu, ChengXiang Zhai |
| 2017 | ICTIR | Information Retrieval Evaluation as Search Simulation: A General Formal Framework for IR Evaluation. | Yinan Zhang, Xueqing Liu, ChengXiang Zhai |
| 2017 | WWW | Numerical Facet Range Partition: Evaluation Metric and Methods. | Xueqing Liu, ChengXiang Zhai, Wei Han, Onur Gngr |
| 2017 | WWW | Modeling the Influence of Popular Trending Events on User Search Behavior. | Shubhra Kanti Karmaker Santu, Liangda Li, Dae Hoon Park, Yi Chang, ChengXiang Zhai |
| 2017 | SIGIR | Axiomatic Thinking for Information Retrieval: And Related Tasks. | Enrique Amig, Hui Fang, Stefano Mizzaro, ChengXiang Zhai |
| 2017 | SIGIR | On Application of Learning to Rank for E-Commerce Search. | Shubhra Kanti Karmaker Santu, Parikshit Sondhi, ChengXiang Zhai |
| 2017 | SIGIR | Probabilistic Topic Models for Text Data Retrieval and Analysis. | ChengXiang Zhai |
| 2017 | WSDM | Constructing and Embedding Abstract Event Causality Networks from Text Snippets. | Sendong Zhao, Quan Wang, Sean Massung, Bing Qin, Ting Liu, Bin Wang, ChengXiang Zhai |
| 2016 | ACL | MeTA: A Unified Toolkit for Text Retrieval and Analysis. | Sean Massung, Chase Geigle, ChengXiang Zhai |
| 2016 | CIKM | Mobile App Retrieval for Social Media Users via Inference of Implicit Intent in Social Media Text. | Dae Hoon Park, Yi Fang, Mengwen Liu, ChengXiang Zhai |
| 2016 | CIKM | Generative Feature Language Models for Mining Implicit Features from Customer Reviews. | Shubhra Kanti Karmaker Santu, Parikshit Sondhi, ChengXiang Zhai |
| 2016 | SIGIR | Learning Query and Document Relevance from a Web-scale Click Graph. | Shan Jiang, Yuening Hu, Changsung Kang, Tim Daly Jr., Dawei Yin, Yi Chang, ChengXiang Zhai |
| 2016 | SIGIR | A Sequential Decision Formulation of the Interface Card Model for Interactive IR. | Yinan Zhang, ChengXiang Zhai |
| 2015 | CIKM | Mining Coordinated Intent Representation for Entity Search and Recommendation. | Huizhong Duan, ChengXiang Zhai |
| 2015 | IJCNN | Joint adaptive loss and l2/l0-norm minimization for unsupervised feature selection. | Mingjie Qian, ChengXiang Zhai |
| 2015 | ICTIR | Axiomatic Analysis of Smoothing Methods in Language Models for Pseudo-Relevance Feedback. | Hussein Hazimeh, ChengXiang Zhai |
| 2015 | ICWE | Beomap: Ad Hoc Topic Maps for Enhanced Exploration of Social Media Data. | Martin Leginus, ChengXiang Zhai, Peter Dolog |
| 2015 | SIGIR | Retrieval of Relevant Opinion Sentences for New Products. | Dae Hoon Park, Hyun Duk Kim, ChengXiang Zhai, Lifan Guo |
| 2015 | SIGIR | Leveraging User Reviews to Improve Accuracy for Mobile App Retrieval. | Dae Hoon Park, Mengwen Liu, ChengXiang Zhai, Haohong Wang |
| 2015 | SIGIR | Towards a Game-Theoretic Framework for Information Retrieval. | ChengXiang Zhai |
| 2015 | SDM | SpecLDA: Modeling Product Reviews and Specifications to Generate Augmented Specifications. | Dae Hoon Park, ChengXiang Zhai, Lifan Guo |
| 2014 | CIKM | Revisiting the Divergence Minimization Feedback Model. | Yuanhua Lv, ChengXiang Zhai |
| 2014 | CIKM | Unsupervised Feature Selection for Multi-View Clustering on Text-Image Web News Data. | Mingjie Qian, ChengXiang Zhai |
| 2014 | CIKM | Mining Semi-Structured Online Knowledge Bases to Answer Natural Language Questions on Community QA Websites. | Parikshit Sondhi, ChengXiang Zhai |
| 2014 | SIGIR | VIRLab: a web-based virtual lab for learning and studying information retrieval models. | Hui Fang, Hao Wu, Peilin Yang, ChengXiang Zhai |
| 2014 | SIGIR | Axiomatic analysis and optimization of information retrieval models. | Hui Fang, ChengXiang Zhai |
| 2014 | SIGIR | VIRLab: A Platform for Privacy-Preserving Evaluation for Information Retrieval Models. | Hui Fang, ChengXiang Zhai |
| 2014 | SIGIR | A two-dimensional click model for query auto-completion. | Yanen Li, Anlei Dong, Hongning Wang, Hongbo Deng, Yi Chang, ChengXiang Zhai |
| 2014 | WSDM | User modeling in search logs via a nonparametric bayesian approach. | Hongning Wang, ChengXiang Zhai, Feng Liang, Anlei Dong, Yi Chang |
| 2014 | SDM | A Constrained Hidden Markov Model Approach for Non-Explicit Citation Context Extraction. | Parikshit Sondhi, ChengXiang Zhai |
| 2013 | CIKM | A probabilistic mixture model for mining and analyzing product search log. | Huizhong Duan, ChengXiang Zhai, Jinxing Cheng, Rohit Kumar |
| 2013 | CIKM | FindiLike: a preference driven entity search engine for evaluating entity retrieval and opinion summarization. | Kavita Ganesan, ChengXiang Zhai |
| 2013 | CIKM | Compact explanatory opinion summarization. | Hyun Duk Kim, Mal Castellanos, Meichun Hsu, ChengXiang Zhai, Umeshwar Dayal, Riddhiman Ghosh |
| 2013 | CIKM | Mining causal topics in text data: iterative topic modeling with time series feedback. | Hyun Duk Kim, Mal Castellanos, Meichun Hsu, ChengXiang Zhai, Thomas A. Rietz, Daniel Diermeier |
| 2013 | CIKM | Unsupervised identification of synonymous query intent templates for attribute intents. | Yanen Li, Bo-June Paul Hsu, ChengXiang Zhai |
| 2013 | CIKM | Mining entity attribute synonyms via compact clustering. | Yanen Li, Bo-June Paul Hsu, ChengXiang Zhai, Kuansan Wang |
| 2013 | ICTIR | Statistical Translation Language Model for Twitter Search. | Maryam Karimzadehgan, ChengXiang Zhai, Miles Efron |
| 2013 | ICTIR | Information Retrieval with Time Series Query. | Hyun Duk Kim, Danila Nikitin, ChengXiang Zhai, Mal Castellanos, Meichun Hsu |
| 2013 | ICTIR | Exploiting Forum Thread Structures to Improve Thread Clustering. | Kumaresh Pattabiraman, Parikshit Sondhi, ChengXiang Zhai |
| 2013 | ICTIR | Axiomatic Analysis and Optimization of Information Retrieval Models. | ChengXiang Zhai, Hui Fang |
| 2013 | WWW | Content-aware click modeling. | Hongning Wang, ChengXiang Zhai, Anlei Dong, Yi Chang |
| 2013 | SIGIR | Ranking explanatory sentences for opinion summarization. | Hyun Duk Kim, Mal Castellanos, Meichun Hsu, ChengXiang Zhai, Umeshwar Dayal, Riddhiman Ghosh |
| 2012 | CIKM | Click patterns: an empirical representation of complex query intents. | Huizhong Duan, Emre Kiciman, ChengXiang Zhai |
| 2012 | CIKM | InCaToMi: integrative causal topic miner between textual and non-textual time series data. | Hyun Duk Kim, ChengXiang Zhai, Thomas A. Rietz, Daniel Diermeier, Meichun Hsu, Mal Castellanos, Carlos Ceja Limon |
| 2012 | CIKM | Unsupervised discovery of opposing opinion networks from forum discussions. | Yue Lu, Hongning Wang, ChengXiang Zhai, Dan Roth |
| 2012 | CIKM | Query likelihood with negative query generation. | Yuanhua Lv, ChengXiang Zhai |
| 2012 | CIKM | Mining long-lasting exploratory user interests from search history. | Bin Tan, Yuanhua Lv, ChengXiang Zhai |
| 2012 | CIKM | BiasTrust: teaching biased users about controversial topics. | V. G. Vinod Vydiswaran, ChengXiang Zhai, Dan Roth, Peter Pirolli |
| 2012 | ECIR | Score Transformation in Linear Combination for Multi-criteria Relevance Ranking. | Shima Gerani, ChengXiang Zhai, Fabio Crestani |
| 2012 | ECIR | Axiomatic Analysis of Translation Language Model for Information Retrieval. | Maryam Karimzadehgan, ChengXiang Zhai |
| 2012 | ECIR | A Log-Logistic Model-Based Interpretation of TF Normalization of BM25. | Yuanhua Lv, ChengXiang Zhai |
| 2012 | ECIR | Reliability Prediction of Webpages in the Medical Domain. | Parikshit Sondhi, V. G. Vinod Vydiswaran, ChengXiang Zhai |
| 2012 | EMNLP | A Discriminative Model for Query Spelling Correction with Latent Structural SVM. | Huizhong Duan, Yanen Li, ChengXiang Zhai, Dan Roth |
| 2012 | KDD | SympGraph: a framework for mining clinical notes through symptom relation graphs. | Parikshit Sondhi, Jimeng Sun, Hanghang Tong, ChengXiang Zhai |
| 2012 | WWW | FindiLike: preference driven entity search. | Kavita Ganesan, ChengXiang Zhai |
| 2012 | WWW | Micropinion generation: an unsupervised approach to generating ultra-concise summaries of opinions. | Kavita Ganesan, ChengXiang Zhai, Evelyne Viegas |
| 2012 | WWW | CloudSpeller: query spelling correction by using a unified hidden markov model with web-scale resources. | Yanen Li, Huizhong Duan, ChengXiang Zhai |
| 2012 | SIGIR | A generalized hidden Markov model with discriminative training for query spelling correction. | Yanen Li, Huizhong Duan, ChengXiang Zhai |
| 2012 | WSDM | Tapping into knowledge base for concept feedback: leveraging conceptnet to improve search results for difficult queries. | Alexander Kotov, ChengXiang Zhai |
| 2011 | ACL | Structural Topic Model for Latent Topical Structure Analysis. | Hongning Wang, Duo Zhang, ChengXiang Zhai |
| 2011 | CIKM | Automatic query reformulation with syntactic operators to alleviate search difficulty. | Huizhong Duan, Rui Li, ChengXiang Zhai |
| 2011 | CIKM | Improving retrieval accuracy of difficult queries through generalizing negative document language models. | Maryam Karimzadehgan, ChengXiang Zhai |
| 2011 | CIKM | Interactive sense feedback for difficult queries. | Alexander Kotov, ChengXiang Zhai |
| 2011 | CIKM | Lower-bounding term frequency normalization. | Yuanhua Lv, ChengXiang Zhai |
| 2011 | CIKM | Adaptive term frequency normalization for BM25. | Yuanhua Lv, ChengXiang Zhai |
| 2011 | ICTIR | Axiomatic Analysis and Optimization of Information Retrieval Models. | ChengXiang Zhai |
| 2011 | KDD | Content-driven trust propagation framework. | V. G. Vinod Vydiswaran, ChengXiang Zhai, Dan Roth |
| 2011 | KDD | Latent aspect rating analysis without aspect keyword supervision. | Hongning Wang, Yue Lu, ChengXiang Zhai |
| 2011 | WWW | Automatic construction of a context-aware sentiment lexicon: an optimization approach. | Yue Lu, Mal Castellanos, Umeshwar Dayal, ChengXiang Zhai |
| 2011 | SIGIR | Unsupervised query segmentation using clickthrough for information retrieval. | Yanen Li, Bo-June Paul Hsu, ChengXiang Zhai, Kuansan Wang |
| 2011 | SIGIR | When documents are very long, BM25 fails! | Yuanhua Lv, ChengXiang Zhai |
| 2011 | SIGIR | A boosting approach to improving pseudo-relevance feedback. | Yuanhua Lv, ChengXiang Zhai, Wan Chen |
| 2011 | SIGIR | Learning online discussion structures by conditional random fields. | Hongning Wang, Chi Wang, ChengXiang Zhai, Jiawei Han |
| 2011 | SIGIR | Beyond search: statistical topic models for text analysis. | ChengXiang Zhai |
| 2011 | WSDM | Mining named entities with temporally correlated bursts from multilingual web news streams. | Alexander Kotov, ChengXiang Zhai, Richard Sproat |
| 2010 | ACL | Cross-Lingual Latent Topic Extraction. | Duo Zhang, Qiaozhu Mei, ChengXiang Zhai |
| 2010 | CIKM | Exploration-exploitation tradeoff in interactive relevance feedback. | Maryam Karimzadehgan, ChengXiang Zhai |
| 2010 | CIKM | Improving one-class collaborative filtering by incorporating rich user information. | Yanen Li, Jia Hu, ChengXiang Zhai, Ye Chen |
| 2010 | CIKM | PTM: probabilistic topic mapping model for mining parallel document collections. | Duo Zhang, Jimeng Sun, ChengXiang Zhai, Abhijit Bose, Nikos Anerousis |
| 2010 | COLING | Opinosis: A Graph Based Approach to Abstractive Summarization of Highly Redundant Opinions. | Kavita Ganesan, ChengXiang Zhai, Jiawei Han |
| 2010 | COLING | Exploiting Structured Ontology to Organize Scattered Online Opinions. | Yue Lu, Huizhong Duan, Hongning Wang, ChengXiang Zhai |
| 2010 | COLING | Shallow Information Extraction from Medical Forum Data. | Parikshit Sondhi, Manish Gupta, ChengXiang Zhai, Julia Hockenmaier |
| 2010 | ECIR | Aggregation of Multiple Judgments for Evaluating Ordered Lists. | Hyun Duk Kim, ChengXiang Zhai, Jiawei Han |
| 2010 | EMNLP | Summarizing Contrastive Viewpoints in Opinionated Text. | Michael J. Paul, ChengXiang Zhai, Roxana Girju |
| 2010 | WWW | Towards natural question guided search. | Alexander Kotov, ChengXiang Zhai |
| 2010 | SIGIR | Estimation of statistical translation models based on mutual information for ad hoc information retrieval. | Maryam Karimzadehgan, ChengXiang Zhai |
| 2010 | SIGIR | Positional relevance model for pseudo-relevance feedback. | Yuanhua Lv, ChengXiang Zhai |
| 2009 | CIKM | Constrained multi-aspect expertise matching for committee review assignment. | Maryam Karimzadehgan, ChengXiang Zhai |
| 2009 | CIKM | Generating comparative summaries of contradictory opinions in text. | Hyun Duk Kim, ChengXiang Zhai |
| 2009 | CIKM | Adaptive relevance feedback in information retrieval. | Yuanhua Lv, ChengXiang Zhai |
| 2009 | CIKM | A comparative study of methods for estimating query language models with pseudo feedback. | Yuanhua Lv, ChengXiang Zhai |
| 2009 | CIKM | Beyond hyperlinks: organizing information footprints in search logs to support effective browsing. | Xuanhui Wang, Bin Tan, Azadeh Shakery, ChengXiang Zhai |
| 2009 | ICDM | Parallel PathFinder Algorithms for Mining Structures from Graphs. | Samson Hauguel, ChengXiang Zhai, Jiawei Han |
| 2009 | ISI | Assured Information Sharing Life Cycle. | Tim Finin, Anupam Joshi, Hillol Kargupta, Yelena Yesha, Joel Sachs, Elisa Bertino, Ninghui Li, Chris Clifton, Gene Spafford, Bhavani Thuraisingham, Murat Kantarcioglu, Alain Bensoussan, Nathan Berg, Latifur Khan, Jiawei Han, ChengXiang Zhai, Ravi S. Sandhu, Shouhuai Xu, Jim Massaro, Lada A. Adamic |
| 2009 | WWW | Rated aspect summarization of short comments. | Yue Lu, ChengXiang Zhai, Neel Sundaresan |
| 2009 | SIGIR | Positional language models for information retrieval. | Yuanhua Lv, ChengXiang Zhai |
| 2009 | SIGIR | Massive Implicit Feedback: Organizing Search Logs into Topic Maps for Collaborative Surfing. | Xuanhui Wang, ChengXiang Zhai |
| 2009 | SDM | Topic Cube: Topic Modeling for OLAP on Multidimensional Text Databases. | Duo Zhang, ChengXiang Zhai, Jiawei Han |
| 2008 | ACL | Generating Impact-Based Summaries for Scientific Literature. | Qiaozhu Mei, ChengXiang Zhai |
| 2008 | CIKM | Multi-aspect expertise matching for review assignment. | Maryam Karimzadehgan, ChengXiang Zhai, Geneva G. Belford |
| 2008 | CIKM | Mining term association patterns from search logs for effective query reformulation. | Xuanhui Wang, ChengXiang Zhai |
| 2008 | DASFAA | Ranking Database Queries with User Feedback: A Neural Network Approach. | Ganesh Agarwal, Nevedita Mallick, Srinivasan Turuvekere, ChengXiang Zhai |
| 2008 | KDD | Mining multi-faceted overviews of arbitrary topics in a text collection. | Xu Ling, Qiaozhu Mei, ChengXiang Zhai, Bruce R. Schatz |
| 2008 | WWW | Topic modeling with network regularization. | Qiaozhu Mei, Deng Cai, Duo Zhang, ChengXiang Zhai |
| 2008 | SIGIR | A general optimization framework for smoothing language models on graph structures. | Qiaozhu Mei, Duo Zhang, ChengXiang Zhai |
| 2008 | SIGIR | A study of methods for negative relevance feedback. | Xuanhui Wang, Hui Fang, ChengXiang Zhai |
| 2007 | ACL | Instance Weighting for Domain Adaptation in NLP. | Jing Jiang, ChengXiang Zhai |
| 2007 | CIKM | A two-stage approach to domain adaptation for statistical classifiers. | Jing Jiang, ChengXiang Zhai |
| 2007 | CIKM | Improve retrieval accuracy for difficult queries using negative feedback. | Xuanhui Wang, Hui Fang, ChengXiang Zhai |
| 2007 | ECIR | Probabilistic Models for Expert Finding. | Hui Fang, ChengXiang Zhai |
| 2007 | ICDE | Collaborative Wrapping: A Turbo Framework for Web Data Extraction. | Shui-Lung Chuang, Kevin Chen-Chuan Chang, ChengXiang Zhai |
| 2007 | KDD | Automatic labeling of multinomial topic models. | Qiaozhu Mei, Xuehua Shen, ChengXiang Zhai |
| 2007 | KDD | Mining correlated bursty topic patterns from coordinated text streams. | Xuanhui Wang, ChengXiang Zhai, Xiao Hu, Richard Sproat |
| 2007 | NAACL | A Systematic Exploration of the Feature Space for Relation Extraction. | Jing Jiang, ChengXiang Zhai |
| 2007 | NAACL | Statistical Language Models for Information Retrieval. | ChengXiang Zhai |
| 2007 | WWW | Topic sentiment mixture: modeling facets and opinions in weblogs. | Qiaozhu Mei, Xu Ling, Matthew Wondra, Hang Su, ChengXiang Zhai |
| 2007 | SIGIR | A study of Poisson query generation model for information retrieval. | Qiaozhu Mei, Hui Fang, ChengXiang Zhai |
| 2007 | SIGIR | Term feedback for information retrieval with language models. | Bin Tan, Atulya Velivelli, Hui Fang, ChengXiang Zhai |
| 2007 | SIGIR | An exploration of proximity measures in information retrieval. | Tao Tao, ChengXiang Zhai |
| 2007 | SIGIR | Learn from web search logs to organize search results. | Xuanhui Wang, ChengXiang Zhai |
| 2007 | VLDB | Context-Aware Wrapping: Synchronized Data Extraction. | Shui-Lung Chuang, Kevin Chen-Chuan Chang, ChengXiang Zhai |
| 2006 | ACL | Named Entity Transliteration with Comparable Corpora. | Richard Sproat, Tao Tao, ChengXiang Zhai |
| 2006 | CIKM | A probabilistic relevance propagation model for hypertext retrieval. | Azadeh Shakery, ChengXiang Zhai |
| 2006 | CIKM | Best-k queries on database systems. | Tao Tao, ChengXiang Zhai |
| 2006 | EMNLP | Unsupervised Named Entity Transliteration Using Temporal and Phonetic Correlation. | Tao Tao, Su-Youn Yoon, Andrew Fister, Richard Sproat, ChengXiang Zhai |
| 2006 | KDD | Generating semantic annotations for frequent patterns with context analysis. | Qiaozhu Mei, Dong Xin, Hong Cheng, Jiawei Han, ChengXiang Zhai |
| 2006 | KDD | A mixture model for contextual text mining. | Qiaozhu Mei, ChengXiang Zhai |
| 2006 | KDD | Mining long-term search history to improve search accuracy. | Bin Tan, Xuehua Shen, ChengXiang Zhai |
| 2006 | NAACL | Exploiting Domain Structure for Named Entity Recognition. | Jing Jiang, ChengXiang Zhai |
| 2006 | NAACL | Language Model Information Retrieval with Document Expansion. | Tao Tao, Xuanhui Wang, Qiaozhu Mei, ChengXiang Zhai |
| 2006 | PSB | Automatically Generating Gene Summaries from Biomedical Literature. | Xu Ling, Jing Jiang, Xin He, Qiaozhu Mei, ChengXiang Zhai, Bruce R. Schatz |
| 2006 | WWW | A probabilistic approach to spatiotemporal theme pattern mining on weblogs. | Qiaozhu Mei, Chao Liu, Hang Su, ChengXiang Zhai |
| 2006 | SIGIR | Semantic term matching in axiomatic approaches to information retrieval. | Hui Fang, ChengXiang Zhai |
| 2006 | SIGIR | Regularized estimation of mixture models for robust pseudo-relevance feedback. | Tao Tao, ChengXiang Zhai |
| 2006 | SIGIR | Latent semantic analysis for multiple-type interrelated data objects. | Xuanhui Wang, Jian-Tao Sun, Zheng Chen, ChengXiang Zhai |
| 2005 | CIKM | Accurately extracting coherent relevant passages using hidden Markov models. | Jing Jiang, ChengXiang Zhai |
| 2005 | CIKM | Implicit user modeling for personalized search. | Xuehua Shen, Bin Tan, ChengXiang Zhai |
| 2005 | CIKM | Accurate language model estimation with document expansion. | Tao Tao, Xuanhui Wang, Qiaozhu Mei, ChengXiang Zhai |
| 2005 | KDD | Discovering evolutionary theme patterns from text: an exploration of temporal text mining. | Qiaozhu Mei, ChengXiang Zhai |
| 2005 | KDD | Mining comparable bilingual text corpora for cross-language information integration. | Tao Tao, ChengXiang Zhai |
| 2005 | SIGIR | An exploration of axiomatic approaches to information retrieval. | Hui Fang, ChengXiang Zhai |
| 2005 | SIGIR | Context-sensitive information retrieval using implicit feedback. | Xuehua Shen, Bin Tan, ChengXiang Zhai |
| 2005 | SIGIR | UCAIR: a personalized search toolbar. | Xuehua Shen, Bin Tan, ChengXiang Zhai |
| 2005 | SIGIR | Active feedback in ad hoc information retrieval. | Xuehua Shen, ChengXiang Zhai |
| 2004 | KDD | A cross-collection mixture model for comparative text mining. | ChengXiang Zhai, Atulya Velivelli, Bei Yu |
| 2004 | SIGIR | A formal study of information retrieval heuristics. | Hui Fang, Tao Tao, ChengXiang Zhai |
| 2004 | SIGIR | ACES: a contextual engine for search. | Xuehua Shen, Smitha Sriram, ChengXiang Zhai |
| 2004 | SIGIR | A session-based search engine. | Smitha Sriram, Xuehua Shen, ChengXiang Zhai |
| 2004 | SIGIR | A two-stage mixture model for pseudo feedback. | Tao Tao, ChengXiang Zhai |
| 2003 | CIKM | Collaborative filtering with decoupled models for preferences and ratings. | Rong Jin, Luo Si, ChengXiang Zhai, James P. Callan |
| 2003 | CIKM | Text classification from positive and unlabeled documents. | Hwanjo Yu, ChengXiang Zhai, Jiawei Han |
| 2003 | SIGIR | Error analysis of difficult TREC topics. | Xiao Hu, Sindhura Bandhakavi, ChengXiang Zhai |
| 2003 | SIGIR | Exploiting query history for document ranking in interactive information retrieval. | Xuehua Shen, ChengXiang Zhai |
| 2003 | SIGIR | Beyond independent relevance: methods and evaluation metrics for subtopic retrieval. | ChengXiang Zhai, William W. Cohen, John D. Lafferty |
| 2003 | UAI | Preference-based Graphic Models for Collaborative Filtering. | Rong Jin, Luo Si, ChengXiang Zhai |
| 2002 | SIGIR | Title language model for information retrieval. | Rong Jin, Alexander G. Hauptmann, ChengXiang Zhai |
| 2002 | SIGIR | Two-stage language models for information retrieval. | ChengXiang Zhai, John D. Lafferty |
| 2001 | CIKM | Model-based Feedback in the Language Modeling Approach to Information Retrieval. | ChengXiang Zhai, John D. Lafferty |
| 2001 | SIGIR | Document Language Models, Query Models, and Risk Minimization for Information Retrieval. | John D. Lafferty, ChengXiang Zhai |
| 2001 | SIGIR | A Study of Smoothing Methods for Language Models Applied to Ad Hoc Information Retrieval. | ChengXiang Zhai, John D. Lafferty |
| 2000 | SIGIR | Exploration of a heuristic approach to threshold learning in adaptive filtering. | ChengXiang Zhai, Peter Jansen, David A. Evans |
| 1996 | ACL | Noun-Phrase Analysis in Unrestricted Text for Information Retrieval. | David A. Evans, ChengXiang Zhai |