| 2025 | ACL | MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs. | Kaustubh Deshpande, Ved Sirdeshmukh, Johannes Baptist Mols, Lifeng Jin, Ed-Yeremai Hernandez-Cardona, Dean Lee, Jeremy Kritz, Willow E. Primack, Summer Yue, Chen Xing |
| 2025 | COLING | Entropy Guided Extrapolative Decoding to Improve Factuality in Large Language Models. | Souvik Das, Lifeng Jin, Linfeng Song, Haitao Mi, Baolin Peng, Dong Yu |
| 2024 | ACL | Improving LLM Generations via Fine-Grained Self-Endorsement. | Ante Wang, Linfeng Song, Baolin Peng, Lifeng Jin, Ye Tian, Haitao Mi, Jinsong Su, Dong Yu |
| 2024 | ACL | Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation. | Xiaoying Zhang, Baolin Peng, Ye Tian, Jingyan Zhou, Lifeng Jin, Linfeng Song, Haitao Mi, Helen Meng |
| 2024 | COLING | A Knowledge Plug-and-Play Test Bed for Open-domain Dialogue Generation. | Xiangci Li, Linfeng Song, Lifeng Jin, Haitao Mi, Jessica Ouyang, Dong Yu |
| 2024 | EACL | Inconsistent dialogue responses and how to recover from them. | Mian Zhang, Lifeng Jin, Linfeng Song, Haitao Mi, Dong Yu |
| 2024 | EMNLP | Self-Consistency Boosts Calibration for Math Reasoning. | Ante Wang, Linfeng Song, Ye Tian, Baolin Peng, Lifeng Jin, Haitao Mi, Jinsong Su, Dong Yu |
| 2024 | ICLR | The Trickle-down Impact of Reward Inconsistency on RLHF. | Lingfeng Shen, Sihao Chen, Linfeng Song, Lifeng Jin, Baolin Peng, Haitao Mi, Daniel Khashabi, Dong Yu |
| 2023 | ACL | Bi-level Finetuning with Task-dependent Similarity Structure for Low-resource Training. | Sai Ashish Somayajula, Lifeng Jin, Linfeng Song, Haitao Mi, Dong Yu |
| 2023 | ACL | SafeConv: Explaining and Correcting Conversational Unsafe Behavior. | Mian Zhang, Lifeng Jin, Linfeng Song, Haitao Mi, Wenliang Chen, Dong Yu |
| 2023 | EACL | How do Words Contribute to Sentence Semantics? Revisiting Sentence Embeddings with a Perturbation Method. | Wenlin Yao, Lifeng Jin, Hongming Zhang, Xiaoman Pan, Kaiqiang Song, Dian Yu, Dong Yu, Jianshu Chen |
| 2023 | EACL | Friend-training: Learning from Models of Different but Related Tasks. | Mian Zhang, Lifeng Jin, Linfeng Song, Haitao Mi, Xiabing Zhou, Dong Yu |
| 2022 | AAAI | Hierarchical Context Tagging for Utterance Rewriting. | Lisa Jin, Linfeng Song, Lifeng Jin, Dong Yu, Daniel Gildea |
| 2022 | EMNLP | Cross-lingual Text-to-SQL Semantic Parsing with Representation Mixup. | Peng Shi, Linfeng Song, Lifeng Jin, Haitao Mi, He Bai, Jimmy Lin, Dong Yu |
| 2022 | EMNLP | Dynamic Augmentation Data Selection for Few-shot Text Classification. | Guangliang Liu, Lifeng Jin, Owen Yuan, Jiayu Zhou |
| 2022 | EMNLP | Salience Allocation as Guidance for Abstractive Summarization. | Fei Wang, Kaiqiang Song, Hongming Zhang, Lifeng Jin, Sangwoo Cho, Wenlin Yao, Xiaoyang Wang, Muhao Chen, Dong Yu |
| 2022 | EMNLP | Learning a Grammar Inducer from Massive Uncurated Instructional Videos. | Songyang Zhang, Linfeng Song, Lifeng Jin, Haitao Mi, Kun Xu, Dong Yu, Jiebo Luo |
| 2021 | ACL | Domain-Adaptive Pretraining Methods for Dialogue Understanding. | Han Wu, Kun Xu, Linfeng Song, Lifeng Jin, Haisong Zhang, Linqi Song |
| 2021 | EMNLP | Character-based PCFG Induction for Modeling the Syntactic Acquisition of Morphologically Rich Languages. | Lifeng Jin, Byung-Doh Oh, William Schuler |
| 2021 | EMNLP | Instance-adaptive training with noise-robust losses against noisy labels. | Lifeng Jin, Linfeng Song, Kun Xu, Dong Yu |
| 2021 | EMNLP | Connect-the-Dots: Bridging Semantics between Words and Definitions via Aligning Word Sense Inventories. | Wenlin Yao, Xiaoman Pan, Lifeng Jin, Jianshu Chen, Dian Yu, Dong Yu |
| 2021 | NAACL | Video-aided Unsupervised Grammar Induction. | Songyang Zhang, Linfeng Song, Lifeng Jin, Kun Xu, Dong Yu, Jiebo Luo |
| 2020 | AAAI | Relation Extraction Exploiting Full Dependency Forests. | Lifeng Jin, Linfeng Song, Yue Zhang, Kun Xu, Wei-Yun Ma, Dong Yu |
| 2020 | IJCNLP | Grounded PCFG Induction with Images. | Lifeng Jin, William Schuler |
| 2019 | ACL | Unsupervised Learning of PCFGs with Normalizing Flow. | Lifeng Jin, Finale Doshi-Velez, Timothy Miller, Lane Schwartz, William Schuler |
| 2019 | ACL | Variance of Average Surprisal: A Better Predictor for Quality of Grammar from Unsupervised PCFG Induction. | Lifeng Jin, William Schuler |
| 2019 | IGARSS | Design and Analysis of Radiometric Calibration Mission in-orbit for Environment and Disasters Monitoring Satellite. | Yang Zhu, Dexin Sun, Xuebin Liu, Lifeng Jin, Jun Zhu, Zhaoguang Bai, Jun Dong, Bin Wu, Min Huang, Huan Yin, Qipeng Cao, Jin Hong |
| 2018 | EMNLP | Depth-bounding is effective: Improvements and Evaluation of Unsupervised PCFG Induction. | Lifeng Jin, Finale Doshi-Velez, Timothy Miller, William Schuler, Lane Schwartz |
| 2016 | COLING | Memory-Bounded Left-Corner Unsupervised Grammar Induction on Child-Directed Input. | Cory Shain, William Bryce, Lifeng Jin, Victoria Krakovna, Finale Doshi-Velez, Timothy Miller, William Schuler, Lane Schwartz |
| 2015 | EMNLP | The Overall Markedness of Discourse Relations. | Lifeng Jin, Marie-Catherine de Marneffe |
| 2015 | NAACL | A Comparison of Word Similarity Performance Using Explanatory and Non-explanatory Texts. | Lifeng Jin, William Schuler |
| 2014 | HCI | Sentences Extraction from Digital Publication for Domain-Specific Knowledge Service. | Mao Ye, Lifeng Jin, Zhi Tang, Jianbo Xu |
| 2014 | HCI | A Semantic Recommender System for Learning Based on Encyclopedia of Digital Publication. | Mao Ye, Lifeng Jin, Zhi Tang, Jianbo Xu |