| 2025 | EMNLP | Discourse Heuristics For Paradoxically Moral Self-Correction. | Guangliang Liu, Zimo Qi, Xitong Zhang, Kristen Marie Johnson |
| 2025 | EMNLP | Diagnosing Moral Reasoning Acquisition in Language Models: Pragmatics and Generalization. | Guangliang Liu, Zimo Qi, Xitong Zhang, Lei Jiang, Kristen Marie Johnson |
| 2025 | IJCNLP | On the Convergence of Moral Self-Correction in Large Language Models. | Guangliang Liu, Haitao Mao, Bochuan Cao, Xitong Zhang, Zhiyu Xue, Rongrong Wang, Kristen Marie Johnson |
| 2025 | IJCNLP | Moral Self-correction is Not An Innate Capability in Language Models. | Guangliang Liu, Zimo Qi, Xitong Zhang, Lu Cheng, Kristen Marie Johnson |
| 2025 | NAACL | A Survey to Recent Progress Towards Understanding In-Context Learning. | Haitao Mao, Guangliang Liu, Yao Ma, Rongrong Wang, Kristen Marie Johnson, Jiliang Tang |
| 2024 | ACL | Towards Understanding Task-agnostic Debiasing Through the Lenses of Intrinsic Bias and Forgetfulness. | Guangliang Liu, Milad Afshari, Xitong Zhang, Zhiyu Xue, Avrajit Ghosh, Bidhan Bashyal, Rongrong Wang, Kristen Marie Johnson |
| 2024 | COLING | ABLE: Agency-BeLiefs Embedding to Address Stereotypical Bias through Awareness Instead of Obliviousness. | Michelle YoungJin Kim, Junghwan Kim, Kristen Marie Johnson |
| 2024 | EMNLP | Intrinsic Self-correction for Enhanced Morality: An Analysis of Internal Mechanisms and the Superficial Hypothesis. | Guangliang Liu, Haitao Mao, Jiliang Tang, Kristen Marie Johnson |
| 2023 | ACL | Race, Gender, and Age Biases in Biomedical Masked Language Models. | Michelle Kim, Junghwan Kim, Kristen Marie Johnson |
| 2023 | EMNLP | PAC-tuning: Fine-tuning Pre-trained Language Models with PAC-driven Perturbed Gradient Descent. | Guangliang Liu, Zhiyu Xue, Xitong Zhang, Kristen Marie Johnson, Rongrong Wang |
| 2022 | COLING | CLoSE: Contrastive Learning of Subframe Embeddings for Political Bias Classification of News Media. | Michelle YoungJin Kim, Kristen Marie Johnson |