Skip to content

Zhengxuan Wu

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

18

Venues

8

Active years

2019–2025

Best venue rank

A*

Where they publish

Papers

18 indexed papers, newest first.

YearVenueTitleAuthors
2025ICMLAxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders.Zhengxuan Wu, Aryaman Arora, Atticus Geiger, Zheng Wang, Jing Huang, Dan Jurafsky, Christopher D. Manning, Christopher Potts
2024ACLRAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations.Jing Huang, Zhengxuan Wu, Christopher Potts, Mor Geva, Atticus Geiger
2024CogSciSymbolic Variables in Distributed Networks that Count.Satchel Grant, Zhengxuan Wu, Jay McClelland, Noah D. Goodman
2024EMNLPDancing in Chains: Reconciling Instruction Following and Faithfulness in Language Models.Zhengxuan Wu, Yuhao Zhang, Peng Qi, Yumo Xu, Rujun Han, Yian Zhang, Jifan Chen, Bonan Min, Zhiheng Huang
2024ICMLIn-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation.Shiqi Chen, Miao Xiong, Junteng Liu, Zhengxuan Wu, Teng Xiao, Siyang Gao, Junxian He
2024NAACLpyvene: A Library for Understanding and Improving PyTorch Models via Interventions.Zhengxuan Wu, Atticus Geiger, Aryaman Arora, Jing Huang, Zheng Wang, Noah D. Goodman, Christopher D. Manning, Christopher Potts
2023ACLInducing Character-level Structure in Subword-based Language Models with Type-level Interchange Intervention Training.Jing Huang, Zhengxuan Wu, Kyle Mahowald, Christopher Potts
2023EMNLPOolong: Investigating What Makes Transfer Learning Hard with Controlled Studies.Zhengxuan Wu, Alex Tamkin, Isabel Papadimitriou
2023EMNLPMQuAKE: Assessing Knowledge Editing in Language Models via Multi-Hop Questions.Zexuan Zhong, Zhengxuan Wu, Christopher D. Manning, Christopher Potts, Danqi Chen
2023ICMLCausal Proxy Models for Concept-based Model Explanations.Zhengxuan Wu, Karel D'Oosterlinck, Atticus Geiger, Amir Zur, Christopher Potts
2022ICMLInducing Causal Structure for Interpretable Neural Networks.Atticus Geiger, Zhengxuan Wu, Hanson Lu, Josh Rozner, Elisa Kreiss, Thomas Icard, Noah D. Goodman, Christopher Potts
2022NAACLCausal Distillation for Language Models.Zhengxuan Wu, Atticus Geiger, Joshua Rozner, Elisa Kreiss, Hanson Lu, Thomas Icard, Christopher Potts, Noah D. Goodman
2021AAAIContext-Guided BERT for Targeted Aspect-Based Sentiment Analysis.Zhengxuan Wu, Desmond C. Ong
2021ACLDynaSent: A Dynamic Benchmark for Sentiment Analysis.Christopher Potts, Zhengxuan Wu, Atticus Geiger, Douwe Kiela
2021CHINot Now, Ask Later: Users Weaken Their Behavior Change Regimen Over Time, But Expect To Re-Strengthen It Imminently.Geza Kovacs, Zhengxuan Wu, Michael S. Bernstein
2021NAACLDynabench: Rethinking Benchmarking in NLP.Douwe Kiela, Max Bartolo, Yixin Nie, Divyansh Kaushik, Atticus Geiger, Zhengxuan Wu, Bertie Vidgen, Grusha Prasad, Amanpreet Singh, Pratik Ringshia, Zhiyi Ma, Tristan Thrush, Sebastian Riedel, Zeerak Waseem, Pontus Stenetorp, Robin Jia, Mohit Bansal, Christopher Potts, Adina Williams
2019ACIIAttending to Emotional Narratives.Zhengxuan Wu, Xiyu Zhang, Zhi-Xuan Tan, Jamil Zaki, Desmond C. Ong
2019CHIConservation of Procrastination: Do Productivity Interventions Save Time Or Just Redistribute It?Geza Kovacs, Drew Mylander Gregory, Zilin Ma, Zhengxuan Wu, Golrokh Emami, Jacob Ray, Michael S. Bernstein