Atticus Geiger
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
20
Venues
6
Active years
2019–2026
Best venue rank
A*
Where they publish
Papers
20 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | ACL | Constructing Interpretable Features from Compositional Neuron Groups. | Or David Shafran, Atticus Geiger, Mor Geva |
| 2025 | ACL | Enhancing Automated Interpretability with Output-Centric Feature Descriptions. | Yoav Gur-Arieh, Roy Mayan, Chen Agassy, Atticus Geiger, Mor Geva |
| 2025 | ICLR | HyperDAS: Towards Automating Mechanistic Interpretability with Hypernetworks. | Jiuding Sun, Jing Huang, Sidharth Baskaran, Karel D'Oosterlinck, Christopher Potts, Michael Sklar, Atticus Geiger |
| 2025 | ICML | MIB: A Mechanistic Interpretability Benchmark. | Aaron Mueller, Atticus Geiger, Sarah Wiegreffe, Dana Arad, Ivn Arcuschin, Adam Belfki, Yik Siu Chan, Jaden Fried Fiotto-Kaufman, Tal Haklay, Michael Hanna, Jing Huang, Rohan Gupta, Yaniv Nikankin, Hadas Orgad, Nikhil Prakash, Anja Reusch, Aruna Sankaranarayanan, Shun Shao, Alessandro Stolfo, Martin Tutek, Amir Zur, David Bau, Yonatan Belinkov |
| 2025 | ICML | AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders. | Zhengxuan Wu, Aryaman Arora, Atticus Geiger, Zheng Wang, Jing Huang, Dan Jurafsky, Christopher D. Manning, Christopher Potts |
| 2025 | ICML | How Do Transformers Learn Variable Binding in Symbolic Programs? | Yiwei Wu, Atticus Geiger, Raphal Millire |
| 2024 | ACL | RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations. | Jing Huang, Zhengxuan Wu, Christopher Potts, Mor Geva, Atticus Geiger |
| 2024 | EMNLP | Updating CLIP to Prefer Descriptions Over Captions. | Amir Zur, Elisa Kreiss, Karel D'Oosterlinck, Christopher Potts, Atticus Geiger |
| 2024 | ICLR | Is This the Subspace You Are Looking for? An Interpretability Illusion for Subspace Activation Patching. | Aleksandar Makelov, Georg Lange, Atticus Geiger, Neel Nanda |
| 2024 | NAACL | pyvene: A Library for Understanding and Improving PyTorch Models via Interventions. | Zhengxuan Wu, Atticus Geiger, Aryaman Arora, Jing Huang, Zheng Wang, Noah D. Goodman, Christopher D. Manning, Christopher Potts |
| 2023 | ACL | ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context Learning. | Jingyuan Selena She, Christopher Potts, Samuel R. Bowman, Atticus Geiger |
| 2023 | CogSci | A Semantics for Causing, Enabling, and Preventing Verbs Using Structural Causal Models. | Angela Cao, Atticus Geiger, Elisa Kreiss, Thomas Icard, Tobias Gerstenberg |
| 2023 | ICML | Causal Proxy Models for Concept-based Model Explanations. | Zhengxuan Wu, Karel D'Oosterlinck, Atticus Geiger, Amir Zur, Christopher Potts |
| 2022 | ICML | Inducing Causal Structure for Interpretable Neural Networks. | Atticus Geiger, Zhengxuan Wu, Hanson Lu, Josh Rozner, Elisa Kreiss, Thomas Icard, Noah D. Goodman, Christopher Potts |
| 2022 | NAACL | Causal Distillation for Language Models. | Zhengxuan Wu, Atticus Geiger, Joshua Rozner, Elisa Kreiss, Hanson Lu, Thomas Icard, Christopher Potts, Noah D. Goodman |
| 2021 | ACL | DynaSent: A Dynamic Benchmark for Sentiment Analysis. | Christopher Potts, Zhengxuan Wu, Atticus Geiger, Douwe Kiela |
| 2021 | NAACL | Dynabench: Rethinking Benchmarking in NLP. | Douwe Kiela, Max Bartolo, Yixin Nie, Divyansh Kaushik, Atticus Geiger, Zhengxuan Wu, Bertie Vidgen, Grusha Prasad, Amanpreet Singh, Pratik Ringshia, Zhiyi Ma, Tristan Thrush, Sebastian Riedel, Zeerak Waseem, Pontus Stenetorp, Robin Jia, Mohit Bansal, Christopher Potts, Adina Williams |
| 2020 | CogSci | Relational reasoning and generalization using non-symbolic neural networks. | Atticus Geiger, Alexandra Carstensen, Michael C. Frank, Christopher Potts |
| 2019 | EMNLP | Posing Fair Generalization Tasks for Natural Language Inference. | Atticus Geiger, Ignacio Cases, Lauri Karttunen, Christopher Potts |
| 2019 | NAACL | Recursive Routing Networks: Learning to Compose Modules for Language Understanding. | Ignacio Cases, Clemens Rosenbaum, Matthew Riemer, Atticus Geiger, Tim Klinger, Alex Tamkin, Olivia Li, Sandhini Agarwal, Joshua D. Greene, Dan Jurafsky, Christopher Potts, Lauri Karttunen |