Skip to content

Alessandro Stolfo

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

10

Venues

6

Active years

2023–2025

Best venue rank

A*

Where they publish

Papers

10 indexed papers, newest first.

YearVenueTitleAuthors
2025EMNLPProbing for Arithmetic Errors in Language Models.Yucheng Sun, Alessandro Stolfo, Mrinmaya Sachan
2025ICLRImproving Instruction-Following in Language Models through Activation Steering.Alessandro Stolfo, Vidhisha Balachandran, Safoora Yousefi, Eric Horvitz, Besmira Nushi
2025ICMLMIB: A Mechanistic Interpretability Benchmark.Aaron Mueller, Atticus Geiger, Sarah Wiegreffe, Dana Arad, Ivn Arcuschin, Adam Belfki, Yik Siu Chan, Jaden Fried Fiotto-Kaufman, Tal Haklay, Michael Hanna, Jing Huang, Rohan Gupta, Yaniv Nikankin, Hadas Orgad, Nikhil Prakash, Anja Reusch, Aruna Sankaranarayanan, Shun Shao, Alessandro Stolfo, Martin Tutek, Amir Zur, David Bau, Yonatan Belinkov
2024ICMLDo Language Models Exhibit the Same Cognitive Biases in Problem Solving as Human Learners?Andreas Opedal, Alessandro Stolfo, Haruki Shirakami, Ying Jiao, Ryan Cotterell, Bernhard Schlkopf, Abulhair Saparov, Mrinmaya Sachan
2024NAACLGroundedness in Retrieval-augmented Long-form Generation: An Empirical Study.Alessandro Stolfo
2023ACLDistilling Reasoning Capabilities into Smaller Language Models.Kumar Shridhar, Alessandro Stolfo, Mrinmaya Sachan
2023ACLA Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language Models.Alessandro Stolfo, Zhijing Jin, Kumar Shridhar, Bernhard Schlkopf, Mrinmaya Sachan
2023EACLLongtonotes: OntoNotes with Longer Coreference Chains.Kumar Shridhar, Nicholas Monath, Raghuveer Thirukovalluru, Alessandro Stolfo, Manzil Zaheer, Andrew McCallum, Mrinmaya Sachan
2023EMNLPTowards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models.Yifan Hou, Jiaoda Li, Yu Fei, Alessandro Stolfo, Wangchunshu Zhou, Guangtao Zeng, Antoine Bosselut, Mrinmaya Sachan
2023EMNLPA Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis.Alessandro Stolfo, Yonatan Belinkov, Mrinmaya Sachan