| 2024 | ICLR | Identifying the Risks of LM Agents with an LM-Emulated Sandbox. | Yangjun Ruan, Honghua Dong, Andrew Wang, Silviu Pitis, Yongchao Zhou, Jimmy Ba, Yann Dubois, Chris J. Maddison, Tatsunori Hashimoto |
| 2023 | ICLR | Large Language Models are Human-Level Prompt Engineers. | Yongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, Jimmy Ba |
| 2020 | AAAI | Fixed-Horizon Temporal Difference Methods for Stable Reinforcement Learning. | Kristopher De Asis, Alan Chan, Silviu Pitis, Richard S. Sutton, Daniel Graves |
| 2020 | ICLR | An Inductive Bias for Distances: Neural Nets that Respect the Triangle Inequality. | Silviu Pitis, Harris Chan, Kiarash Jamali, Jimmy Ba |
| 2020 | ICML | Maximum Entropy Gain Exploration for Long Horizon Multi-goal Reinforcement Learning. | Silviu Pitis, Harris Chan, Stephen Zhao, Bradly C. Stadie, Jimmy Ba |
| 2019 | AAAI | Rethinking the Discount Factor in Reinforcement Learning: A Decision Theoretic Approach. | Silviu Pitis |
| 2018 | AAAI | Source Traces for Temporal Difference Learning. | Silviu Pitis |
| 2017 | ICAIL | Methods for retrieving alternative contract language using a prototype. | Silviu Pitis |