| 2021 | How Effective is Continuous Integration in Indicating Single-Statement Bugs? | Jasmine Latendresse, Rabe Abdalkareem, Diego Elias Costa, Emad Shihab |
| 2021 | How Do Software Developers Use GitHub Actions to Automate Their Workflows? | Timothy Kinsman, Mairieli Santos Wessel, Marco Aurlio Gerosa, Christoph Treude |
| 2021 | Denchmark: A Bug Benchmark of Deep Learning-related Software. | Misoo Kim, Youngkyoung Kim, Eunseok Lee |
| 2021 | S3M: Siamese Stack (Trace) Similarity Measure. | Aleksandr Khvorov, Roman Vasiliev, George A. Chernishev, Irving Muller Rodrigues, Dmitrij V. Koznov, Nikita Povarov |
| 2021 | PySStuBs: Characterizing Single-Statement Bugs in Popular Open-Source Python Projects. | Arthur V. Kamienski, Luisa Palechor, Cor-Paul Bezemer, Abram Hindle |
| 2021 | Practitioners' Perceptions of the Goals and Visual Explanations of Defect Prediction Models. | Jirayus Jiarpakdee, Chakkrit Tantithamthavorn, John C. Grundy |
| 2021 | Automatically Selecting Follow-up Questions for Deficient Bug Reports. | Mia Mohammad Imran, Agnieszka Ciborowska, Kostadin Damevski |
| 2021 | The Secret Life of Hackathon Code Where does it come from and where does it go? | Ahmed Imam, Tapajit Dey, Alexander Nolte, Audris Mockus, James D. Herbsleb |
| 2021 | Tracking Hackathon Code Creation and Reuse. | Ahmed Imam, Tapajit Dey |
| 2021 | On the Effectiveness of Deep Vulnerability Detectors to Simple Stupid Bug Detection. | Jiayi Hua, Haoyu Wang |
| 2021 | What Code Is Deliberately Excluded from Test Coverage and Why? | Andr C. Hora |
| 2021 | Googling for Software Development: What Developers Search For and What They Find. | Andr C. Hora |
| 2021 | A Traceability Dataset for Open Source Systems. | Mouna Hammoudi, Christoph Mayr-Dorn, Atif Mashkoor, Alexander Egyed |
| 2021 | A Replication Study on the Usability of Code Vocabulary in Predicting Flaky Tests. | Guillaume Haben, Sarra Habchi, Mike Papadakis, Maxime Cordy, Yves Le Traon |
| 2021 | gambit - An Open Source Name Disambiguation Tool for Version Control Systems. | Christoph Gote, Christian Zingg |
| 2021 | On the Naturalness and Localness of Software Logs. | Sina Gholamian, Paul A. S. Ward |
| 2021 | Leveraging Models to Reduce Test Cases in Software Repositories. | Golnaz Gharachorlu, Nick Sumner |
| 2021 | Waiting around or job half-done? Sentiment in self-admitted technical debt. | Gianmarco Fucci, Nathan Cassee, Fiorella Zampetti, Nicole Novielli, Alexander Serebrenik, Massimiliano Di Penta |
| 2021 | On Improving Deep Learning Trace Analysis with System Call Arguments. | Quentin Fournier, Daniel Aloise, Seyed Vahid Azhari, Franois Tetreault |
| 2021 | Escaping the Time Pit: Pitfalls and Guidelines for Using Time-Based Git Data. | Samuel W. Flint, Jigyasa Chauhan, Robert Dyer |
| 2021 | The Wonderless Dataset for Serverless Computing. | Nafise Eskandani, Guido Salvaneschi |
| 2021 | Revisiting Dockerfiles in Open Source Software Over Time. | Kalvin Eng, Abram Hindle |
| 2021 | Duets: A Dataset of Reproducible Pairs of Java Library-Clients. | Thomas Durieux, Csar Soto-Valero, Benoit Baudry |
| 2021 | An Empirical Study of OSS-Fuzz Bugs. | Zhen Yu Ding, Claire Le Goues |
| 2021 | Sampling Projects in GitHub for MSR Studies. | Ozren Dabic, Emad Aghajani, Gabriele Bavota |