| 2020 | The Software Heritage Graph Dataset: Large-scale Analysis of Public Software Development History. | Antoine Pietri, Diomidis Spinellis, Stefano Zacchiroli |
| 2020 | Determining the Intrinsic Structure of Public Software Development History. | Antoine Pietri, Guillaume Rousseau, Stefano Zacchiroli |
| 2020 | Forking Without Clicking: on How to Identify Software Repository Forks. | Antoine Pietri, Guillaume Rousseau, Stefano Zacchiroli |
| 2020 | What constitutes Software?: An Empirical, Descriptive Study of Artifacts. | Rolf-Helge Pfeiffer |
| 2020 | Developer-Driven Code Smell Prioritization. | Fabiano Pecorelli, Fabio Palomba, Foutse Khomh, Andrea De Lucia |
| 2020 | GitterCom: A Dataset of Open Source Developer Communications in Gitter. | Esteban Parra, Ashley Ellis, Sonia Haiduc |
| 2020 | Behind the Intents: An In-depth Empirical Study on Software Refactoring in Modern Code Review. | Matheus Paixo, Anderson G. Ucha, Ana Carla Bibiano, Daniel Oliveira, Alessandro Garcia, Jens Krinke, Emilio Arvonio |
| 2020 | The Impact of Dynamics of Collaborative Software Engineering on Introverts: A Study Protocol. | Ingrid Nunes, Christoph Treude, Fabio Calefato |
| 2020 | Can We Use SE-specific Sentiment Analysis Tools in a Cross-Platform Setting? | Nicole Novielli, Fabio Calefato, Davide Dongiovanni, Daniela Girardi, Filippo Lanubile |
| 2020 | An Empirical Study of Method Chaining in Java. | Tomoki Nakamaru, Tomomasa Matsunaga, Tetsuro Yamazaki, Soramichi Akiyama, Shigeru Chiba |
| 2020 | On the Prevalence, Impact, and Evolution of SQL Code Smells in Data-Intensive Systems. | Biruk Asmare Muse, Mohammad Masudur Rahman, Csaba Nagy, Anthony Cleve, Foutse Khomh, Giuliano Antoniol |
| 2020 | Using Others' Tests to Identify Breaking Updates. | Suhaib Mujahid, Rabe Abdalkareem, Emad Shihab, Shane McIntosh |
| 2020 | A Complete Set of Related Git Repositories Identified via Community Detection Approaches Based on Shared Commits. | Audris Mockus, Diomidis Spinellis, Zoe Kotti, Gabriel John Dusing |
| 2020 | Capture the Feature Flag: Detecting Feature Flags in Open-Source. | Jens Meinicke, Juan Hoyos, Bogdan Vasilescu, Christian Kstner |
| 2020 | RTPTorrent: An Open-source Dataset for Evaluating Regression Test Prioritization. | Toni Mattis, Patrick Rein, Falco Drsch, Robert Hirschfeld |
| 2020 | Traceability Support for Multi-Lingual Software Projects. | Yalin Liu, Jinfeng Lin, Jane Cleland-Huang |
| 2020 | AndroZooOpen: Collecting Large-scale Open Source Android Apps for the Research Community. | Pei Liu, Li Li, Yanjie Zhao, Xiaoyu Sun, John Grundy |
| 2020 | PUMiner: Mining Security Posts from Developer Question and Answer Websites with PU Learning. | Triet Huynh Minh Le, David Hin, Roland Croft, Muhammad Ali Babar |
| 2020 | TestRoutes: A Manually Curated Method Level Dataset for Test-to-Code Traceability. | Andrs Kicsi, Lszl Vidcs, Tibor Gyimthy |
| 2020 | A Study on the Accuracy of OCR Engines for Source Code Transcription from Programming Screencasts. | Abdulkarim Khormi, Mohammad Alahmadi, Sonia Haiduc |
| 2020 | How Often Do Single-Statement Bugs Occur?: The ManySStuBs4J Dataset. | Rafael-Michael Karampatsis, Charles Sutton |
| 2020 | From Innovations to Prospects: What Is Hidden Behind Cryptocurrencies? | Ang Jia, Ming Fan, Xi Xu, Di Cui, Wenying Wei, Zijiang Yang, Kai Ye, Ting Liu |
| 2020 | The Scent of Deep Learning Code: An Empirical Study. | Hadhemi Jebnoun, Houssem Ben Braiek, Mohammad Masudur Rahman, Foutse Khomh |
| 2020 | Boa Views: Easy Modularization and Sharing of MSR Analyses. | Che Shian Hung, Robert Dyer |
| 2020 | Large-Scale Manual Validation of Bugfixing Changes. | Steffen Herbold, Alexander Trautsch, Benjamin Ledel |