| 2022 | Microsoft CloudMine: Data Mining for the Executive Order on Improving the Nation's Cybersecurity. | Kim Herzig, Luke Ghostling, Maximilian Grothusmann, Sascha Just, Nora Huang, Alan Klimowski, Yashasvini Ramkumar, Myles McLeroy, Kivan Muslu, Hitesh Sajnani, Varsha Vadaga |
| 2022 | GitRank: A Framework to Rank GitHub Repositories. | Niranjan Hasabnis |
| 2022 | Operationalizing Threats to MSR Studies by Simulation-Based Testing. | Johannes Hrtel, Ralf Lmmel |
| 2022 | A Large-Scale Comparison of Python Code in Jupyter Notebooks and Scripts. | Konstantin Grotov, Sergey Titov, Vladimir Sotnikov, Yaroslav Golubev, Timofey Bryksin |
| 2022 | Does Configuration Encoding Matter in Learning Software Performance? An Empirical Study on Encoding Schemes. | Jingzhi Gong, Tao Chen |
| 2022 | Quid Pro Quo: An Exploration of Reciprocity in Code Review. | Carlos Gavidia-Calderon, DongGyun Han, Amel Bennaceur |
| 2022 | Evaluating the effectiveness of local explanation methods on source code-based defect prediction models. | Yuxiang Gao, Yi Zhu, Qiao Yu |
| 2022 | Does This Apply to Me? An Empirical Study of Technical Context in Stack Overflow. | Akalanka Galappaththi, Sarah Nadi, Christoph Treude |
| 2022 | LineVul: A Transformer-based Line-Level Vulnerability Prediction. | Michael Fu, Chakkrit Tantithamthavorn |
| 2022 | Inspect4py: A Knowledge Extraction Framework for Python Code Repositories. | Rosa Filgueira, Daniel Garijo |
| 2022 | How heated is it? Understanding GitHub locked issues. | Isabella Ferreira, Bram Adams, Jinghui Cheng |
| 2022 | Replicating Data Pipelines with GrimoireLab. | Kalvin Eng, Hareem Sahar |
| 2022 | LAGOON: An Analysis Tool for Open Source Communities. | Sourya Dey, Walt Woods |
| 2022 | Detecting Privacy-Sensitive Code Changes with Language Modeling. | Gkalp Demirci, Vijayaraghavan Murali, Imad Ahmad, Rajeev Rao, Gareth Ari Aye |
| 2022 | FixJS: A Dataset of Bug-fixing JavaScript Commits. | Viktor Csuvik, Lszl Vidcs |
| 2022 | Noisy Label Learning for Security Defects. | Roland Croft, Muhammad Ali Babar, Huaming Chen |
| 2022 | To What Extent do Deep Learning-based Code Recommenders Generate Predictions by Cloning Code from the Training Set? | Matteo Ciniselli, Luca Pascarella, Gabriele Bavota |
| 2022 | An Empirical Study on Maintainable Method Size in Java. | Shaiful Alam Chowdhury, Gias Uddin, Reid Holmes |
| 2022 | Bot Detection in GitHub Repositories. | Natarajan Chidambaram, Pooya Rostami Mazrae |
| 2022 | Empirical Standards for Repository Mining. | Preetha Chatterjee, Tushar Sharma, Paul Ralph |
| 2022 | Vul4J: A Dataset of Reproducible Java Vulnerabilities Geared Towards the Study of Program Repair Techniques. | Quang-Cuong Bui, Riccardo Scandariato, Nicols E. Daz Ferreyra |
| 2022 | Extracting Corrective Actions from Code Repositories. | Yegor Bugayenko, Kirill Daniakin, Mirko Farina, Firas Jolha, Artem V. Kruglov, Giancarlo Succi, Witold Pedrycz |
| 2022 | Automatically Prioritizing and Assigning Tasks from Code Repositories in Puzzle Driven Development. | Yegor Bugayenko, Ayomide Bakare, Arina Cheverda, Mirko Farina, Artem V. Kruglov, Yaroslav Plaksin, Giancarlo Succi, Witold Pedrycz |
| 2022 | DaSEA - A Dataset for Software Ecosystem Analysis. | Petya Buchkova, Joakim Hey Hinnerskov, Kasper Olsen, Rolf-Helge Pfeiffer |
| 2022 | To Type or Not to Type? A Systematic Comparison of the Software Quality of JavaScript and TypeScript Applications on GitHub. | Justus Bogner, Manuel Merkel |