| 2025 | SMATCH-M-LLM: Semantic Similarity in Metamodel Matching With Large Language Models. | Nafisa Ahmed, Hin Chi Kwok, Mohammad Hamdaqa, Wesley K. G. Assuno |
| 2025 | Can LLMs Replace Manual Annotation of Software Engineering Artifacts? | Toufique Ahmed, Premkumar T. Devanbu, Christoph Treude, Michael Pradel |
| 2025 | SPRINT: An Assistant for Issue Report Management. | Ahmed Adnan, Antu Saha, Oscar Chaparro |
| 2025 | RepoChat: An LLM-Powered Chatbot for GitHub Repository Question-Answering. | Samuel Abedu, Laurine Menneron, SayedHassan Khatoonabadi, Emad Shihab |
| 2025 | pyMethods2Test: A Dataset of Python Tests Mapped to Focal Methods. | Idriss Abdelmadjid, Robert Dyer |
| 2024 | Enhancing Performance Bug Prediction Using Performance Code Metrics. | Guoliang Zhao, Stefanos Georgiou, Ying Zou, Safwat Hassan, Derek Truong, Toby Corbin |
| 2024 | Does Generative AI Generate Smells Related to Container Orchestration?: An Exploratory Study with Kubernetes Manifests. | Yue Zhang, Rachel Meredith, Wilson Reeves, Julia Coriolano, Muhammad Ali Babar, Akond Rahman |
| 2024 | A Mutation-Guided Assessment of Acceleration Approaches for Continuous Integration: An Empirical Study of YourBase. | Zhili Zeng, Tao Xiao, Maxime Lamothe, Hideaki Hata, Shane McIntosh |
| 2024 | MalwareBench: Malware samples are not enough. | Nusrat Zahan, Philipp Burckhardt, Mikola Lysenko, Feross Aboukhadijeh, Laurie A. Williams |
| 2024 | DevGPT: Studying Developer-ChatGPT Conversations. | Tao Xiao, Christoph Treude, Hideaki Hata, Kenichi Matsumoto |
| 2024 | ChatGPT Chats Decoded: Uncovering Prompt Patterns for Superior Solutions in Software Development Lifecycle. | Liangxuan Wu, Yanjie Zhao, Xinyi Hou, Tianming Liu, Haoyu Wang |
| 2024 | A Large-Scale Empirical Study of Open Source License Usage: Practices and Challenges. | Jiaqi Wu, Lingfeng Bao, Xiaohu Yang, Xin Xia, Xing Hu |
| 2024 | CodeLL: A Lifelong Learning Dataset to Support the Co-Evolution of Data and Language Models of Code. | Martin Weyssow, Claudio Di Sipio, Davide Di Ruscio, Houari A. Sahraoui |
| 2024 | Keep Me Updated: An Empirical Study on Embedded JavaScript Engines in Android Apps. | Elliott Wen, Jiaxiang Zhou, Xiapu Luo, Giovanni Russello, Jens Dietrich |
| 2024 | Global Prosperity or Local Monopoly? Understanding the Geography of App Popularity. | Liu Wang, Conghui Zheng, Haoyu Wang, Xiapu Luo, Gareth Tyson, Yi Wang, Shangguang Wang |
| 2024 | Estimating Usage Of Open Source Projects. | Sophia Vargas, Georg J. P. Link, JaYoung Lee |
| 2024 | Unveiling ChatGPT's Usage in Open Source Projects: A Mining-based Study. | Rosalia Tufano, Antonio Mastropaolo, Federica Pepe, Ozren Dabic, Massimiliano Di Penta, Gabriele Bavota |
| 2024 | A Dataset of Atoms of Confusion in the Android Open Source Project. | Davi Tabosa, Oton Pinheiro, Lincoln S. Rocha, Windson Viana |
| 2024 | SATDAUG - A Balanced and Augmented Dataset for Detecting Self-Admitted Technical Debt. | Edi Sutoyo, Andrea Capiluppi |
| 2024 | Questioning the Questions We Ask About the Impact of AI on Software Engineering : MSR 2024 Keynote. | Margaret-Anne D. Storey |
| 2024 | Comparing Apples to Androids: Discovery, Retrieval, and Matching of iOS and Android Apps for Cross-Platform Analyses. | Magdalena Steinbck, Jakob Bleier, Mikka Rainer, Tobias Urban, Christine Utz, Martina Lindorfer |
| 2024 | The PIPr Dataset of Public Infrastructure as Code Programs. | Daniel Sokolowski, David Spielmann, Guido Salvaneschi |
| 2024 | GitBug-Java: A Reproducible Benchmark of Recent Java Bugs. | Andr Silva, Nuno Saavedra, Martin Monperrus |
| 2024 | On the Anatomy of Real-World R Code for Static Analysis. | Florian Sihler, Lukas Pietzschmann, Raphael Straub, Matthias Tichy, Andor Diera, Abdelhalim Hafedh Dahou |
| 2024 | Quality Assessment of ChatGPT Generated Code and their Use by Developers. | Mohammed Latif Siddiq, Lindsay Roney, Jiahao Zhang, Joanna C. S. Santos |