| 2018 | Understanding the usage, impact, and adoption of non-OSI approved licenses. | Rmulo Manciola Meloca, Gustavo Pinto, Leonardo Baiser, Marco Mattos, Ivanilton Polato, Igor Scaliante Wiese, Daniel M. Germn |
| 2018 | 50K-C: a dataset of compilable, and compiled, Java projects. | Pedro Martins, Rohan Achar, Cristina V. Lopes |
| 2018 | Public git archive: a big code dataset for all. | Vadim Markovtsev, Waren Long |
| 2018 | Natural language or not (NLON): a package for software engineering text analysis pipeline. | Mika V. Mntyl, Fabio Calefato, Malick Claes |
| 2018 | 500+ times faster than deep learning: a case study exploring faster methods for text mining stackoverflow. | Suvodeep Majumder, Nikhila Balaji, Katie Brey, Wei Fu, Tim Menzies |
| 2018 | The Android update problem: an empirical study. | Mehran Mahmoudi, Sarah Nadi |
| 2018 | Automatic classification of software artifacts in open-source applications. | Yuzhan Ma, Sarah Fakhoury, Michael Christensen, Venera Arnaoudova, Waleed Zogaan, Mehdi Mirakhorli |
| 2018 | Characterising deprecated Android APIs. | Li Li, Jun Gao, Tegawend F. Bissyand, Lei Ma, Xin Xia, Jacques Klein |
| 2018 | Exploring the use of automated API migrating techniques in practice: an experience report on Android. | Maxime Lamothe, Weiyi Shang |
| 2018 | An evaluation of open-source software microbenchmark suites for continuous performance assessment. | Christoph Laaber, Philipp Leitner |
| 2018 | Mining the mind, minding the mine: grand challenges in comprehension and mining. | Amy J. Ko |
| 2018 | Mining and extraction of personal software process measures through IDE interaction logs. | Alireza Joonbakhsh, Ashkan Sami |
| 2018 | The hidden cost of code completion: understanding the impact of the recommendation-list length on its efficiency. | Xianhao Jin, Francisco Servant |
| 2018 | A search system for mathematical expressions on software binaries. | Ridhi Jain, Sai Prathik, Venkatesh Vinayakarao, Rahul Purandare |
| 2018 | Prevalence of confusing code in software projects: atoms of confusion in the wild. | Dan Gopstein, Hongwei Henry Zhou, Phyllis G. Frankl, Justin Cappos |
| 2018 | VulinOSS: a dataset of security vulnerabilities in open-source systems. | Antonios Gkortzis, Dimitris Mitropoulos, Diomidis Spinellis |
| 2018 | What are your programming language's energy-delay implications? | Stefanos Georgiou, Maria Kechagia, Panos Louridas, Diomidis Spinellis |
| 2018 | A graph-based dataset of commit history of real-world Android apps. | Franz-Xaver Geiger, Ivano Malavolta, Luca Pascarella, Fabio Palomba, Dario Di Nucci, Alberto Bacchelli |
| 2018 | Jbench: a dataset of data races for concurrency testing. | Jian Gao, Xin Yang, Yu Jiang, Han Liu, Weiliang Ying, Xian Zhang |
| 2018 | Bayesian hierarchical modelling for tailoring metric thresholds. | Neil A. Ernst |
| 2018 | Word embeddings for the software engineering domain. | Vasiliki Efstathiou, Christos Chatzilenas, Diomidis Spinellis |
| 2018 | On the impact of security vulnerabilities in the npm package dependency network. | Alexandre Decan, Tom Mens, Eleni Constantinou |
| 2018 | Large-scale analysis of the co-commit patterns of the active developers in github's top repositories. | Eldan Cohen, Mariano P. Consens |
| 2018 | Towards automatically identifying paid open source developers. | Malick Claes, Mika Mntyl, Miikka Kuutila, Umar Farooq |
| 2018 | Detecting and characterizing developer behavior following opportunistic reuse of code snippets from the web. | Agnieszka Ciborowska, Nicholas A. Kraft, Kostadin Damevski |