| 2026 | LLAMAFUZZ: Large Language Model Enhanced Greybox Fuzzing. | Hongxiang Zhang, Yuyang Rong, Yifeng He, Hao Chen |
| 2026 | Separating Valid from Invalid Inputs for a Digital Aircraft Design Tool. | Malte Christian Struck, Alexander Weinert, Andreas Schuster, Michael Felderer |
| 2026 | APITestGenie: Generating Web API Tests from Requirements and API Specifications with LLMs. | Andr Pereira, Bruno Lima, Joo Pascoal Faria |
| 2026 | Exploring Mocking Techniques for Managing External Dependencies in Service-Based Systems: A Mapping Study. | Benedito de Oliveira, Fernando Castor, Leo Fernandes, Samuel Amorim |
| 2026 | Improving Deep Learning Library Testing with Machine Learning. | Facundo Molina, M. M. Abid Naziri, Feiran (Alex) Qin, Alessandra Gorla, Marcelo d'Amorim |
| 2026 | A Unified Benchmark for Out-of-Distribution Detection for Autonomous Driving Systems. | Xiangyu Li, Jingyu Zhang, Jacky Keung, Xiaoxue Ma, Yihan Liao |
| 2026 | REST-at: An LLM-Based Tool for Automating Traceability between Requirements and Test Cases. | Nicole Leon-Quinstedt, Bao Lindgren, Mert Yurdakul, Francisco Gomes de Oliveira Neto |
| 2026 | ACT: Automated CPS Testing for Open-Source Robotic Platforms. | Aditya A. Krishnan, Donghoon Kim, Hokeun Kim |
| 2026 | Understanding Bug-Reproducing Tests: A First Empirical Study. | Andre Hora, Gordon Fraser |
| 2026 | L-SCALE: Locality-Sensitive Coverage for Automata LEarning. | Mark Leon Giraud, Bastian Engel, Lea Nasarek, Yannis Storrer, Philipp Takacs, Leon Philipp Wittemund |
| 2026 | Search-Based Fuzzing For RESTful APIs That Use MongoDB. | Hernan Ghianni, Man Zhang, Juan P. Galeotti, Andrea Arcuri |
| 2026 | HYDRA: A Hybrid Heuristic-Guided Deep Representation Architecture for Predicting Latent Zero-Day Vulnerabilities in Patched Functions. | Mohammad Farhad, Sabbir Rahman, Shuvalaxmi Dass |
| 2026 | Understanding and Detecting Platform-Specific Violations in Android Auto Apps. | Moshood Abiola Fakorede, Umar Farooq |
| 2026 | From Logs to Lessons: An Exploration of LLM-based Log Summarization for Debugging Automotive Software. | Anton Ekstrm, Hampus Rhedin Stam, Francisco Gomes de Oliveira Neto, Gregory Gay, Sabina Edenlund |
| 2026 | Software Testing Education in the LLM Era: Insights and Emerging Theory. | Samhitha Dwarakanath, Nathalia Nascimento, Everton Guimares |
| 2026 | A Framework for Similarity-based and Resource-aware Orchestration of End-to-End Test Cases. | Cristian Augusto, Antonia Bertolino, Guglielmo De Angelis, Claudio de la Riva, Francesca Lonetti, Jess Morn |
| 2026 | Testing Framework Migration with Large Language Models. | Altino Alves, Joo Eduardo Montandon, Andre Hora |
| 2026 | Understanding on the Edge: LLM-generated Boundary Test Explanations. | Sabinakhon Akbarova, Felix Dobslaw, Robert Feldt |
| 2025 | Simulink Mutation Testing using CodeBERT. | Jingfan Zhang, Delaram Ghobari, Mehrdad Sabetzadeh, Shiva Nejati |
| 2025 | A Taxonomy of Failures in Tool-Augmented LLMs. | Cailin Winston, Ren Just |
| 2025 | ASTRAL: Automated Safety Testing of Large Language Models. | Miriam Ugarte, Pablo Valle, Jos Antonio Parejo, Sergio Segura, Aitor Arrieta |
| 2025 | A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal Verification. | Norbert Tihanyi, Yiannis Charalambous, Ridhi Jain, Mohamed Amine Ferrag, Lucas C. Cordeiro |
| 2025 | An Adaptive Testing Approach Based on Field Data. | Samira Silva, Ricardo Caldas, Patrizio Pelliccione, Antonia Bertolino |
| 2025 | AsserT5: Test Assertion Generation Using a Fine-Tuned Code Language Model. | Severin Primbs, Benedikt Fein, Gordon Fraser |
| 2025 | Automated Test Generation for Integration Testing. | Elson Kurian, Giovanni Denaro, Pietro Braione, Luca Guglielmo |