| 2016 | Synchronization Debugging of Hybrid Parallel Programs. | Olaf Krzikalla, Ralph Mller-Pfefferkorn, Wolfgang E. Nagel |
| 2016 | Nasty-MPI: Debugging Synchronization Errors in MPI-3 One-Sided Applications. | Roger Kowalewski, Karl Frlinger |
| 2016 | Multi-objective Optimization Framework for VMI Distribution in Federated Cloud Repositories. | Dragi Kimovski, Nishant Saurabh, Sandi Gec, Vlado Stankovski, Radu Prodan |
| 2016 | Penalized Graph Partitioning for Static and Dynamic Load Balancing. | Tim Kiefer, Dirk Habich, Wolfgang Lehner |
| 2016 | Towards a Methodology to Form Microservices from Monolithic Ones. | Gabor Kecskemeti, Attila Kertsz, Attila Csaba Marosi |
| 2016 | Ultra-Fast Detection of Higher-Order Epistatic Interactions on GPUs. | Daniel Jnger, Christian Hundt, Jorge Gonzlez-Domnguez, Bertil Schmidt |
| 2016 | A Context-Aware Primitive for Nested Recursive Parallelism. | Herbert Jordan, Peter Thoman, Peter Zangerl, Thomas Heller, Thomas Fahringer |
| 2016 | Non-preemptive Scheduling with Setup Times: A PTAS. | Klaus Jansen, Felix Land |
| 2016 | Automatic Verification of Self-consistent MPI Performance Guidelines. | Sascha Hunold, Alexandra Carpen-Amarie, Felix Donatus Lbbe, Jesper Larsson Trff |
| 2016 | Automatic OpenCL Task Adaptation for Heterogeneous Architectures. | Pierre Huchant, Marie Christine Counilh, Denis Barthou |
| 2016 | Computational Considerations for a Global Human Well-Being Simulation. | Aaron Howell, Paul R. Brenner |
| 2016 | Towards the Next Generation of Large-Scale Network Archives. | Stijn Heldens, Ana Lucia Varbanescu, Wing Lung Ngai, Tim Hegeman, Alexandru Iosup |
| 2016 | A Massively-Parallel, Fault-Tolerant Solver for High-Dimensional PDEs. | Mario Heene, Alfredo Parra-Hinojosa, Hans-Joachim Bungartz, Dirk Pflger |
| 2016 | Effective Minimally-Invasive GPU Acceleration of Distributed Sparse Matrix Factorization. | Anshul Gupta, Natalia Gimelshein, Seid Koric, Steven C. Rennich |
| 2016 | Lightweight and Accurate Silent Data Corruption Detection in Ordinary Differential Equation Solvers. | Pierre-Louis Guhur, Hong Zhang, Tom Peterka, Emil M. Constantinescu, Franck Cappello |
| 2016 | Performance and Power-Aware Classification for Frequency Scaling of GPGPU Applications. | Joo Guerreiro, Aleksandar Ilic, Nuno Roma, Pedro Toms |
| 2016 | Seamless HPC Integration of Data-Intensive KNIME Workflows via UNICORE. | Richard Grunzke, Florian Jug, Bernd Schuller, Ren Jkel, Gene Myers, Wolfgang E. Nagel |
| 2016 | Multicore vs Manycore: The Energy Cost of Concurrency. | Martin Groen, Vincent Gramoli |
| 2016 | The Information Needed for Reproducing Shared Memory Experiments. | Vincent Gramoli |
| 2016 | Reducing Response Time with Preheated Caches. | Mathias Gottschlag, Frank Bellosa |
| 2016 | The ICARUS White Paper: A Scalable, Energy-Efficient, Solar-Powered HPC Center Based on Low Power GPUs. | Markus Geveler, Dirk Ribbrock, Daniel Donner, Hannes Ruelmann, Christoph Hppke, David Schneider, Daniel Tomaschewski, Stefan Turek |
| 2016 | Scheduling MapReduce Jobs Under Multi-round Precedences. | Dimitris Fotakis, Ioannis Milis, Orestis Papadigenopoulos, Vasilis Vassalos, Georgios Zois |
| 2016 | Improving Performance of Distributed Graph Traversals via Application-Aware Plug-In Work Scheduler. | Jesun Sahariar Firoz, Marcin Zalewski, Martina Barnas, Andrew Lumsdaine |
| 2016 | Resampling with Feedback - A New Paradigm of Using Workload Data for Performance Evaluation. | Dror G. Feitelson |
| 2016 | Improving Memory Accesses for Heterogeneous Parallel Multi-objective Feature Selection on EEG Classification. | Juan Jos Escobar, Julio Ortega, Jess Gonzlez, Miguel Damas |