| 2017 | Cross-ISA machine emulation for multicores. | Emilio G. Cota, Paolo Bonzini, Alex Benne, Luca P. Carloni |
| 2017 | Formalizing the concurrency semantics of an LLVM fragment. | Soham Chakraborty, Viktor Vafeiadis |
| 2017 | Taming warp divergence. | Jayvant Anantpur, R. Govindarajan |
| 2017 | Software prefetching for indirect memory accesses. | Sam Ainsworth, Timothy M. Jones |
| 2016 | Exploiting mixed SIMD parallelism by reducing data reorganization overhead. | Hao Zhou, Jingling Xue |
| 2016 | Atomicity violation checker for task parallel programs. | Adarsh Yoga, Santosh Nagarakatte |
| 2016 | gpucc: an open-source GPGPU compiler. | Jingyue Wu, Artem Belevich, Eli Bendersky, Mark Heffernan, Chris Leary, Jacques A. Pienaar, Bjarke Roune, Rob Springer, Xuetian Weng, Robert Hundt |
| 2016 | Towards automatic significance analysis for approximate computing. | Vassilis Vassiliadis, Jan Riehme, Jens Deussen, Konstantinos Parasyris, Christos D. Antonopoulos, Nikolaos Bellas, Spyros Lalis, Uwe Naumann |
| 2016 | Inference of peak density of indirect branches to detect ROP attacks. | Mateus Tymburib, Rubens E. A. Moreira, Fernando Magno Quinto Pereira |
| 2016 | Sparse flow-sensitive pointer analysis for multithreaded programs. | Yulei Sui, Peng Di, Jingling Xue |
| 2016 | An Image Processing Language: External and Shallow/Deep Embeddings. | Robert J. Stewart |
| 2016 | A basic linear algebra compiler for structured matrices. | Daniele G. Spampinato, Markus Pschel |
| 2016 | Industrial Application of Domain Specific Languages Combined with Formal Techniques. | Mathijs Schuts, Jozef Hooman |
| 2016 | StructSlim: a lightweight profiler to guide structure splitting. | Probir Roy, Xu Liu |
| 2016 | Symbolic range analysis of pointers. | Vitor Paisante, Maroua Maalej, Leonardo Barbosa e Oliveira, Laure Gonnord, Fernando Magno Quinto Pereira |
| 2016 | Communication-aware mapping of stream graphs for multi-GPU platforms. | Dong Nguyen, Jongeun Lee |
| 2016 | Are there Domain Specific Languages? | Greg Michaelson |
| 2016 | Portable and transparent software managed scheduling on accelerators for fair resource sharing. | Christos Margiolas, Michael F. P. O'Boyle |
| 2016 | Why So Many?: A Brief Tour of Haskell DSLs for Parallel Programming. | Patrick Maier |
| 2016 | Cheetah: detecting false sharing efficiently and effectively. | Tongping Liu, Xu Liu |
| 2016 | Automatic Generation of Code Analysis Tools: The CastQL Approach. | Christakis Lezos, Grigoris Dimitroulakos, Ioannis Latifis, Konstantinos Masselos |
| 2016 | IPAS: intelligent protection against silent output corruption in scientific applications. | Ignacio Laguna, Martin Schulz, David F. Richards, Jon Calhoun, Luke N. Olson |
| 2016 | Re-constructing high-level information for language-specific binary re-optimization. | Toshihiko Koju, Reid Copeland, Motohiro Kawahito, Moriyoshi Ohara |
| 2016 | NRG-loops: adjusting power from within applications. | Melanie Kambadur, Martha A. Kim |
| 2016 | Portable performance on asymmetric multicore processors. | Ivan Jibaja, Ting Cao, Stephen M. Blackburn, Kathryn S. McKinley |