| 2015 | EMEURO: a framework for generating multi-purpose accelerators via deep learning. | Lawrence C. McAfee, Kunle Olukotun |
| 2015 | Data provenance tracking for concurrent programs. | Brandon Lucia, Luis Ceze |
| 2015 | Automatic data placement into GPU on-chip memory resources. | Chao Li, Yi Yang, Zhen Lin, Huiyang Zhou |
| 2015 | Exploiting Dynamic Parallelism to Efficiently Support Irregular Nested Loops on GPUs. | Da Li, Hancheng Wu, Michela Becchi |
| 2015 | A graph-based higher-order intermediate representation. | Roland Leia, Marcel Kster, Sebastian Hack |
| 2015 | Locality-centric thread scheduling for bulk-synchronous programming models on CPU architectures. | Hee-Seok Kim, Izzat El Hajj, John A. Stratton, Steven S. Lumetta, Wen-mei W. Hwu |
| 2015 | Improving GPGPU energy-efficiency through concurrent kernel execution and DVFS. | Qing Jiao, Mian Lu, Huynh Phung Huynh, Tulika Mitra |
| 2015 | Optimizing binary translation of dynamically generated code. | Byron Hawkins, Brian Demsky, Derek Bruening, Qin Zhao |
| 2015 | Checking correctness of code generator architecture specifications. | Niranjan Hasabnis, Rui Qiao, R. Sekar |
| 2015 | Hardware-Aware Automatic Code-Transformation to Support Compilers in Exploiting the Multi-Level Parallel Potential of Modern CPUs. | Dustin Feld, Thomas Soddemann, Michael Jnger, Sven Mallach |
| 2015 | Characterizing and enhancing global memory data coalescing on GPUs. | Naznin Fauzia, Louis-Nol Pouchet, P. Sadayappan |
| 2015 | A parallel abstract interpreter for JavaScript. | Kyle Dewey, Vineeth Kashyap, Ben Hardekopf |
| 2015 | Cycle-based Model to Evaluate Consistency Protocols within a Multi-protocol Compilation Tool-chain. | Hamza Chaker, Loc Cudennec, Safae Dahmani, Guy Gogniat, Martha Johanna Seplveda |
| 2015 | Runtime Support for Multiple Offload-Based Programming Models on Embedded Manycore Accelerators. | Alessandro Capotondi, Germain Haugou, Andrea Marongiu, Luca Benini |
| 2015 | HELIX-UP: relaxing program semantics to unleash parallelization. | Simone Campanoni, Glenn H. Holloway, Gu-Yeon Wei, David M. Brooks |
| 2015 | The Basic Building Blocks of Parallel Tasks. | Rohit Atre, Ali Jannesari, Felix Wolf |
| 2015 | Getting in control of your control flow with control-data isolation. | William Arthur, Ben Mehne, Reetuparna Das, Todd M. Austin |
| 2014 | DeltaPath: Precise and Scalable Calling Context Encoding. | Qiang Zeng, Junghwan Rhee, Hui Zhang, Nipun Arora, Guofei Jiang, Peng Liu |
| 2014 | Accelerating Dynamic Detection of Uses of Undefined Values with Static Value-Flow Analysis. | Ding Ye, Yulei Sui, Jingling Xue |
| 2014 | LeakChecker: Practical Static Memory Leak Detection for Managed Languages. | Dacong Yan, Guoqing Xu, Shengqian Yang, Atanas Rountev |
| 2014 | Software Transactional Memory for GPU Architectures. | Yunlong Xu, Rui Wang, Nilanjan Goswami, Tao Li, Lan Gao, Depei Qian |
| 2014 | Red Fox: An Execution Environment for Relational Query Processing on GPUs. | Haicheng Wu, Gregory F. Diamos, Tim Sheard, Molham Aref, Sean Baxter, Michael Garland, Sudhakar Yalamanchili |
| 2014 | Energy efficient data access techniques. | David B. Whalley |
| 2014 | Optimizing R VM: Allocation Removal and Path Length Reduction via Interpreter-level Specialization. | Haichuan Wang, Peng Wu, David A. Padua |
| 2014 | DrDebug: Deterministic Replay based Cyclic Debugging with Dynamic Slicing. | Yan Wang, Harish Patil, Cristiano Pereira, Gregory Lueck, Rajiv Gupta, Iulian Neamtiu |