| 2015 | Neural acceleration for GPU throughput processors. | Amir Yazdanbakhsh, Jongse Park, Hardik Sharma, Pejman Lotfi-Kamran, Hadi Esmaeilzadeh |
| 2015 | Characterizing, modeling, and improving the QoE of mobile devices with low battery level. | Kaige Yan, Xingyao Zhang, Xin Fu |
| 2015 | Enabling coordinated register allocation and thread-level parallelism optimization for GPUs. | Xiaolong Xie, Yun Liang, Xiuhong Li, Yudong Wu, Guangyu Sun, Tao Wang, Dongrui Fan |
| 2015 | Automated Power and Latency Management in Heterogeneous 3D NoCs. | Awet Yemane Weldezion, Masoumeh Ebrahimi, Masoud Daneshtalab, Hannu Tenhunen |
| 2015 | Control flow coalescing on a hybrid dataflow/von Neumann GPGPU. | Dani Voitsechov, Yoav Etsion |
| 2015 | TimeTrader: exploiting latency tail to save datacenter energy for online search. | Balajee Vamanan, Hamza Bin Sohail, Jahangir Hasan, T. N. Vijaykumar |
| 2015 | The application slowdown model: quantifying and controlling the impact of inter-application interference at shared caches and main memory. | Lavanya Subramanian, Vivek Seshadri, Arnab Ghosh, Samira Manabi Khan, Onur Mutlu |
| 2015 | More is less: improving the energy efficiency of data movement via opportunistic use of sparse codes. | Yanwei Song, Engin Ipek |
| 2015 | Efficiently enforcing strong memory ordering in GPUs. | Abhayendra Singh, Shaizeen Aga, Satish Narayanasamy |
| 2015 | Efficient GPU synchronization without scopes: saying no to complex consistency models. | Matthew D. Sinclair, Johnathan Alsop, Sarita V. Adve |
| 2015 | Efficiently prefetching complex address patterns. | Manjunath Shevgoor, Sahil Koladiya, Rajeev Balasubramonian, Chris Wilkerson, Seth H. Pugsley, Zeshan Chishti |
| 2015 | Avoiding information leakage in the memory controller with fixed service policies. | Ali Shafiee, Akhila Gundu, Manjunath Shevgoor, Rajeev Balasubramonian, Mohit Tiwari |
| 2015 | The inner most loop iteration counter: a new dimension in branch history. | Andr Seznec, Joshua San Miguel, Jorge Albericio |
| 2015 | Gather-scatter DRAM: in-DRAM address translation to improve the spatial locality of non-unit strided accesses. | Vivek Seshadri, Thomas Mullins, Amirali Boroumand, Onur Mutlu, Phillip B. Gibbons, Michael A. Kozuch, Todd C. Mowry |
| 2015 | Long term parking (LTP): criticality-aware resource allocation in OOO processors. | Andreas Sembrant, Trevor E. Carlson, Erik Hagersten, David Black-Schaffer, Arthur Perais, Andr Seznec, Pierre Michaud |
| 2015 | ThyNVM: enabling software-transparent crash consistency in persistent memory systems. | Jinglei Ren, Jishen Zhao, Samira Manabi Khan, Jongmoo Choi, Yongwei Wu, Onur Mutlu |
| 2015 | A fast and accurate analytical technique to compute the AVF of sequential bits in a processor. | Steven Raasch, Arijit Biswas, Jon Stephan, Paul Racunas, Joel S. Emer |
| 2015 | Large pages and lightweight memory management in virtualized environments: can you have it both ways? | Binh Pham, Jn Vesel, Gabriel H. Loh, Abhishek Bhattacharjee |
| 2015 | DynaMOS: dynamic schedule migration for heterogeneous cores. | Shruti Padmanabha, Andrew Lukefahr, Reetuparna Das, Scott A. Mahlke |
| 2015 | Methodology to verify, debug and evaluate performances of NoC based interconnects. | Patrick Oury, Nick Heaton, Stewart Penman |
| 2015 | Border control: sandboxing accelerators. | Lena E. Olson, Jason Power, Mark D. Hill, David A. Wood |
| 2015 | Modeling the implications of DRAM failures and protection techniques on datacenter TCO. | Panagiota Nikolaou, Yiannakis Sazeides, Lorena Ndreu, Marios Kleanthous |
| 2015 | MORC: a manycore-oriented compressed cache. | Tri Minh Nguyen, David Wentzlaff |
| 2015 | The CRISP performance model for dynamic voltage and frequency scaling in a GPGPU. | Rajib Nath, Dean M. Tullsen |
| 2015 | Rethinking Memory System Design (along with Interconnects). | Onur Mutlu |