| 2016 | Anatomy of microarchitecture-level reliability assessment: Throughput and accuracy. | Athanasios Chatzidimitriou, Dimitris Gizopoulos |
| 2016 | NyuziRaster: Optimizing rasterizer performance and energy in the Nyuzi open source GPU. | Jeff Bush, Mohammad A. Khasawneh, Khaled Z. Mahmoud, Timothy N. Miller |
| 2016 | Agave: A benchmark suite for exploring the complexities of the Android software stack. | Martin K. Brown, Zachary Yannes, Michael Lustig, Mazdak Sanati, Sally A. McKee, Gary S. Tyson, Steven K. Reinhardt |
| 2016 | MLC PCM main memory with accelerated read. | Mohammad Arjomand, Amin Jadidi, Mahmut T. Kandemir, Anand Sivasubramaniam, Chita R. Das |
| 2016 | GSI: A GPU Stall Inspector to characterize the sources of memory stalls for tightly coupled GPUs. | Johnathan Alsop, Matthew D. Sinclair, Rakesh Komuravelli, Sarita V. Adve |
| 2016 | DVFS performance prediction for managed multithreaded applications. | Shoaib Akram, Jennifer B. Sartor, Lieven Eeckhout |
| 2016 | An automated framework for characterizing and subsetting GPGPU workloads. | Vignesh Adhinarayanan, Wu-chun Feng |
| 2016 | Characterization and architectural implications of big data workloads. | Lei Wang, Rui Ren, Jianfeng Zhan, Zhen Jia |
| 2015 | Graph-matching-based simulation-region selection for multiple binaries. | Charles Yount, Harish Patil, Mohammad S. Islam, Aditya Srikanth |
| 2015 | DRAW: investigating benefits of adaptive fetch group size on GPU. | Myung Kuk Yoon, Yunho Oh, Sangpil Lee, Seung-Hun Kim, Deokho Kim, Won Woo Ro |
| 2015 | Self-monitoring overhead of the Linux perf_ event performance counter interface. | Vincent M. Weaver |
| 2015 | Emulating cache organizations on real hardware using performance cloning. | Yipeng Wang, Yan Solihin |
| 2015 | Estimation-based profiling for code placement optimization in sensor network programs. | Lipeng Wan, Qing Cao, Wenjun Zhou |
| 2015 | QTrace: a framework for customizable full system instrumentation. | Xin Tong, Andreas Moshovos |
| 2015 | Micro-architecture independent analytical processor performance and power modeling. | Sam Van den Steen, Sander De Pestel, Moncef Mechri, Stijn Eyerman, Trevor E. Carlson, David Black-Schaffer, Erik Hagersten, Lieven Eeckhout |
| 2015 | Eliminating on-chip traffic waste: are we there yet? | Robert Smolinski, Rakesh Komuravelli, Hyojin Sung, Sarita V. Adve |
| 2015 | Can RDMA benefit online data processing workloads on memcached and MySQL? | Dipti Shankar, Xiaoyi Lu, Jithin Jose, Md. Wasi-ur-Rahman, Nusrat S. Islam, Dhabaleswar K. Panda |
| 2015 | Message from the program chair. | Jose Renau |
| 2015 | Factors affecting scalability of multithreaded Java applications on manycore systems. | Junjie Qian, Du Li, Witawas Srisa-an, Hong Jiang, Sharad C. Seth |
| 2015 | Micro-architecture independent branch behavior characterization. | Sander De Pestel, Stijn Eyerman, Lieven Eeckhout |
| 2015 | DELPHI: a framework for RTL-based architecture design evaluation using DSENT models. | Michael Papamichael, Cagla Cakir, Chen Sun, Chia-Hsin Owen Chen, James C. Hoe, Ken Mai, Li-Shiuan Peh, Vladimir Stojanovic |
| 2015 | A modeling framework for reuse distance-based estimation of cache performance. | Xiaoyue Pan, Bengt Jonsson |
| 2015 | DNOC: an accurate and fast virtual channel and deflection routing network-on-chip simulator. | Gadi Oxman, Shlomo Weiss |
| 2015 | Characterization and cross-platform analysis of high-throughput accelerators. | Keitaro Oka, Wenhao Jia, Margaret Martonosi, Koji Inoue |
| 2015 | Hierarchical cycle accounting: a new method for application performance tuning. | Andrzej Nowak, David Levinthal, Willy Zwaenepoel |