| 2016 | Performance analysis of accelerated biophysically-meaningful neuron simulations. | Georgios Smaragdos, Georgios Chatzikonstantis, Sofia Nomikou, Dimitrios Rodopoulos, Ioannis Sourdis, Dimitrios Soudris, Chris I. De Zeeuw, Christos Strydis |
| 2016 | Addressing service interruptions in memory with thread-to-rank assignment. | Manjunath Shevgoor, Rajeev Balasubramonian, Niladrish Chatterjee, Jung-Sik Kim |
| 2016 | Storage consolidation: Not always a panacea, but can we ease the pain? | Narges Shahidi, Mohammad Arjomand, Anand Sivasubramaniam, Mahmut T. Kandemir, Chita R. Das |
| 2016 | Splash-3: A properly synchronized benchmark suite for contemporary research. | Christos Sakalis, Carl Leonardsson, Stefanos Kaxiras, Alberto Ros |
| 2016 | Demystifying cloud benchmarking. | Tapti Palit, Yongming Shen, Michael Ferdman |
| 2016 | CoolSim: Eliminating traditional cache warming with fast, virtualized profiling. | Nikos Nikoleris, Andreas Sandberg, Erik Hagersten, Trevor E. Carlson |
| 2016 | A comprehensive performance analysis of HSA and OpenCL 2.0. | Saoni Mukherjee, Yifan Sun, Paul Blinzer, Amir Kavyan Ziabari, David R. Kaeli |
| 2016 | Message from the program chair. | Andreas Moshovos |
| 2016 | Characterizing Hadoop applications on microservers for performance and energy efficiency optimizations. | Maria Malik, Avesta Sasan, Rajiv V. Joshi, Setareh Rafatirah, Houman Homayoun |
| 2016 | Compositional model of coherence and NUMA effects for optimizing thread and data placement. | Hao Luo, Jacob Brock, Pengcheng Li, Chen Ding, Chencheng Ye |
| 2016 | FastCap: An efficient and fair algorithm for power capping in many-core systems. | Yanpei Liu, Guilherme Cox, Qingyuan Deng, Stark C. Draper, Ricardo Bianchini |
| 2016 | Characterization and bottleneck analysis of a 64-bit ARMv8 platform. | Michael A. Laurenzano, Ananta Tiwari, Allyson Cauble-Chantrenne, Adam Jundt, William A. Ward Jr., Roy L. Campbell, Laura Carrington |
| 2016 | MofySim: A mobile full-system simulation framework for energy consumption and performance analysis. | Minho Ju, Hyeonggyu Kim, Soontae Kim |
| 2016 | NoMali: Simulating a realistic graphics driver stack using a stub GPU. | Ren de Jong, Andreas Sandberg |
| 2016 | Elastic traces for fast and accurate system performance exploration. | Radhika Jagtap, Stephan Diestelhorst, Andreas Hansson |
| 2016 | JIT-assisted fast-forward embedding and instrumentation to enable fast, accurate, and agile simulation. | Berkin Ilbeyi, Christopher Batten |
| 2016 | Message from the general chair. | Erik Hagersten |
| 2016 | TaskPoint: Sampled simulation of task-based programs. | Thomas Grass, Alejandro Rico, Marc Casas, Miquel Moret, Eduard Ayguad |
| 2016 | X-Mem: A cross-platform and extensible memory characterization tool for the cloud. | Mark Gottscho, Sriram Govindan, Bikash Sharma, Mohammed Shoaib, Puneet Gupta |
| 2016 | Analyzing the energy-efficiency of sparse matrix multiplication on heterogeneous systems: A comparative study of GPU, Xeon Phi and FPGA. | Heiner Giefers, Peter W. J. Staar, Costas Bekas, Christoph Hagleitner |
| 2016 | OpenSoC Fabric: On-chip network generator. | Farzad Fatollahi-Fard, David Donofrio, George Michelogiannakis, John Shalf |
| 2016 | Evaluating asymmetric multiprocessing for mobile applications. | Songchun Fan, Benjamin C. Lee |
| 2016 | Interactive visualization of cross-layer performance anomalies in dynamic task-parallel applications and systems. | Andi Drebes, Antoniu Pop, Karine Heydemann, Albert Cohen |
| 2016 | AnyCore: A synthesizable RTL model for exploring and fabricating adaptive superscalar cores. | Rangeen Basu Roy Chowdhury, Anil K. Kannepalli, Sungkwan Ku, Eric Rotenberg |
| 2016 | Workload characterization and optimization of TPC-H queries on Apache Spark. | Tatsuhiro Chiba, Tamiya Onodera |