| 2014 | GPU-Qin: A methodology for evaluating the error resilience of GPGPU applications. | Bo Fang, Karthik Pattabiraman, Matei Ripeanu, Sudhanva Gurumurthi |
| 2014 | A software based profiling method for obtaining speedup stacks on commodity multi-cores. | David Eklov, Nikos Nikoleris, Erik Hagersten |
| 2014 | Life lessons and datacenter performance analysis. | Amer Diwan |
| 2014 | Accelerating network-on-chip simulation via sampling. | Wenbo Dai, Natalie D. Enright Jerger |
| 2014 | Bridging the energy-efficiency gap in a future of massive data. | Fred Chong |
| 2014 | Optimized hardware for suboptimal software: The case for SIMD-aware benchmarks. | Juan M. Cebrian, Magnus Jahre, Lasse Natvig |
| 2014 | BarrierPoint: Sampled simulation of multi-threaded applications. | Trevor E. Carlson, Wim Heirman, Kenzo Van Craeynest, Lieven Eeckhout |
| 2014 | Reverse engineering of cache replacement policies in Intel microprocessors and their evaluation. | Andreas Abel, Jan Reineke |
| 2013 | Wall-clock based synchronization: A parallel simulation technology for cluster systems. | Xiaodong Zhu, Junmin Wu, Guoliang Chen, Tao Li |
| 2013 | A statistical machine learning based modeling and exploration framework for run-time cross-stack energy optimization. | Changshu Zhang, Arun Ravindran |
| 2013 | EMERALD: Characterization of emerging applications and algorithms for low-power devices. | Chuanjun Zhang, Glenn G. Ko, Jungwook Choi, Shang-nien Tsai, Minje Kim, Abner Guzmn-Rivera, Rob A. Rutenbar, Paris Smaragdis, Mi Sun Park, Vijaykrishnan Narayanan, Hongyi Xin, Onur Mutlu, Bin Li, Li Zhao, Mei Chen |
| 2013 | PAPI 5: Measuring power, energy, and the cloud. | Vincent M. Weaver, Daniel Terpstra, Heike McCraw, Matt Johnson, Kiran Kasichayanula, James Ralph, John Nelson, Philip Mucci, Tushar Mohan, Shirley Moore |
| 2013 | Non-determinism and overcount on modern hardware performance counter implementations. | Vincent M. Weaver, Daniel Terpstra, Shirley Moore |
| 2013 | XAMP: An eXtensible Analytical Model Platform. | Yipeng Wang, Yan Solihin |
| 2013 | Selecting benchmark combinations for the evaluation of multicore throughput. | Ricardo A. Velsquez, Pierre Michaud, Andr Seznec |
| 2013 | Quantifying the energy efficiency of FFT on heterogeneous platforms. | Yash Ukidave, Amir Kavyan Ziabari, Perhaad Mistry, Gunar Schirner, David R. Kaeli |
| 2013 | QTrace: An interface for customizable full system instrumentation. | Xin Tong, Jack Luo, Andreas Moshovos |
| 2013 | Understanding the implications of virtual machine management on processor microarchitecture design. | Xiufeng Sui, Tao Sun, Tao Li, Lixin Zhang |
| 2013 | Peta Thread Computing [Keynote I]. | Michael Shebanow |
| 2013 | ISA-independent workload characterization and its implications for specialized architectures. | Yakun Sophia Shao, David M. Brooks |
| 2013 | Interactive analysis of large distributed systems with scalable topology-based visualization. | Lucas Mello Schnorr, Arnaud Legrand, Jean-Marc Vincent |
| 2013 | Power/performance evaluation of energy efficient Ethernet (EEE) for High Performance Computing. | Karthikeyan P. Saravanan, Paul M. Carpenter, Alex Ramrez |
| 2013 | Trace filtering of multithreaded applications for CMP memory simulation. | Alejandro Rico, Alex Ramrez, Mateo Valero |
| 2013 | A mathematical hard disk timing model for full system simulation. | Benjamin S. Parsons, Vijay S. Pai |
| 2013 | Increasing the Transparent Page Sharing in Java. | Kazunori Ogata, Tamiya Onodera |