| 2011 | Optimizing the datacenter for data-centric workloads. | Stijn Polfliet, Frederick Ryckbosch, Lieven Eeckhout |
| 2011 | MDR: performance model driven runtime for heterogeneous parallel platforms. | Jacques A. Pienaar, Anand Raghunathan, Srimat T. Chakradhar |
| 2011 | Controlling cache utilization of HPC applications. | Swann Perarnau, Marc Tchiboukdjian, Guillaume Huard |
| 2011 | SRC: FenixOS - a research operating system focused on high scalability and reliability. | Stavros Passas, Sven Karlsson |
| 2011 | F | Jin Ouyang, Chuan Yang, Dimin Niu, Yuan Xie, Zhiwen Liu |
| 2011 | Automatic SIMD vectorization of fast fourier transforms for the larrabee and AVX instruction sets. | Daniel S. McFarlin, Volodymyr Arbatov, Franz Franchetti, Markus Pschel |
| 2011 | Poster: implications of merging phases on scalability of multi-core architectures. | Madhavan Manivannan, Ben H. H. Juurlink, Per Stenstrm |
| 2011 | SRC: an automatic code overlaying technique for multicores with explicitly-managed memory hierarchies. | Choonki Jang |
| 2011 | An execution strategy and optimized runtime support for parallelizing irregular reductions on modern GPUs. | Xin Huo, Vignesh T. Ravi, Wenjing Ma, Gagan Agrawal |
| 2011 | Performance impact and interplay of SSD parallelism through advanced commands, allocation strategy and data granularity. | Yang Hu, Hong Jiang, Dan Feng, Lei Tian, Hao Luo, Shu Ping Zhang |
| 2011 | Optimizing throughput/power trade-offs in hardware transactional memory using DVFS and intelligent scheduling. | Clay Hughes, Tao Li |
| 2011 | Generic topology mapping strategies for large-scale parallel architectures. | Torsten Hoefler, Marc Snir |
| 2011 | Challenges and opportunities in renewable energy and energy efficiency. | Steven W. Hammond |
| 2011 | Using GPUs to compute large out-of-card FFTs. | Liang Gu, Jakob Siegel, Xiaoming Li |
| 2011 | Performance modeling as the key to extreme scale computing. | William D. Gropp |
| 2011 | ZEBRA: a data-centric, hybrid-policy hardware transactional memory design. | J. Rubn Titos Gil, Anurag Negi, Manuel E. Acacio, Jos M. Garca, Per Stenstrm |
| 2011 | Modeling the performance of an algebraic multigrid cycle on HPC platforms. | Hormozd Gahvari, Allison H. Baker, Martin Schulz, Ulrike Meier Yang, Kirk E. Jordan, William Gropp |
| 2011 | SRC: virtual i/o caching: dynamic storage cache management for concurrent workloads. | Michael R. Frasca, Ramya Prabhakar |
| 2011 | SRC: facilitating efficient parallelization of information storage and retrieval on large data sets. | Steven Feldman |
| 2011 | Cost-effectively offering private buffers in SoCs and CMPs. | Zhen Fang, Li Zhao, Ravishankar R. Iyer, Carlos Flores Fajardo, German Fabila Garcia, Seung Eun Lee, Bin Li, Steve R. King, Xiaowei Jiang, Srihari Makineni |
| 2011 | SRC: Damaris - using dedicated i/o cores for scalable post-petascale HPC simulations. | Matthieu Dorier |
| 2011 | High performance linpack benchmark: a fault tolerant implementation without checkpointing. | Teresa Davies, Christer Karlsson, Hui Liu, Chong Ding, Zizhong Chen |
| 2011 | SRC: soft error detection and recovery for high performance linpack. | Teresa Davies, Zizhong Chen |
| 2011 | SRC: automatic extraction of SST/macro skeleton models. | Amruth Rudraiah Dakshinamurthy |
| 2011 | A QHD-capable parallel H.264 decoder. | Chi Ching Chi, Ben H. H. Juurlink |