| 2011 | Large-Scale Simulator for Global Data Infrastructure Optimization. | Sergio Herrero-Lopez, John R. Williams, Abel Sanchez |
| 2011 | Automatic Task Re-organization in MapReduce. | Zhenhua Guo, Marlon E. Pierce, Geoffrey C. Fox, Mo Zhou |
| 2011 | Performance Analysis and Benchmarking of the Intel SCC. | Philipp Gschwandtner, Thomas Fahringer, Radu Prodan |
| 2011 | RDMA Based Replication of Multiprocessor Virtual Machines over High-Performance Interconnects. | Balazs Gerofi, Yutaka Ishikawa |
| 2011 | Parallel I/O Performance for Application-Level Checkpointing on the Blue Gene/P System. | Jing Fu, Misun Min, Robert Latham, Christopher D. Carothers |
| 2011 | AA-Dedupe: An Application-Aware Source Deduplication Approach for Cloud Backup Services in the Personal Computing Environment. | Yinjin Fu, Hong Jiang, Nong Xiao, Lei Tian, Fang Liu |
| 2011 | Performance Characterization and Optimization of Atomic Operations on AMD GPUs. | Marwa K. Elteir, Heshan Lin, Wu-chun Feng |
| 2011 | Multicore/GPGPU Portable Computational Kernels via Multidimensional Arrays. | H. Carter Edwards, Daniel Sunderland, Chris Amsler, Sam Mish |
| 2011 | High Performance Dense Linear System Solver with Soft Error Resilience. | Peng Du, Piotr Luszczek, Jack J. Dongarra |
| 2011 | Optimizing Network I/O Virtualization with Efficient Interrupt Coalescing and Virtual Receive Side Scaling. | Yaozu Dong, Dongxiao Xu, Yang Zhang, Guangdeng Liao |
| 2011 | Application I/O and Data Management. | William W. Dai |
| 2011 | Experience on Comparison of Operating Systems Scalability on the Multi-core Architecture. | Yan Cui, Yingxin Wang, Yu Chen, Yuanchun Shi |
| 2011 | Heterogeneous Cloud Computing. | Stephen P. Crago, Kyle Dunn, Patrick Eads, Lorin Hochstein, Dong-In Kang, Mikyung Kang, Devendra Modium, Karandeep Singh, Jinwoo Suh, John Paul Walters |
| 2011 | Performance Behavior Prediction Scheme for Shared-Memory Parallel Applications. | John Corredor, Juan Carlos Moure, Dolores Rexachs, Daniel Franco, Emilio Luque |
| 2011 | FastQuery: A Parallel Indexing System for Scientific Data. | Jerry Chi-Yuan Chou, Kesheng Wu, Prabhat |
| 2011 | Exploring Fine-Grained Task-Based Execution on Multi-GPU Systems. | Long Chen, Oreste Villa, Guang R. Gao |
| 2011 | HEaRS: A Hierarchical Energy-Aware Resource Scheduler for Virtualized Data Centers. | Hui Chen, Meina Song, Junde Song, Ada Gavrilovska, Karsten Schwan |
| 2011 | Automatically Selecting the Number of Aggregators for Collective I/O Operations. | Mohamad Chaarawi, Edgar Gabriel |
| 2011 | Predictive and Distributed Routing Balancing for High Speed Interconnection Networks. | Carlos Nunez Castillo, Diego Lugones, Daniel Franco, Emilio Luque |
| 2011 | Achieving Scalable Parallelization for the Hessenberg Factorization. | Anthony M. Castaldo, R. Clint Whaley |
| 2011 | Design of HPC Node with Heterogeneous Processors. | Zheng Cao, Hongwei Tang, Qiang Li, Bo Li, Fei Chen, Kai Wang, Xuejun An, Ninghui Sun |
| 2011 | A Sampling-Based Approach for Communication Libraries Auto-Tuning. | Elisabeth Brunet, Franois Trahay, Alexandre Denis, Raymond Namyst |
| 2011 | On Scalability for MPI Runtime Systems. | George Bosilca, Thomas Hrault, Ala Rezmerita, Jack J. Dongarra |
| 2011 | Performance Portability of a GPU Enabled Factorization with the DAGuE Framework. | George Bosilca, Aurlien Bouteiller, Thomas Hrault, Pierre Lemarinier, Narapat Ohm Saengpatsa, Stanimire Tomov, Jack J. Dongarra |
| 2011 | Implementation of Multigrid on QPACE. | Matthias Bolten, Daniel Brinkers, Ulrich Rde, Markus Strmer |