| 2012 | Hierarchical Clustering Strategies for Fault Tolerance in Large Scale HPC Systems. | Leonardo Arturo Bautista-Gomez, Thomas Ropars, Naoya Maruyama, Franck Cappello, Satoshi Matsuoka |
| 2012 | Replication Based QoS Framework for Flash Arrays. | Nihat Altiparmak, Ali Saman Tosun |
| 2011 | Investigating Scenario-Conscious Asynchronous Rendezvous over RDMA. | Judicael A. Zounmevo, Ahmad Afsahi |
| 2011 | Evaluating Performance Impacts of Delayed Failure Repairing on Large-Scale Systems. | Zhou Zhou, Wei Tang, Ziming Zheng, Zhiling Lan, Narayan Desai |
| 2011 | Data Partitioning on Heterogeneous Multicore Platforms. | Ziming Zhong, Vladimir Rychkov, Alexey L. Lastovetsky |
| 2011 | GPApriori: GPU-Accelerated Frequent Itemset Mining. | Fan Zhang, Yan Zhang, Jason D. Bakos |
| 2011 | Frequent Itemset Mining on Large-Scale Shared Memory Machines. | Yan Zhang, Fan Zhang, Jason D. Bakos |
| 2011 | BMF: Bitmapped Mass Fingerprinting for Fast Protein Identification. | Weikuan Yu, K. John Wu, Wei-Shinn Ku, Cong Xu, Juan Gao |
| 2011 | Implementing High Performance Remote Method Invocation in CCA. | Jian Yin, Khushbu Agarwal, Manoj Krishnan, Daniel G. Chavarra-Miranda, Ian Gorton, Tom Epperly |
| 2011 | Improving PCM Endurance with Randomized Address Remapping in Hybrid Memory System. | Gang Wu, Jian Gao, Huxing Zhang, Yaozu Dong |
| 2011 | Performance Emulation of Cell-Based AMR Cosmology Simulations. | Jingjin Wu, Roberto E. Gonzlez, Zhiling Lan, Nickolay Y. Gnedin, Andrey V. Kravtsov, Douglas H. Rudd, Yongen Yu |
| 2011 | Improving I/O Forwarding Throughput with Data Compression. | Benjamin Welton, Dries Kimpe, Jason Cope, Christina M. Patrick, Kamil Iskra, Robert B. Ross |
| 2011 | Optimized Non-contiguous MPI Datatype Communication for GPU Clusters: Design, Implementation and Evaluation with MVAPICH2. | Hao Wang, Sreeram Potluri, Miao Luo, Ashish Kumar Singh, Xiangyong Ouyang, Sayantan Sur, Dhabaleswar K. Panda |
| 2011 | Play It Again, SimMR! | Abhishek Verma, Ludmila Cherkasova, Roy H. Campbell |
| 2011 | Automatic Hybrid OpenMP + MPI Program Generation for Dynamic Programming Problems. | Denny R. Vandenberg, Quentin F. Stout |
| 2011 | EDO: Improving Read Performance for Scientific Applications through Elastic Data Organization. | Yuan Tian, Scott Klasky, Hasan Abbasi, Jay F. Lofstead, Ray W. Grout, Norbert Podhorszki, Qing Liu, Yandong Wang, Weikuan Yu |
| 2011 | Automatic Computer System Characterization for a Parallelizing Compiler. | Alan Sussman, Norman Lo, Timothy Anderson |
| 2011 | Improving MapReduce Performance via Heterogeneity-Load-Aware Partition Function. | Huifeng Sun, Junliang Chen, Chuanchang Liu, Zibin Zheng, Nan Yu, Zhi Yang |
| 2011 | Design and Evaluation of Network Topology-/Speed- Aware Broadcast Algorithms for InfiniBand Clusters. | Hari Subramoni, Krishna Chaitanya Kandalla, Jrme Vienne, Sayantan Sur, Bill Barth, Karen A. Tomko, Robert T. McLay, Karl W. Schulz, Dhabaleswar K. Panda |
| 2011 | Quartile and Outlier Detection on Heterogeneous Clusters Using Distributed Radix Sort. | Kyle Spafford, Jeremy S. Meredith, Jeffrey S. Vetter |
| 2011 | An ISO-Energy-Efficient Approach to Scalable System Power-Performance Optimization. | Shuaiwen Song, Matthew Grove, Kirk W. Cameron |
| 2011 | MPI Alltoall Personalized Exchange on GPGPU Clusters: Design Alternatives and Benefit. | Ashish Kumar Singh, Sreeram Potluri, Hao Wang, Krishna Chaitanya Kandalla, Sayantan Sur, Dhabaleswar K. Panda |
| 2011 | An Energy-Efficient Scheme for Cloud Resource Provisioning Based on CloudSim. | Yuxiang Shi, Xiaohong Jiang, Kejiang Ye |
| 2011 | Design and Implementation of Broadcast Algorithms for Extreme-Scale Systems. | Pavel Shamis, Richard L. Graham, Manjunath Gorentla Venkata, Joshua Ladd |
| 2011 | Multiphase LBM Distributed over Multiple GPUs. | Carlos Rosales |