| 2015 | Vectorized Big Integer Operations for Cryptosystems on the Intel MIC Architecture. | Cheng Chang, Shun Yao, Dantong Yu |
| 2015 | Improving Communication Throughput by Multipath Load Balancing on Blue Gene/Q. | Huy Bui, Preeti Malakar, Venkatram Vishwanath, Todd S. Munson, Eun-Sung Jung, Andrew E. Johnson, Michael E. Papka, Jason Leigh |
| 2015 | On the Use of Commodity Ethernet Technology in Exascale HPC Systems. | Mariano Benito, Enrique Vallejo, Ramn Beivide |
| 2015 | Which Verification for Soft Error Detection? | Leonardo Bautista-Gomez, Anne Benoit, Aurlien Cavelan, Saurabh K. Raina, Yves Robert, Hongyang Sun |
| 2015 | A Performance Model for GPU-Accelerated FDTD Applications. | Paul F. Baumeister, Thorsten Hater, Jiri Kraus, Dirk Pleiter, Pierre Wahl |
| 2015 | Scaling Computation on GPUs Using Powerlists. | Anshu S. Anand, R. K. Shyamasundar |
| 2015 | A Simple BSP-based Model to Predict Execution Time in GPU Applications. | Marcos Amaris, Daniel Cordeiro, Alfredo Goldman, Raphael Y. de Camargo |
| 2015 | Big data in life sciences and public health. | Srinivas Aluru |
| 2015 | On the Resilience of Parallel Sparse Hybrid Solvers. | Emmanuel Agullo, Luc Giraud, Mawussi Zounon |
| 2015 | Task-Based Multifrontal QR Solver for GPU-Accelerated Multicore Architectures. | Emmanuel Agullo, Alfredo Buttari, Abdou Guermouche, Florent Lopez |
| 2015 | FlexCore: A Reconfigurable Processor Supporting Flexible, Dynamic Morphing. | Furat Afram, Kanad Ghose |
| 2015 | Holistic Management of Sustainable Geo-Distributed Data Centers. | Zahra Abbasi, Sandeep K. S. Gupta |
| 2014 | High performance MPI library over SR-IOV enabled infiniband clusters. | Jie Zhang, Xiaoyi Lu, Jithin Jose, Mingzhe Li, Rong Shi, Dhabaleswar K. Panda |
| 2014 | Combining HoL-blocking avoidance and differentiated services in high-speed interconnects. | Pedro Ybenes, Jess Escudero-Sahuquillo, Crispn Gmez Requena, Pedro Javier Garca, Francisco J. Alfaro, Francisco J. Quiles, Jos Duato |
| 2014 | An early experience of regional ocean modelling on intel many integrated core architecture. | Srikanth Yalavarthi, Akshara Kaginalkar |
| 2014 | Smart multi-task scheduling for OpenCL programs on CPU/GPU heterogeneous platforms. | Yuan Wen, Zheng Wang, Michael F. P. O'Boyle |
| 2014 | A high performance broadcast design with hardware multicast and GPUDirect RDMA for streaming applications on Infiniband clusters. | Akshay Venkatesh, Hari Subramoni, Khaled Hamidouche, Dhabaleswar K. Panda |
| 2014 | Parallel AMG solver for three dimensional unstructured grids using GPU. | K. Ravi Tej, Naveen Sivadasan, Vatsalya Sharma, Raja Banerjee |
| 2014 | Xevolver: An XML-based code translation framework for supporting HPC application migration. | Hiroyuki Takizawa, Shoichi Hirasawa, Yasuharu Hayashi, Ryusuke Egawa, Hiroaki Kobayashi |
| 2014 | Optimization of scan algorithms on multi- and many-core processors. | Qiao Sun, Chao Yang |
| 2014 | DRIVE: Using implicit caching hints to achieve disk I/O reduction in virtualized environments. | Sujesha Sudevalayam, Purushottam Kulkarni |
| 2014 | Matrix-matrix multiplication on a large register file architecture with indirection. | Dheeraj Sreedhar, Jeff H. Derby, Robert K. Montoye, Charles L. Johnson |
| 2014 | Queueing-based storage performance modeling and placement in OpenStack environments. | Yang Song, Rakesh Jain, Ramani Routray |
| 2014 | Simple parallel biconnectivity algorithms for multicore platforms. | George M. Slota, Kamesh Madduri |
| 2014 | Designing efficient small message transfer mechanism for inter-node MPI communication on InfiniBand GPU clusters. | Rong Shi, Sreeram Potluri, Khaled Hamidouche, Jonathan L. Perkins, Mingzhe Li, Davide Rossetti, Dhabaleswar K. Panda |