| 2017 | S-Aligner: Ultrascalable Read Mapping on Sunway Taihu Light. | Xiaohui Duan, Kai Xu, Yuandong Chan, Christian Hundt, Bertil Schmidt, Pavan Balaji, Weiguo Liu |
| 2017 | Manala: A Flexible Flow Control Library for Asynchronous Task Communication. | Matthieu Dreher, Kiran Sasikumar, Subramanian Sankaranarayanan, Tom Peterka |
| 2017 | Justice: A Deadline-Aware, Fair-Share Resource Allocator for Implementing Multi-Analytics. | Stratos Dimopoulos, Chandra Krintz, Rich Wolski |
| 2017 | TGE: Machine Learning Based Task Graph Embedding for Large-Scale Topology Mapping. | Jong Youl Choi, Jeremy Logan, Matthew Wolf, George Ostrouchov, Tahsin M. Kur, Qing Liu, Norbert Podhorszki, Scott Klasky, Melissa Romanus, Qian Sun, Manish Parashar, Randy Michael Churchill, Choong-Seock Chang |
| 2017 | DH-Falcon: A Language for Large-Scale Graph Processing on Distributed Heterogeneous Systems. | Unnikrishnan Cheramangalath, Rupesh Nasre, Y. N. Srikant |
| 2017 | AMM: Scalable Memory Reuse Model to Predict the Performance of Physics Codes. | Gopinath Chennupati, Nandakishore Santhi, Stephan J. Eidenbenz, Sunil Thulasidasan |
| 2017 | Mitigating the Write Amplification Problem of Write-Optimized File Systems on Flash Storage. | Shuo-Han Chen, Jun-Long Lin, Tseng-Yi Chen, Tsan-sheng Hsu, Hsin-Wen Wei, Wei-Kuan Shih |
| 2017 | keybin: Key-Based Binning for Distributed Clustering. | Xinyu Chen, Jeremy Benson, Trilce Estrada |
| 2017 | Evaluating Effect of Write Combining on PCIe Throughput to Improve HPC Interconnect Performance. | Mahesh Chaudhari, Kedar Kulkarni, Shreeya Badhe, Vandana Inamdar |
| 2017 | A Malleable and Fault-Tolerant Task Pool Framework for X10. | Marco Bungart, Claudia Fohry |
| 2017 | Co-locating Graph Analytics and HPC Applications. | Kevin A. Brown, Satoshi Matsuoka |
| 2017 | Dynamic Co-Scheduling Driven by Main Memory Bandwidth Utilization. | Jens Breitbart, Simon Pickartz, Stefan Lankes, Josef Weidendorfer, Antonello Monti |
| 2017 | Monitoring Infrastructure: The Challenges of Moving Beyond Petascale. | Amanda Bonnie, Mike Mason, Daniel Illescas |
| 2017 | Flexible Data Aggregation for Performance Profiling. | David Bhme, David Beckingsale, Martin Schulz |
| 2017 | Parallelized Recovery of Hundreds of Millions Small Data Objects. | Kevin Beineke, Stefan Nothaas, Michael Schttner |
| 2017 | Parallel and Efficient Sensitivity Analysis of Microscopy Image Segmentation Workflows in Hybrid Systems. | Willian Barreiros, George Teodoro, Tahsin M. Kur, Jun Kong, Alba C. M. A. Melo, Joel H. Saltz |
| 2017 | Understanding the Role of GPGPU-Accelerated SoC-Based ARM Clusters. | Reza Azimi, Tyler Fox, Sherief Reda |
| 2017 | Assuming Failure Independence: Are We Right to be Wrong? | Guillaume Aupy, Yves Robert, Frdric Vivien |
| 2017 | Optimizing the Datapath for Key-value Middleware with NVMe SSDs over RDMA Interconnects. | Zhongqi An, Zhengyu Zhang, Qiang Li, Jing Xing, Hao Du, Zhan Wang, Zhigang Huo, Jie Ma |
| 2017 | The Effect of Resource Allocation and System Events on VM Consolidation. | Maruf Ahmed, Albert Y. Zomaya |
| 2017 | Measuring Minimum Switch Port Metric Retrieval Time and Impact for Multi-layer InfiniBand Fabrics. | Michael Aguilar, Benjamin A. Allan, Sergei Polevitzky |
| 2017 | YAViT (Yet Another Viz Tool): Raising the Level of Abstraction in End-User HPC Interactions. | Omar Aaziz, Ujjwal Panthi, Jonathan Cook |
| 2017 | Contention-Aware Kernel-Assisted MPI Collectives for Multi-/Many-Core Systems. | Sourav Chakraborty, Hari Subramoni, Dhabaleswar K. Panda |
| 2016 | Exploring Plan-Based Scheduling for Large-Scale Computing Systems. | Xingwu Zhang, Zhou Zhou, Xu Yang, Zhiling Lan, Jia Wang |
| 2016 | FlashStager: Improving the Performance of SSD-Based Data Staging Systems via Write Redirection. | Xuechen Zhang, Fang Zheng, Karsten Schwan, Matthew Wolf |