| 2016 | Neural Network-Based Task Scheduling with Preemptive Fan Control. | Bilge Acun, Eun Kyung Lee, Yoonho Park, Laxmikant V. Kal |
| 2016 | Using Scientific Workflows for Science and Engineering Optimisation. | David Abramson |
| 2016 | Scalemine: scalable parallel frequent subgraph mining in a single large graph. | Ehab Abdelhamid, Ibrahim Abdelaziz, Panos Kalnis, Zuhair Khayyat, Fuad T. Jamour |
| 2016 | Optimized Distributed Work-Stealing. | Vivek Kumar, Karthik Murthy, Vivek Sarkar, Yili Zheng |
| 2016 | SERF: efficient scheduling for fast deep neural network serving via judicious parallelism. | Feng Yan, Yuxiong He, Olatunji Ruwase, Evgenia Smirni |
| 2015 | Feature frequency profiles for automatic sample identification using PySpark. | Gregory J. Zynda, Niall Gaffney, Mehmet M. Dalkilic, Matthew W. Vaughn |
| 2015 | In Situ Analysis as a Parallel I/O Problem. | Sean Ziegeler, Chuck Atkins, Andrew C. Bauer, Lucas Pettey |
| 2015 | DeltaFS: exascale file systems scale better without dedicated servers. | Qing Zheng, Kai Ren, Garth A. Gibson, Bradley W. Settlemyer, Gary Grider |
| 2015 | Co-sites: the autonomous distributed dataflows in collaborative scientific discovery. | Yanwei Zhang, Matthew Wolf, Karsten Schwan, Qing Liu, Greg Eisenhauer, Scott Klasky |
| 2015 | Dynamic parallelism for simple and efficient GPU graph algorithms. | Peter Zhang, Eric Holk, John Matty, Samantha Misurda, Marcin Zalewski, Jonathan Chu, Scott McMillan, Andrew Lumsdaine |
| 2015 | SJM: an SCM-based journaling mechanism with write reduction for file systems. | Lingfang Zeng, Binbing Hou, Dan Feng, Kenneth B. Kent |
| 2015 | Generalised vectorisation for sparse matrix: vector multiplication. | A. N. Yzelman |
| 2015 | Performance optimization for the k-nearest neighbors kernel on x86 architectures. | Chenhan D. Yu, Jianyu Huang, Woody Austin, Bo Xiao, George Biros |
| 2015 | Optimizing deep learning hyper-parameters through an evolutionary algorithm. | Steven R. Young, Derek C. Rose, Thomas P. Karnowski, Seung-Hwan Lim, Robert M. Patton |
| 2015 | Mixed-precision block gram Schmidt orthogonalization. | Ichitaro Yamazaki, Stanimire Tomov, Jakub Kurzak, Jack J. Dongarra, Jesse L. Barlow |
| 2015 | Randomized algorithms to update partial singular value decomposition on a hybrid CPU/GPU cluster. | Ichitaro Yamazaki, Jakub Kurzak, Piotr Luszczek, Jack J. Dongarra |
| 2015 | Big data analytics on traditional HPC infrastructure using two-level storage. | Pengfei Xuan, Jeffrey Denton, Pradip K. Srimani, Rong Ge, Feng Luo |
| 2015 | A low-cost adaptive data separation method for the flash translation layer of solid state drives. | Wei Xie, Yong Chen, Philip C. Roth |
| 2015 | Interlanguage parallel scripting for distributed-memory scientific computing. | Justin M. Wozniak, Timothy G. Armstrong, Ketan Maheshwari, Daniel S. Katz, Michael Wilde, Ian T. Foster |
| 2015 | Workflow provenance: an analysis of long term storage costs. | Simon Woodman, Hugo Hiden, Paul Watson |
| 2015 | SCinet: 25 years of extreme networking. | Linda Winkler |
| 2015 | Guided profiling for auto-tuning array layouts on GPUs. | Nicolas Weber, Sandra C. Amend, Michael Goesele |
| 2015 | Automatic and transparent I/O optimization with storage integrated application runtime support. | Noah Watkins, Zhihao Jia, Galen M. Shipman, Carlos Maltzahn, Alex Aiken, Patrick S. McCormick |
| 2015 | A practical approach to reconciling availability, performance, and capacity in provisioning extreme-scale storage systems. | Lipeng Wan, Feiyi Wang, Sarp Oral, Devesh Tiwari, Sudharshan S. Vazhkudai, Qing Cao |
| 2015 | HydraDB: a resilient RDMA-driven key-value middleware for in-memory cluster computing. | Yandong Wang, Li Zhang, Jian Tan, Min Li, Yuqing Gao, Xavier Guerin, Xiaoqiao Meng, Shicong Meng |