| 2015 | SSMART: smart scheduling of multi-architecture tasks on heterogeneous systems. | Judit Planas, Rosa M. Badia, Eduard Ayguad, Jess Labarta |
| 2015 | Reducing overhead in the Uintah framework to support short-lived tasks on GPU-heterogeneous architectures. | Brad Peterson, Harish Kumar Dasari, Alan Humphrey, James C. Sutherland, Tony Saad, Martin Berzins |
| 2015 | Accelerating dynamically typed languages with a virtual function cache. | Lillian Pentecost, John A. Stratton |
| 2015 | A data streaming model in MPI. | Ivy Bo Peng, Stefano Markidis, Erwin Laure, Daniel J. Holmes, Mark Bull |
| 2015 | VOCL-FT: introducing techniques for efficient soft error coprocessor recovery. | Antonio J. Pea, Wesley Bland, Pavan Balaji |
| 2015 | Early experiences with node-level power capping on the Cray XC40 platform. | Kevin T. Pedretti, Stephen L. Olivier, Kurt B. Ferreira, Galen M. Shipman, Wei Shu |
| 2015 | Database assisted distribution to improve fault tolerance for multiphysics applications. | Robert S. Pavel, Allen L. McPherson, Timothy C. Germann, Christoph Junghans |
| 2015 | BD-CATS: big data clustering at trillion particle scale. | Md. Mostofa Ali Patwary, Surendra Byna, Nadathur Rajagopalan Satish, Narayanan Sundaram, Zarija Lukic, Vadim Roytershteyn, Michael J. Anderson, Yushu Yao, Prabhat, Pradeep Dubey |
| 2015 | High-performance algebraic multigrid solver optimized for multi-core based distributed parallel systems. | Jongsoo Park, Mikhail Smelyanskiy, Ulrike Meier Yang, Dheevatsa Mudigere, Pradeep Dubey |
| 2015 | ELF: maximizing memory-level parallelism for GPUs with coordinated warp and fetch scheduling. | Jason Jong Kyu Park, Yongjun Park, Scott A. Mahlke |
| 2015 | Runtime-driven shared last-level cache management for task-parallel programs. | Abhisek Pan, Vijay S. Pai |
| 2015 | Exploring dynamic parallelism in OpenMP. | Guray Ozen, Eduard Ayguad, Jess Labarta |
| 2015 | Hysteresis-based optimization of data transfer throughput. | Md. S. Q. Zulkar Nine, Kemal Guner, Tevfik Kosar |
| 2015 | High speed scientific data transfers using software defined networking. | Harvey B. Newman, Azher Mughal, Dorian Kcira, Iosif Legrand, Ramiro Voicu, Julian J. Bunn |
| 2015 | Introducing high performance computing concepts into engineering undergraduate curriculum: a success story. | B. Neelima, Jiajia Li |
| 2015 | Simulating stencil-based application on future Xeon Phi processor. | Chitra Natarajan, Carl J. Beckmann, Anthony Nguyen, Mauricio Araya-Polo, Tryggve Fossum, Detlef Hohl |
| 2015 | GraphBIG: understanding graph computing in the context of industrial solutions. | Lifeng Nai, Yinglong Xia, Ilie Gabriel Tanase, Hyesoon Kim, Ching-Yung Lin |
| 2015 | Automatic sharing classification and timely push for cache-coherent systems. | Malek Musleh, Vijay S. Pai |
| 2015 | Design and implementation of control sequence generator for SDN-enhanced MPI. | Baatarsuren Munkhdorj, Keichi Takahashi, Dashdavaa Khureltulga, Yasuhiro Watashiba, Yoshiyuki Kido, Susumu Date, Shinji Shimojo |
| 2015 | Lighthouse: an automated solver selection tool. | Pate Motter, Kanika Sood, Elizabeth R. Jessup, Boyana Norris |
| 2015 | Contemporary challenges for data-intensive scientific workflow management systems. | Ryan Mork, Paul Martin, Zhiming Zhao |
| 2015 | Profile-based power shifting in interconnection networks with on/off links. | Shinobu Miwa, Hiroshi Nakamura |
| 2015 | Towards the development of hierarchical data motion power cost models. | Tiffany M. Mintz, Oluwatosin O. Alabi |
| 2015 | Trends in system cost and performance balances and implications for the future of HPC. | John D. McCalpin |
| 2015 | Performance of random sampling for computing low-rank approximations of a dense matrix on GPUs. | Tho Mary, Ichitaro Yamazaki, Jakub Kurzak, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra |