| 2018 | Analysis of Explicit vs. Implicit Tasking in OpenMP Using Kripke. | Charles Jin, Muthu Manikandan Baskaran |
| 2018 | iSpan: parallel identification of strongly connected components with spanning trees. | Yuede Ji, Hang Liu, H. Howie Huang |
| 2018 | Chapel Aggregation Library (CAL). | Louis Jenkins, Marcin Zalewski, Michael Ferguson |
| 2018 | Framework for scalable intra-node collective operations using shared memory. | Surabhi Jain, Rashid Kaleem, Marc Gamell Balmana, Akhil Langer, Dmitry Durnov, Alexander Sannikov, Maria Garzaran |
| 2018 | Lessons learned from analyzing dynamic promotion for user-level threading. | Shintaro Iwasaki, Abdelhalim Amer, Kenjiro Taura, Pavan Balaji |
| 2018 | A fast scalable implicit solver for nonlinear time-evolution earthquake city problem on low-ordered unstructured finite elements with artificial intelligence and transprecision computing. | Tsuyoshi Ichimura, Kohei Fujita, Takuma Yamaguchi, Akira Naruse, Jack C. Wells, Thomas C. Schulthess, Tjerk P. Straatsma, Christopher Zimmer, Maxime Martinasso, Kengo Nakajima, Muneo Hori, Lalith Maddegedara |
| 2018 | Partial redundancy in HPC systems with non-uniform node reliabilities. | Zaeem Hussain, Taieb Znati, Rami G. Melhem |
| 2018 | Algorithm Selection of MPI Collectives Using Machine Learning Techniques. | Sascha Hunold, Alexandra Carpen-Amarie |
| 2018 | TriCore: parallel triangle counting on GPUs. | Yang Hu, Hang Liu, H. Howie Huang |
| 2018 | A Comprehensive Informative Metric for Analyzing HPC System Status Using the LogSCAN Platform. | Yawei Hui, Byung-Hoon Park, Christian Engelmann |
| 2018 | Compiler-aided Type Tracking for Correctness Checking of MPI Applications. | Alexander Hck, Jan-Patrick Lehr, Sebastian Kreutzer, Joachim Protze, Christian Terboven, Christian H. Bischof, Matthias S. Mller |
| 2018 | PARCOACH Extension for a Full-Interprocedural Collectives Verification. | Pierre Huchant, Emmanuelle Saillard, Denis Barthou, Hugo Brunie, Patrick Carribault |
| 2018 | Accelerating quantum chemistry with vectorized and batched integrals. | Hua Huang, Edmond Chow |
| 2018 | Bandwidth Scheduling for Big Data Transfer with Deadline Constraint between Data Centers. | Aiqin Hou, Chase Q. Wu, Dingyi Fang, Liudong Zuo, Michelle M. Zhu, Xiaoyang Zhang, Ruimin Qiao, Xiaoyan Yin |
| 2018 | Integration of Burst Buffer in High-level Parallel I/O Library for Exa-scale Computing Era. | Kaiyuan Hou, Reda Al-Bahrani, Esteban Rangel, Ankit Agrawal, Robert Latham, Robert B. Ross, Alok N. Choudhary, Wei-keng Liao |
| 2018 | Tracking Network Flows with P4. | Joseph Hill, Mitchel Aloserij, Paola Grosso |
| 2018 | GASNet-EX Performance Improvements Due to Specialization for the Cray Aries Network. | Paul H. Hargrove, Dan Bonachea |
| 2018 | Harnessing GPU tensor cores for fast FP16 arithmetic to speed up mixed-precision iterative refinement solvers. | Azzam Haidar, Stanimire Tomov, Jack J. Dongarra, Nicholas J. Higham |
| 2018 | Software Prefetching for Unstructured Mesh Applications. | Ioan Hadade, Timothy M. Jones, Feng Wang, Luca di Mare |
| 2018 | Employing Student Retention Strategies for an Introductory GPU Programming Course. | Julian Gutierrez, Fritz Previlon, David R. Kaeli |
| 2018 | FlipTracker: understanding natural error resilience in HPC applications. | Luanzheng Guo, Dong Li, Ignacio Laguna, Martin Schulz |
| 2018 | Dynamic data race detection for OpenMP programs. | Yizi Gu, John M. Mellor-Crummey |
| 2018 | Fast Detection of Elephant Flows with Dirichlet-Categorical Inference. | Aditya Gudibanda, Jordi Ros-Giralt, Alan Commike, Richard Lethin |
| 2018 | High-Performance GPU Implementation of PageRank with Reduced Precision Based on Mantissa Segmentation. | Thomas Grtzmacher, Hartwig Anzt, Florian Scheidegger, Enrique S. Quintana-Ort |
| 2018 | Jupyter Notebooks and User-Friendly HPC Access. | Ben Glick, Jens Mache |