| 2019 | Optimizing GPU programs by register demotion: poster. | Putt Sakdhnagool, Amit Sabne, Rudolf Eigenmann |
| 2019 | Processing transactions in a predefined order. | Mohamed M. Saad, Masoomeh Javidi Kishi, Shihao Jing, Sandeep Hans, Roberto Palmieri |
| 2019 | Deciphering Predictive Schedulers for Heterogeneous-ISA Multicore Architectures. | Andreas Prodromou, Ashish Venkat, Dean M. Tullsen |
| 2019 | Formal Verification through Combinatorial Topology: the CAS-Extended Model. | Christina L. Peterson, Damian Dechev |
| 2019 | Task-DAG Support in Single-Source PHAST Library: Enabling Flexible Assignment of Tasks to CPUs and GPUs in Heterogeneous Architectures. | Biagio Peccerillo, Sandro Bartolini |
| 2019 | High performance distributed deep learning: a beginner's guide. | Dhabaleswar K. Panda, Ammar Ahmad Awan, Hari Subramoni |
| 2019 | Checking linearizability using hitting families. | Burcu Kulahcioglu Ozkan, Rupak Majumdar, Filip Niksic |
| 2019 | GOPipe: a granularity-oblivious programming framework for pipelined stencil executions on GPU. | Chanyoung Oh, Zhen Zheng, Xipeng Shen, Jidong Zhai, Youngmin Yi |
| 2019 | Compiler-assisted adaptive program scheduling in big.LITTLE systems: poster. | Marcelo Novaes, Vinicius Petrucci, Abdoulaye Gamati, Fernando Magno Quinto Pereira |
| 2019 | Process Barrier for Predictable and Repeatable Concurrent Execution. | Masataka Nishi |
| 2019 | Automated multi-dimensional elasticity for streaming runtimes: poster. | Xiang Ni, Scott Schneider, Raju Pavuluri, Jonathan Kaus, Kun-Lung Wu |
| 2019 | Don't Forget About Synchronization!: A Case Study of K-Means on GPU. | Jacob Nelson, Roberto Palmieri |
| 2019 | Programming quantum computers: a primer with IBM Q and D-Wave exercises. | Frank Mueller, Greg Byrd, Patrick Dreher |
| 2019 | A pattern based algorithmic autotuner for graph processing on GPUs. | Ke Meng, Jiajia Li, Guangming Tan, Ninghui Sun |
| 2019 | libMPNode: An OpenMP Runtime For Parallel Processing Across Incoherent Domains. | Robert Lyerly, Sang-Hoon Kim, Binoy Ravindran |
| 2019 | A coordinated tiling and batching framework for efficient GEMM on GPUs. | Xiuhong Li, Yun Liang, Shengen Yan, Liancheng Jia, Yinghan Li |
| 2019 | GPOP: a cache and memory-efficient framework for graph processing over partitions. | Kartik Lakhotia, Rajgopal Kannan, Sourav Pati, Viktor K. Prasanna |
| 2019 | Wait-free Dynamic Transactions for Linked Data Structures. | Pierre LaBorde, Lance Lebanoff, Christina L. Peterson, Deli Zhang, Damian Dechev |
| 2019 | Corrected trees for reliable group communication. | Martin Kttler, Maksym Planeta, Jan Bierbaum, Carsten Weinhold, Hermann Hrtig, Amnon Barak, Torsten Hoefler |
| 2019 | Lock-free channels for programming via communicating sequential processes: poster. | Nikita Koval, Dan Alistarh, Roman Elizarov |
| 2019 | Scheduling HPC workloads on heterogeneous-ISA architectures: poster. | Mohamed Lamine Karaoui, Anthony Carno, Robert Lyerly, Sang-Hoon Kim, Pierre Olivier, Changwoo Min, Binoy Ravindran |
| 2019 | High-throughput image alignment for connectomics using frugal snap judgments: poster. | Tim Kaler, Brian Wheatman, Sarah Wooders |
| 2019 | Brie: A Specialized Trie for Concurrent Datalog. | Herbert Jordan, Pavle Subotic, David Zhao, Bernhard Scholz |
| 2019 | A specialized B-tree for concurrent datalog evaluation. | Herbert Jordan, Pavle Subotic, David Zhao, Bernhard Scholz |
| 2019 | Creating repeatable, reusable experimentation pipelines with popper: tutorial. | Ivo Jimenez, Jay F. Lofstead, Carlos Maltzahn |