| 2020 | GPGPU performance estimation for frequency scaling using cross-benchmarking. | Qiang Wang, Chengjian Liu, Xiaowen Chu |
| 2020 | Optimizing GPU programs by partial evaluation. | Aleksey Tyurin, Daniil Berezun, Semyon V. Grigorev |
| 2020 | waveSZ: a hardware-algorithm co-design of efficient lossy compression for scientific data. | Jiannan Tian, Sheng Di, Chengming Zhang, Xin Liang, Sian Jin, Dazhao Cheng, Dingwen Tao, Franck Cappello |
| 2020 | ArcherGear: data race equivalencing for expeditious HPC debugging. | Samuel Thayer, Ganesh Gopalakrishnan, Ian Briggs, Michael Bentley, Dong H. Ahn, Ignacio Laguna, Gregory L. Lee |
| 2020 | XIndex: a scalable learned index for multicore data storage. | Chuzhe Tang, Youyun Wang, Zhiyuan Dong, Gansen Hu, Zhaoguo Wang, Minjie Wang, Haibo Chen |
| 2020 | ELSE: an efficient link-time static instrumentation tool for embedded system. | Xiaoxin Tang |
| 2020 | Practical parallel hypergraph algorithms. | Julian Shun |
| 2020 | Revisiting linpack algorithm on large-scale CPU-GPU heterogeneous systems. | Chaoyang Shui, Xianzhi Yu, Yujin Yan, Yinshan Wang, Ke Meng, Guangming Tan |
| 2020 | Functional faults. | Gali Sheffi, Erez Petrank |
| 2020 | Scalable top-k retrieval with Sparta. | Gali Sheffi, Dmitry Basin, Edward Bortnikov, David Carmel, Idit Keidar |
| 2020 | A supernodal all-pairs shortest path algorithm. | Piyush Sao, Ramakrishnan Kannan, Prasun Gera, Richard W. Vuduc |
| 2020 | On the fly MHP analysis. | Sonali Saha, V. Krishna Nandivada |
| 2020 | Fast concurrent data sketches. | Arik Rinberg, Alexander Spiegelman, Edward Bortnikov, Eshcar Hillel, Idit Keidar, Lee Rhodes, Hadar Serviansky |
| 2020 | Bounded incoherence: a programming model for non-cache-coherent shared memory architectures. | Yuxin Ren, Gabriel Parmer, Dejan S. Milojicic |
| 2020 | High-level hardware feature extraction for GPU performance prediction of stencils. | Toomas Remmelg, Bastian Hagedorn, Lu Li, Michel Steuwer, Sergei Gorlatch, Christophe Dubach |
| 2020 | Exploring accelerator and parallel graph algorithmic choices for temporal graphs. | Akif Rehman, Masab Ahmad, Omer Khan |
| 2020 | Unveiling kernel concurrency in multiresolution filters on GPUs with an image processing DSL. | Bo Qiao, Oliver Reiche, Jrgen Teich, Frank Hannig |
| 2020 | Scheduling Irregular Dataflow Pipelines on SIMD Architectures. | Tom Plano, Jeremy Buhler |
| 2020 | SIMD-based Exact Parallel Fuzzy Dilation Operator for Fast Computing of Fuzzy Spatial Relations. | Rgis Pierrard, Laurent Cabaret, Jean-Philippe Poli, Cline Hudelot |
| 2020 | Automated test generation for OpenCL kernels using fuzzing and constraint solving. | Chao Peng, Ajitha Rajan |
| 2020 | spECK: accelerating GPU sparse matrix-matrix multiplication through lightweight analysis. | Mathias Parger, Martin Winter, Daniel Mlakar, Markus Steinberger |
| 2020 | Scaling concurrent queues by using HTM to profit from failed atomic operations. | Or Ostrovsky, Adam Morrison |
| 2020 | Universal wait-free memory reclamation. | Ruslan Nikolaev, Binoy Ravindran |
| 2020 | Automatic generation of specialized direct convolutions for mobile GPUs. | Naums Mogers, Valentin Radu, Lu Li, Jack Turner, Michael F. P. O'Boyle, Christophe Dubach |
| 2020 | Oak: a scalable off-heap allocated key-value map. | Hagar Meir, Dmitry Basin, Edward Bortnikov, Anastasia Braginsky, Yonatan Gottesman, Idit Keidar, Eran Meir, Gali Sheffi, Yoav Zuriel |