| 2020 | CapelliniSpTRSV: A Thread-Level Synchronization-Free Sparse Triangular Solve on GPUs. | Jiya Su, Feng Zhang, Weifeng Liu, Bingsheng He, Ruofan Wu, Xiaoyong Du, Rujia Wang |
| 2020 | E-LAS: Design and Analysis of Completion-Time Agnostic Scheduling for Distributed Deep Learning Cluster. | Abeda Sultana, Li Chen, Fei Xu, Xu Yuan |
| 2020 | HCAPP: Scalable Power Control for Heterogeneous 2.5D Integrated Systems. | Kramer Straube, Jason Lowe-Power, Christopher Nitta, Matthew K. Farrens, Venkatesh Akella |
| 2020 | Revisiting Sparse Dynamic Programming for the 0/1 Knapsack Problem. | Tarequl Islam Sifat, Nirmal Prajapati, Sanjay V. Rajopadhye |
| 2020 | Towards High-Efficiency Data Centers via Job-Aware Network Scheduling. | Yang Shi, Mei Wen, Chunyuan Zhang |
| 2020 | GraBi: Communication-Efficient and Workload-Balanced Partitioning for Bipartite Graphs. | Feng Sheng, Qiang Cao, Hong Jiang, Jie Yao |
| 2020 | Toward Large-Scale Image Segmentation on Summit. | Sudip K. Seal, Seung-Hwan Lim, Dali Wang, Jacob D. Hinkle, Dalton D. Lunga, Aristeidis Tsaris |
| 2020 | The Art of CPU-Pinning: Evaluating and Improving the Performance of Virtualization and Containerization Platforms. | Davood Ghatreh Samani, Chavit Denninnart, Josef Bacik, Mohsen Amini Salehi |
| 2020 | Enabling performance portability of data-parallel OpenMP applications on asymmetric multicore processors. | Juan Carlos Saez, Fernando Castro, Manuel Prieto-Matas |
| 2020 | Polo: Receiver-Driven Congestion Control for Low Latency over Commodity Network Fabric. | Chang Ruan, Jianxin Wang, Wanchun Jiang, Tao Zhang |
| 2020 | Optimizing Linearizable Bulk Operations on Data Structures. | Matthew Rodriguez, Michael F. Spear |
| 2020 | First Time Miss : Low Overhead Mitigation for Shared Memory Cache Side Channels. | Kartik Ramkrishnan, Stephen McCamant, Pen-Chung Yew, Antonia Zhai |
| 2020 | SPECcast: A Methodology for Fast Performance Evaluation with SPEC CPU 2017 Multiprogrammed Workloads. | Pablo Prieto, Pablo Abad Fidalgo, Jose Angel Herrero, Jos-ngel Gregorio, Valentin Puente |
| 2020 | DNNARA: A Deep Neural Network Accelerator using Residue Arithmetic and Integrated Photonics. | Jiaxin Peng, Yousra Al-Kabani, Shuai Sun, Volker J. Sorger, Tarek A. El-Ghazawi |
| 2020 | Algorithm-Based Checkpoint-Recovery for the Conjugate Gradient Method. | Carlos Pachajoa, Christina Pacher, Markus Levonyak, Wilfried N. Gansterer |
| 2020 | Generating Robust Parallel Programs via Model Driven Prediction of Compiler Optimizations for Non-determinism. | Girish Mururu, Kaushik Ravichandran, Ada Gavrilovska, Santosh Pande |
| 2020 | Experiences on the characterization of parallel applications in embedded systems with Extrae/Paraver. | Adrian Munera, Sara Royuela, Germn Llort, Estanislao Mercadal, Franck Wartel, Eduardo Quiones |
| 2020 | Fast Spectral Graph Layout on Multicore Platforms. | Ashirbad Mishra, Shad Kirmani, Kamesh Madduri |
| 2020 | Selective Coflow Completion for Time-sensitive Distributed Applications with Poco. | Shouxi Luo, Pingzhi Fan, Huanlai Xing, Hongfang Yu |
| 2020 | Efficient Block Algorithms for Parallel Sparse Triangular Solve. | Zhengyang Lu, Yuyao Niu, Weifeng Liu |
| 2020 | Large-scale Simulations of Peridynamics on Sunway Taihulight Supercomputer. | Xinyuan Li, Huang Ye, Jian Zhang |
| 2020 | Deep Reinforcement Learning based Elasticity-compatible Heterogeneous Resource Management for Time-critical Computing. | Zixia Liu, Liqiang Wang, Gang Quan |
| 2020 | Memory-Centric Communication Mechanism for Real-time Autonomous Navigation Applications. | Wei Liu, Yifan Gong, Hao Wu, Jidong Zhai, Jiangming Jin |
| 2020 | A Rack-Aware Pipeline Repair Scheme for Erasure-Coded Distributed Storage Systems. | Tong Liu, Shakeel Alibhai, Xubin He |
| 2020 | Impact of Memory DoS Attacks on Cloud Applications and Real-Time Detection Schemes. | Zhuozhao Li, Tanmoy Sen, Haiying Shen, Mooi Choo Chuah |