| 2019 | Blockchain abstract data type: poster. | Emmanuelle Anceaume, Antonella Del Pozzo, Romaric Ludinard, Maria Potop-Butucaru, Sara Tucci Piergiovanni |
| 2019 | Provably and practically efficient granularity control. | Umut A. Acar, Vitaly Aksenov, Arthur Charguraud, Mike Rainey |
| 2019 | Implementing parallel and concurrent tree structures. | Yihan Sun, Guy E. Blelloch |
| 2018 | FlashR: parallelize and scale R for machine learning using SSDs. | Da Zheng, Disa Mhembere, Joshua T. Vogelstein, Carey E. Priebe, Randal C. Burns |
| 2018 | Bridging the gap between deep learning and sparse matrix format selection. | Yue Zhao, Jiajia Li, Chunhua Liao, Xipeng Shen |
| 2018 | SIMD code generation for stencils on brick decompositions. | Tuowen Zhao, Mary W. Hall, Protonu Basu, Samuel Williams, Hans Johansen |
| 2018 | Initial Steps toward Making GPU a First-Class Computing Resource: Sharing and Resource Management. | Jun Yang |
| 2018 | Efficient parallel determinacy race detection for two-dimensional dags. | Yifan Xu, I-Ting Angelina Lee, Kunal Agrawal |
| 2018 | VerifiedFT: a verified, high-performance precise dynamic race detector. | James R. Wilcox, Cormac Flanagan, Stephen N. Freund |
| 2018 | MaxPair: Enhance OpenCL Concurrent Kernel Execution by Weighted Maximum Matching. | Yuan Wen, Michael F. P. O'Boyle, Christian Fensch |
| 2018 | Interval-based memory reclamation. | Haosen Wen, Joseph Izraelevitz, Wentao Cai, H. Alan Beadle, Michael L. Scott |
| 2018 | Lazygraph: lazy data coherency for replicas in distributed graph-parallel computation. | Lei Wang, Liangji Zhuang, Junhang Chen, Huimin Cui, Fang Lv, Ying Liu, Xiaobing Feng |
| 2018 | Superneurons: dynamic GPU memory management for training deep neural networks. | Linnan Wang, Jinmian Ye, Yiyang Zhao, Wei Wu, Ang Li, Shuaiwen Leon Song, Zenglin Xu, Tim Kraska |
| 2018 | swSpTRSV: a fast sparse triangular solve with sparse level tile layout on sunway architectures. | Xinliang Wang, Weifeng Liu, Wei Xue, Li Wu |
| 2018 | Intra-Task Parallelism in Automotive Real-Time Systems. | Remko van Wagensveld, Tobias Wgemann, Niklas Hehenkamp, Ramin Tavakoli Kolagari, Ulrich Margull, Ralph Mader |
| 2018 | A microbenchmark to study GPU performance models. | Vasily Volkov |
| 2018 | Reduction to Band Form for the Singular Value Decomposition on Graphics Accelerators. | Andrs E. Toms, Rafael Rodrguez-Snchez, Sandra Cataln, Enrique S. Quintana-Ort |
| 2018 | vSensor: leveraging fixed-workload snippets of programs for performance variance detection. | Xiongchao Tang, Jidong Zhai, Xuehai Qian, Bingsheng He, Wei Xue, Wenguang Chen |
| 2018 | PAM: parallel augmented maps. | Yihan Sun, Daniel Ferizovic, Guy E. Blelloch |
| 2018 | Ikra-Cpp: A C++/CUDA DSL for Object-Oriented Programming with Structure-of-Arrays Layout. | Matthias Springer, Hidehiko Masuhara |
| 2018 | An Evaluation of Vectorization and Cache Reuse Tradeoffs on Modern CPUs. | Du Shen, Milind Chabbi, Xu Liu |
| 2018 | SIMDization of Small Tensor Multiplication Kernels for Wide SIMD Vector Processors. | Christopher Rodrigues, Amarin Phaosawasdi, Peng Wu |
| 2018 | Automated code acceleration targeting heterogeneous openCL devices. | Heinrich Riebler, Gavin Vaz, Tobias Kenter, Christian Plessl |
| 2018 | A predictable synchronisation algorithm. | Stefan Reif, Wolfgang Schrder-Preikschat |
| 2018 | Register optimizations for stencils on GPUs. | Prashant Singh Rawat, Fabrice Rastello, Aravind Sukumaran-Rajam, Louis-Nol Pouchet, Atanas Rountev, P. Sadayappan |