| 2018 | Oversubscribed Command Queues in GPUs. | Sooraj Puthoor, Xulong Tang, Joseph Gross, Bradford M. Beckmann |
| 2018 | Cache-tries: concurrent lock-free hash tries with constant-time operations. | Aleksandar Prokopec |
| 2018 | Untitled record | Manuel Pter, Jesper Larsson Trff |
| 2018 | A Data Layout Transformation for Vectorizing Compilers. | Arsne Prard-Gayot, Richard Membarth, Philipp Slusallek, Simon Moll, Roland Leia, Sebastian Hack |
| 2018 | Transparent GPU memory management for DNNs. | Jung-Ho Park, Hyungmin Cho, Wookeun Jung, Jaejin Lee |
| 2018 | Investigating automatic vectorization for real-time 3D scene understanding. | Alexandru Nica, Emanuele Vespa, Pablo Gonzlez de Aledo, Paul H. J. Kelly |
| 2018 | Quantifying and reducing execution variance in STM via model driven commit optimization. | Girish Mururu, Ada Gavrilovska, Santosh Pande |
| 2018 | Usuba: Optimizing & Trustworthy Bitslicing Compiler. | Darius Mercadier, Pierre-variste Dagand, Lionel Lacassagne, Gilles Muller |
| 2018 | DisCVar: discovering critical variables using algorithmic differentiation for transient faults. | Harshitha Menon, Kathryn Mohror |
| 2018 | Combining PREM compilation and ILP scheduling for high-performance and predictable MPSoC execution. | Joel Matejka, Bjrn Forsberg, Michal Sojka, Zdenek Hanzlek, Luca Benini, Andrea Marongiu |
| 2018 | Understanding Parallelization Tradeoffs for Linear Pipelines. | Aristeidis Mastoras, Thomas R. Gross |
| 2018 | Layrub: layer-centric GPU memory reuse and data migration in extreme-scale deep learning systems. | Bo Liu, Wenbin Jiang, Hai Jin, Xuanhua Shi, Yang Ma |
| 2018 | Register-based implementation of the sparse general matrix-matrix multiplication on GPUs. | Junhong Liu, Xin He, Weifeng Liu, Guangming Tan |
| 2018 | High-performance genomic analysis framework with in-memory computing. | Xueqi Li, Guangming Tan, Bingchen Wang, Ninghui Sun |
| 2018 | Designing scalable FPGA architectures using high-level synthesis. | Johannes de Fine Licht, Michaela Blott, Torsten Hoefler |
| 2018 | Small SIMD Matrices for CERN High Throughput Computing. | Florian Lemaitre, Benjamin Couturier, Lionel Lacassagne |
| 2018 | HPVM: heterogeneous parallel virtual machine. | Maria Kotsifakou, Prakalp Srivastava, Matthew D. Sinclair, Rakesh Komuravelli, Vikram S. Adve, Sarita V. Adve |
| 2018 | Safe privatization in transactional memory. | Artem Khyzha, Hagit Attiya, Alexey Gotsman, Noam Rinetzky |
| 2018 | A scalable queue for work distribution on GPUs. | Bernhard Kerbl, Joerg H. Mueller, Michael Kenzel, Dieter Schmalstieg, Markus Steinberger |
| 2018 | Vectorization of a spectral finite-element numerical kernel. | Sylvain Jubertie, Fabrice Dupros, Florent De Martin |
| 2018 | Two concurrent data structures for efficient datalog query processing. | Herbert Jordan, Bernhard Scholz, Pavle Subotic |
| 2018 | VAIL: A Victim-Aware Cache Policy for Improving Lifetime of Hybrid Memory. | Youchuang Jia, Fang Zhou, Xiang Gao, Song Wu, Hai Jin, Xiaofei Liao, Pingpeng Yuan |
| 2018 | Optimizing N-dimensional, winograd-based convolution for manycore CPUs. | Zhen Jia, Aleksandar Zlateski, Frdo Durand, Kai Li |
| 2018 | Revealing parallel scans and reductions in sequential loops through function reconstruction. | Peng Jiang, Gagan Agrawal |
| 2018 | An effective fusion and tile size model for optimizing image processing pipelines. | Abhinav Jangda, Uday Bondhugula |