| 2019 | Debugging Support for Pattern-Matching Languages and Accelerators. | Matthew Casias, Kevin Angstadt, Tommy Tracy II, Kevin Skadron, Westley Weimer |
| 2019 | Which Graph Representation to Select for Static Graph-Algorithms on a CUDA-capable GPU. | Thorsten Bla, Michael Philippsen |
| 2019 | ProbeGuard: Mitigating Probing Attacks Through Reactive Program Transformations. | Koustubha Bhat, Erik van der Kouwe, Herbert Bos, Cristiano Giuffrida |
| 2019 | AcMC | Subho S. Banerjee, Zbigniew T. Kalbarczyk, Ravishankar K. Iyer |
| 2019 | DCNS: Automated Detection Of Conservative Non-Sleep Defects in the Linux Kernel. | Jia-Ju Bai, Julia Lawall, Wende Tan, Shi-Min Hu |
| 2019 | PUMA: A Programmable Ultra-efficient Memristor-based Accelerator for Machine Learning Inference. | Aayush Ankit, Izzat El Hajj, Sai Rahul Chalamalasetti, Geoffrey Ndu, Martin Foltin, R. Stanley Williams, Paolo Faraboschi, Wen-mei W. Hwu, John Paul Strachan, Kaushik Roy, Dejan S. Milojicic |
| 2019 | FlatFlash: Exploiting the Byte-Accessibility of SSDs within a Unified Memory-Storage Hierarchy. | Ahmed H. M. O. Abulila, Vikram Sharma Mailthody, Zaid Qureshi, Jian Huang, Nam Sung Kim, Jinjun Xiong, Wen-Mei W. Hwu |
| 2019 | A Framework for Memory Oversubscription Management in Graphics Processing Units. | Chen Li, Rachata Ausavarungnirun, Christopher J. Rossbach, Youtao Zhang, Onur Mutlu, Yang Guo, Jun Yang |
| 2019 | uops.info: Characterizing Latency, Throughput, and Port Usage of Instructions on Intel Microarchitectures. | Andreas Abel, Jan Reineke |
| 2019 | PMTest: A Fast and Flexible Testing Framework for Persistent Memory Programs. | Sihang Liu, Yizhou Wei, Jishen Zhao, Aasheesh Kolli, Samira Manabi Khan |
| 2018 | Optimizing Deep Learning Workloads on ARM GPU with TVM. | Lanmin Zheng, Tianqi Chen |
| 2018 | Wonderland: A Novel Abstraction-Based Out-Of-Core Graph Processing System. | Mingxing Zhang, Yongwei Wu, Youwei Zhuo, Xuehai Qian, Chengying Huan, Kang Chen |
| 2018 | Liquid Silicon-Monona: A Reconfigurable Memory-Oriented Computing Fabric with Scalable Multi-Context Support. | Yue Zha, Jing Li |
| 2018 | Datasize-Aware High Dimensional Configurations Auto-Tuning of In-Memory Cluster Computing. | Zhibin Yu, Zhendong Bei, Xuehai Qian |
| 2018 | Filtering Translation Bandwidth with Virtual Caching. | Hongil Yoon, Jason Lowe-Power, Gurindar S. Sohi |
| 2018 | Sugar: Secure GPU Acceleration in Web Browsers. | Zhihao Yao, Zongheng Ma, Yingtong Liu, Ardalan Amiri Sani, Aparna Chandramowlishwaran |
| 2018 | Espresso: Brewing Java For More Non-Volatility with Non-volatile Memory. | Mingyu Wu, Ziming Zhao, Haoyu Li, Heting Li, Haibo Chen, Binyu Zang, Haibing Guan |
| 2018 | Watching for Software Inefficiencies with Witch. | Shasha Wen, Xu Liu, John Byrne, Milind Chabbi |
| 2018 | Enhancing Cross-ISA DBT Through Automatically Learned Translation Rules. | Wenwen Wang, Stephen McCamant, Antonia Zhai, Pen-Chung Yew |
| 2018 | Understanding and Auto-Adjusting Performance-Sensitive Configurations. | Shu Wang, Chi Li, Henry Hoffmann, Shan Lu, William Sentosa, Achmad Imam Kistijantoro |
| 2018 | Darwin: A Genomics Co-processor Provides up to 15, 000X Acceleration on Long Read Assembly. | Yatish Turakhia, Gill Bejerano, William J. Dally |
| 2018 | VAULT: Reducing Paging Overheads in SGX with Efficient Integrity Verification Structures. | Meysam Taassori, Ali Shafiee, Rajeev Balasubramonian |
| 2018 | LTRF: Enabling High-Capacity Register Files for GPUs via Hardware/Software Cooperative Register Prefetching. | Mohammad Sadrosadati, Amirhossein Mirhosseini, Seyed Borna Ehsani, Hamid Sarbazi-Azad, Mario Drumond, Babak Falsafi, Rachata Ausavarungnirun, Onur Mutlu |
| 2018 | Tigr: Transforming Irregular Graphs for GPU-Friendly Graph Processing. | Amir Hossein Nodehi Sabet, Junqiao Qiu, Zhijia Zhao |
| 2018 | Sulong, and Thanks for All the Bugs: Finding Errors in C Programs by Abstracting from the Native Execution Model. | Manuel Rigger, Roland Schatz, Ren Mayrhofer, Matthias Grimmer, Hanspeter Mssenbck |