| 2016 | vDNN: Virtualized deep neural networks for scalable, memory-efficient neural network design. | Minsoo Rhu, Natalia Gimelshein, Jason Clemons, Arslan Zulfiqar, Stephen W. Keckler |
| 2016 | Register sharing for equality prediction. | Arthur Perais, Fernando A. Endo, Andr Seznec |
| 2016 | Dictionary sharing: An efficient cache compression scheme for compressed caches. | Biswabandan Panda, Andr Seznec |
| 2016 | Reliability-Aware Task Scheduling using Clustered Replication for Multi-core Real-Time systems. | Alireza Namazi, Meisam Abdollahi, Saeed Safari, Siamak Mohammadi, Masoud Daneshtalab |
| 2016 | The microarchitecture of a real-time robot motion planning accelerator. | Sean Murray, William Floyd-Jones, Ying Qi, George Dimitri Konidaris, Daniel J. Sorin |
| 2016 | Improved Flow Control for Minimal Fully Adaptive Routing in 2D Mesh NoC. | Alireza Monemi, Chia Yee Ooi, Muhammad Nadzir Marsono, Maurizio Palesi |
| 2016 | Chameleon: Versatile and practical near-DRAM acceleration architecture for large memory systems. | Hadi Asghari Moghaddam, Young Hoon Son, Jung Ho Ahn, Nam Sung Kim |
| 2016 | The Bunker Cache for spatio-value approximation. | Joshua San Miguel, Jorge Albericio, Natalie D. Enright Jerger, Aamer Jaleel |
| 2016 | A MAC protocol for Reliable Broadcast Communications in Wireless Network-on-Chip. | Albert Mestres, Sergi Abadal, Josep Torrellas, Eduard Alarcn, Albert Cabellos-Aparicio |
| 2016 | Keynotes: Internet of Things: History and hype, technology and policy. | Margaret Martonosi |
| 2016 | Low-cost soft error resilience with unified data verification and fine-grained recovery for acoustic sensor based detection. | Qingrui Liu, Changhee Jung, Dongyoon Lee, Devesh Tiwari |
| 2016 | Extending Gem5-Garnet for Efficient and Accurate Trace-driven NoC Simulation. | Ren-Min Li, Chung-Ta King, Bhaskar Das |
| 2016 | PoisonIvy: Safe speculation for secure memory. | Tamara Silbergleit Lehman, Andrew D. Hilton, Benjamin C. Lee |
| 2016 | Delegated persist ordering. | Aasheesh Kolli, Jeff Rosen, Stephan Diestelhorst, Ali G. Saidi, Steven Pelley, Sihang Liu, Peter M. Chen, Thomas F. Wenisch |
| 2016 | Path confidence based lookahead prefetching. | Jinchun Kim, Seth H. Pugsley, Paul V. Gratz, A. L. Narasimha Reddy, Chris Wilkerson, Zeshan Chishti |
| 2016 | Contention-based congestion management in large-scale networks. | Gwangsun Kim, Changhyun Kim, Jiyun Jeong, Mike Parker, John Kim |
| 2016 | Quantifying and improving the efficiency of hardware-based mobile malware detectors. | Mikhail Kazdagli, Vijay Janapa Reddi, Mohit Tiwari |
| 2016 | pTask: A smart prefetching scheme for OS intensive applications. | Prathmesh Kallurkar, Smruti R. Sarangi |
| 2016 | Stripes: Bit-serial deep neural network computing. | Patrick Judd, Jorge Albericio, Tayler H. Hetherington, Tor M. Aamodt, Andreas Moshovos |
| 2016 | NEUTRAMS: Neural network transformation and co-design under neuromorphic hardware constraints. | Yu Ji, Youhui Zhang, Shuangchen Li, Ping Chi, Cihang Jiang, Peng Qu, Yuan Xie, Wenguang Chen |
| 2016 | Cache-emulated register file: An integrated on-chip memory architecture for high performance GPGPUs. | Naifeng Jing, Jianfei Wang, Fengfeng Fan, Wenkang Yu, Li Jiang, Chao Li, Xiaoyao Liang |
| 2016 | Data-centric execution of speculative parallel programs. | Mark C. Jeffrey, Suvinay Subramanian, Maleen Abeydeera, Joel S. Emer, Daniel Snchez |
| 2016 | Continuous shape shifting: Enabling loop co-optimization via near-free dynamic code rewriting. | Animesh Jain, Michael A. Laurenzano, Lingjia Tang, Jason Mars |
| 2016 | Concise loads and stores: The case for an asymmetric compute-memory architecture for approximation. | Animesh Jain, Parker Hill, Shih-Chieh Lin, Muneeb Khan, Md. Enamul Haque, Michael A. Laurenzano, Scott A. Mahlke, Lingjia Tang, Jason Mars |
| 2016 | Towards efficient server architecture for virtualized network function deployment: Implications and implementations. | Yang Hu, Tao Li |