| 2013 | Optimizing Google's warehouse scale computers: The NUMA experience. | Lingjia Tang, Jason Mars, Xiao Zhang, Robert Hagmann, Robert Hundt, Eric Tune |
| 2013 | A novel system architecture for web scale applications using lightweight CPUs and virtualized I/O. | Kshitij Sudan, Saisanthosh Balakrishnan, Sean Lie, Min Xu, Dhiraj Mallick, Gary Lauterbach, Rajeev Balasubramonian |
| 2013 | MISE: Providing performance predictability and improving fairness in shared main memory systems. | Lavanya Subramanian, Vivek Seshadri, Yoongu Kim, Ben Jaiyen, Onur Mutlu |
| 2013 | Cache coherence for GPU architectures. | Inderpreet Singh, Arrvindh Shriraman, Wilson W. L. Fung, Mike O'Connor, Tor M. Aamodt |
| 2013 | Modeling performance variation due to cache sharing. | Andreas Sandberg, Andreas Sembrant, Erik Hagersten, David Black-Schaffer |
| 2013 | Sonic Millip3De: A massively parallel 3D-stacked accelerator for 3D ultrasound. | Richard Sampson, Ming Yang, Siyuan Wei, Chaitali Chakrabarti, Thomas F. Wenisch |
| 2013 | Energy-efficient interconnect via Router Parking. | Ahmad Samih, Ren Wang, Anil Krishna, Christian Maciocco, Tsung-Yuan Charlie Tai, Yan Solihin |
| 2013 | How to implement effective prediction and forwarding for fusable dynamic multicore architectures. | Behnam Robatmili, Dong Li, Hadi Esmaeilzadeh, Madhu Saravana Sibi Govindan, Aaron Smith, Andrew Putnam, Doug Burger, Stephen W. Keckler |
| 2013 | The dual-path execution model for efficient GPU control flow. | Minsoo Rhu, Mattan Erez |
| 2013 | Optimizing virtual machine scheduling in NUMA multicore systems. | Jia Rao, Kun Wang, Xiaobo Zhou, Cheng-Zhong Xu |
| 2013 | Rainbow: Efficient memory dependence recording with high replay parallelism for relaxed memory model. | Xuehai Qian, He Huang, Benjamn Sahelices, Depei Qian |
| 2013 | Bridging the semantic gap: Emulating biological neuronal behaviors with simple digital neurons. | Andrew Nere, Atif Hashmi, Mikko H. Lipasti, Giulio Tononi |
| 2013 | A case for Refresh Pausing in DRAM memory systems. | Prashant J. Nair, Chia-Chen Chou, Moinuddin K. Qureshi |
| 2013 | Macho: A failure model-oriented adaptive cache architecture to enable near-threshold voltage scaling. | Tayyeb Mahmood, Soontae Kim, Seokin Hong |
| 2013 | Reducing GPU offload latency via fine-grained CPU-GPU synchronization. | Daniel Lustig, Margaret Martonosi |
| 2013 | Exploring high-performance and energy proportional interface for phase change memory systems. | Zhongqi Li, Ruijin Zhou, Tao Li |
| 2013 | Enabling distributed generation powered sustainable high-performance data center. | Chao Li, Ruijin Zhou, Tao Li |
| 2013 | Tiered-latency DRAM: A low latency and low cost DRAM architecture. | Donghyuk Lee, Yoongu Kim, Vivek Seshadri, Jamie Liu, Lavanya Subramanian, Onur Mutlu |
| 2013 | Skinflint DRAM system: Minimizing DRAM chip writes for low power. | Yebin Lee, Soontae Kim, Seokin Hong, Jongmin Lee |
| 2013 | Breaking the on-chip latency barrier using SMART. | Tushar Krishna, Chia-Hsin Owen Chen, Woo-Cheol Kwon, Li-Shiuan Peh |
| 2013 | Layout-conscious random topologies for HPC off-chip interconnects. | Michihiro Koibuchi, Ikki Fujiwara, Hiroki Matsutani, Henri Casanova |
| 2013 | Improving multi-core performance using mixed-cell cache architecture. | Samira Manabi Khan, Alaa R. Alameldeen, Chris Wilkerson, Jaydeep P. Kulkarni, Daniel A. Jimnez |
| 2013 | SCRAP: Architecture for signature-based protection from Code Reuse Attacks. | Mehmet Kayaalp, Timothy Schmitt, Junaid Nomani, Dmitry Ponomarev, Nael B. Abu-Ghazaleh |
| 2013 | EnergySmart: Toward energy-efficient manycores for Near-Threshold Computing. | Ulya R. Karpuzcu, Abhishek A. Sinkar, Nam Sung Kim, Josep Torrellas |
| 2013 | Adaptive Reliability Chipkill Correct (ARCC). | Xun Jian, Rakesh Kumar |