| 2014 | Scale-out NUMA. | Stanko Novakovic, Alexandros Daglis, Edouard Bugnion, Babak Falsafi, Boris Grot |
| 2014 | Deterministic galois: on-demand, portable and parameterless. | Donald Nguyen, Andrew Lenharth, Keshav Pingali |
| 2014 | Data-parallel finite-state machines. | Todd Mytkowicz, Madanlal Musuvathi, Wolfram Schulte |
| 2014 | Price theory based power management for heterogeneous multi-cores. | Thannirmalai Somu Muthukaruppan, Anuj Pathania, Tulika Mitra |
| 2014 | Fence-free work stealing on bounded TSO processors. | Adam Morrison, Yehuda Afek |
| 2014 | Disengaged scheduling for fair, protected access to fast computational accelerators. | Konstantinos Menychtas, Kai Shen, Michael L. Scott |
| 2014 | Exploiting GPU Hardware Saturation for Fast Compiler Optimization. | Alberto Magni, Christophe Dubach, Michael F. P. O'Boyle |
| 2014 | Speculative hardware/software co-designed floating-point multiply-add fusion. | Marc Lupon, Enric Gibert, Grigorios Magklis, Sridhar Samudrala, Ral Martnez, Kyriakos Stavrou, David R. Ditzel |
| 2014 | ad-heap: an Efficient Heap Data Structure for Asymmetric Multicore Processors. | Weifeng Liu, Brian Vinter |
| 2014 | NVM duet: unified working memory and persistent store architecture. | Ren-Shuo Liu, De-Yu Shen, Chia-Lin Yang, Shun-Chih Yu, Cheng-Yuan Michael Wang |
| 2014 | SI-TM: reducing transactional memory abort rates through snapshot isolation. | Heiner Litz, David R. Cheriton, Amin Firoozshahian, Omid Azizi, John P. Stevenson |
| 2014 | K2: a mobile operating system for heterogeneous coherence domains. | Felix Xiaozhu Lin, Zhen Wang, Lin Zhong |
| 2014 | Locality-oblivious cache organization leveraging single-cycle multi-hop NoCs. | Woo-Cheol Kwon, Tushar Krishna, Li-Shiuan Peh |
| 2014 | Ubik: efficient cache sharing with strict qos for latency-critical workloads. | Harshad Kasture, Daniel Snchez |
| 2014 | Triple-A: a Non-SSD based autonomic all-flash array for high performance storage systems. | Myoungsoo Jung, Wonil Choi, John Shalf, Mahmut T. Kandemir |
| 2014 | Application-aware Memory System for Fair and Efficient Execution of Concurrent GPGPU Applications. | Adwait Jog, Evgeny Bolotin, Zvika Guz, Mike Parker, Stephen W. Keckler, Mahmut T. Kandemir, Chita R. Das |
| 2014 | Heterogeneous-race-free memory models. | Derek R. Hower, Blake A. Hechtman, Bradford M. Beckmann, Benedict R. Gaster, Mark D. Hill, Steven K. Reinhardt, David A. Wood |
| 2014 | RelaxReplay: record and replay for relaxed-consistency multiprocessors. | Nima Honarmand, Josep Torrellas |
| 2014 | Integrated 3D-stacked server designs for increasing physical density of key-value stores. | Anthony Gutierrez, Michael Cieslak, Bharan Giridhar, Ronald G. Dreslinski, Luis Ceze, Trevor N. Mudge |
| 2014 | Neuromorphic processing: a new frontier in scaling computer architecture. | Jeff Gehlhaar |
| 2014 | Efficient Instrumentation of GPGPU Applications Using Information Flow Analysis and Symbolic Execution. | Naila Farooqui, Karsten Schwan, Sudhakar Yalamanchili |
| 2014 | The benefit of SMT in the multi-core era: flexibility towards degrees of thread-level parallelism. | Stijn Eyerman, Lieven Eeckhout |
| 2014 | Power Modeling for Heterogeneous Processors. | Tahir Diop, Natalie D. Enright Jerger, Jason Helge Anderson |
| 2014 | Finding the limit: examining the potential and complexity of compilation scheduling for JIT-based runtime systems. | Yufei Ding, Mingzhou Zhou, Zhijia Zhao, Sarah Eisenstat, Xipeng Shen |
| 2014 | Quasar: resource-efficient and QoS-aware cluster management. | Christina Delimitrou, Christos Kozyrakis |