| 2014 | Scalable analysis of multicore data reuse and sharing. | Miquel Perics, Kenjiro Taura, Satoshi Matsuoka |
| 2014 | Author retrospective for bloom filtering cache misses for accurate data speculation and prefetching. | Jih-Kwon Peir, Shih-Chang Kevin Lai, Shih-Lien Lu, Jared Stark, Konrad Lai |
| 2014 | Load balancing n-body simulations with highly non-uniform density. | Olga Pearce, Todd Gamblin, Bronis R. de Supinski, Tom Arsenlis, Nancy M. Amato |
| 2014 | Reduction of operating system jitter caused by page reclaim. | Yoshihiro Oyama, Shun Ishiguro, Jun Murakami, Shin Sasaki, Ryo Matsumiya, Osamu Tatebe |
| 2014 | Author's retrospective for: improving the performance of speculatively parallel applications on the hydra CMP. | Kunle Olukotun, Lance Hammond, Mark Willey |
| 2014 | Evaluation of methods to integrate analysis into a large-scale shock shock physics code. | Ron A. Oldfield, Kenneth Moreland, Nathan Fabian, David H. Rogers |
| 2014 | Accelerating cache coherence mechanism with speculation. | Jun Ohno, Kei Hiraki |
| 2014 | Verifying micro-architecture simulators using event traces. | Hui Meen Nyew, Nilufer Onder, Soner nder, Zhenlin Wang |
| 2014 | Reducing energy consumption of NoC by router bypassing. | Takahiro Naruko |
| 2014 | Author retrospective improving data cache performance by pre-executing instructions under a cache miss. | Trevor N. Mudge |
| 2014 | Collective memory transfers for multi-core chips. | George Michelogiannakis, Alexander Williams, Samuel Williams, John Shalf |
| 2014 | Author retrospective: compilation techniques for block-cyclic distributions. | John M. Mellor-Crummey, Seema Hiranandani, Ajay Sethi |
| 2014 | Multi-stage coordinated prefetching for present-day processors. | Sanyam Mehta, Zhenman Fang, Antonia Zhai, Pen-Chung Yew |
| 2014 | Author retrospective for optimizing for parallelism and data locality. | Kathryn S. McKinley |
| 2014 | Author retrospective for search and replication in unstructured peer-to-peer networks. | Qin Lv, Pei Cao, Edith Cohen, Kai Li, Scott Shenker |
| 2014 | Revealing applications' access pattern in collective I/O for cache management. | Yin Lu, Yong Chen, Robert Latham, Yu Zhuang |
| 2014 | HPC for the human brain project. | Thomas Lippert |
| 2014 | Block value based insertion policy for high performance last-level caches. | Lingda Li, Junlin Lu, Xu Cheng |
| 2014 | Author's retrospective for array privatization for parallel execution of loops. | Zhiyuan Li |
| 2014 | Overhead of a decentralized gossip algorithm on the performance of HPC applications. | Ely Levy, Amnon Barak, Amnon Shiloh, Matthias Lieber, Carsten Weinhold, Hermann Hrtig |
| 2014 | An optimal distributed load balancing algorithm for homogeneous work units. | Akhil Langer |
| 2014 | Author retrospective for anatomy of a message in the alewife multiprocessor. | John Kubiatowicz |
| 2014 | Author retrospective for semantical interprocedural parallelization: an overview of the PIPS project. | Franois Irigoin, Pierre Jouvelot, Rmi Triolet |
| 2014 | On the conditions for efficient interoperability with threads: an experience with PGAS languages using cray communication domains. | Khaled Z. Ibrahim, Katherine A. Yelick |
| 2014 | A programming system for xeon phis with runtime SIMD parallelization. | Xin Huo, Bin Ren, Gagan Agrawal |