| 2013 | Ligra: a lightweight graph processing framework for shared memory. | Julian Shun, Guy E. Blelloch |
| 2013 | A compiler infrastructure for embedded heterogeneous MPSoCs. | Weihua Sheng, Stefan Schrmans, Maximilian Odendahl, Mark Bertsch, Vitaliy Volevach, Rainer Leupers, Gerd Ascheid |
| 2013 | Betweenness centrality: algorithms and implementations. | Dimitrios Prountzos, Keshav Pingali |
| 2013 | Parallel programming with big operators. | Changhee Park, Guy L. Steele Jr., Jean-Baptiste Tristan |
| 2013 | Scalable data race detection for partitioned global address space programs. | Chang-Seo Park, Koushik Sen, Costin Iancu |
| 2013 | Decomposition techniques for optimal design-space exploration of streaming applications. | Shobana Padmanabhan, Yixin Chen, Roger D. Chamberlain |
| 2013 | Morph algorithms on GPUs. | Rupesh Nasre, Martin Burtscher, Keshav Pingali |
| 2013 | Fast concurrent queues for x86 processors. | Adam Morrison, Yehuda Afek |
| 2013 | Distributed merge trees. | Dmitriy Morozov, Gunther H. Weber |
| 2013 | Parallel schedule synthesis for attribute grammars. | Leo A. Meyerovich, Matthew E. Torok, Eric Atkinson, Rastislav Bodk |
| 2013 | RaceFree: an efficient multi-threading model for determinism. | Kai Lu, Xu Zhou, Xiaoping Wang, Wenzhe Zhang, Gen Li |
| 2013 | Multi-level parallel computing of reverse time migration for seismic imaging on blue Gene/Q. | Ligang Lu, Karen A. Magerlein |
| 2013 | A generate-test-aggregate parallel programming library: systematic parallel programming for MapReduce. | Yu Liu, Kento Emoto, Zhenjiang Hu |
| 2013 | Data layout optimization for GPGPU architectures. | Jun Liu, Wei Ding, Ohyoung Jang, Mahmut T. Kandemir |
| 2013 | Adoption protocols for fanout-optimal fault-tolerant termination detection. | Jonathan Lifflander, Phil Miller, Laxmikant V. Kal |
| 2013 | Correct and efficient work-stealing for weak memory models. | Nhat Minh L, Antoniu Pop, Albert Cohen, Francesco Zappa Nardelli |
| 2013 | Empirical measurement of instruction level parallelism for four generations of ARM CPUs. | Martin J. Johnson, Ken A. Hawick |
| 2013 | A pattern-supported parallelization approach. | Ralf Jahr, Mike Gerdes, Theo Ungerer |
| 2013 | The tasks with effects model for safe concurrency. | Stephen Heumann, Vikram S. Adve, Shengjie Wang |
| 2013 | Scheduling directives for shared-memory many-core processor systems. | Oded Green, Yitzhak Birk |
| 2013 | Automatic problem size sensitive task partitioning on heterogeneous parallel systems. | Ivan Grasso, Klaus Kofler, Biagio Cosenza, Thomas Fahringer |
| 2013 | Ownership passing: efficient distributed memory programming on multi-core systems. | Andrew Friedley, Torsten Hoefler, Greg Bronevetsky, Andrew Lumsdaine, Ching-Chen Ma |
| 2013 | Low power cache architectures with hybrid approach of filtering unnecessary way accesses. | Lingjun Fan, Shinan Wang, Yasong Zheng, Weisong Shi, Dongrui Fan |
| 2013 | Expressing graph algorithms using generalized active messages. | Nick Edmonds, Jeremiah Willcock, Andrew Lumsdaine |
| 2013 | Towards an energy estimator for fault tolerance protocols. | Mohammed el Mehdi Diouri, Olivier Glck, Laurent Lefvre, Franck Cappello |