| 2013 | Scalable statistics counters. | Dave Dice, Yossi Lev, Mark Moir |
| 2013 | Using hardware transactional memory to correct and simplify and readers-writer lock algorithm. | Dave Dice, Yossi Lev, Yujie Liu, Victor Luchangco, Mark Moir |
| 2013 | Relational algorithms for multi-bulk-synchronous processors. | Gregory Frederick Diamos, Haicheng Wu, Jin Wang, Ashwin Sanjay Lele, Sudhakar Yalamanchili |
| 2013 | Parallel suffix array and least common prefix for the GPU. | Mrinal Deo, Sean Keely |
| 2013 | Scalable deterministic replay in a parallel full-system emulator. | Yufei Chen, Haibo Chen |
| 2013 | Online-ABFT: an online algorithm based fault tolerance scheme for soft error detection in iterative methods. | Zizhong Chen |
| 2013 | ZOOMM: a parallel web browser engine for multicore mobile devices. | Calin Cascaval, Seth Fowler, Pablo Montesinos-Ortego, Wayne Piekarski, Mehrdad Reshadi, Behnam Robatmili, Michael Weber, Vrajesh Bhavsar |
| 2013 | Runtime elision of transactional barriers for captured memory. | Fernando Miguel Carvalho, Joo P. Cachopo |
| 2013 | NUMA-aware reader-writer locks. | Irina Calciu, David Dice, Yossi Lev, Victor Luchangco, Virendra J. Marathe, Nir Shavit |
| 2013 | TeamWork: synchronizing threads globally to detect real deadlocks for multithreaded programs. | Yan Cai, Ke Zhai, Shangru Wu, Wing Kwong Chan |
| 2013 | Auto-tuning methodology to represent landform attributes on multicore and multi-GPU systems. | Murilo Boratto, Pedro Alonso, Domingo Gimnez, Marcos Barreto, Karolyne Oliveira |
| 2013 | Bulk synchronous visualization. | Lars Ailo Bongo |
| 2013 | TigerQuoll: parallel event-based JavaScript. | Daniele Bonetta, Walter Binder, Cesare Pautasso |
| 2013 | Data-only flattening for nested data parallelism. | Lars Bergstrom, Matthew Fluet, Mike Rainey, John H. Reppy, Stephen Rosen, Adam Shaw |
| 2013 | From relational verification to SIMD loop synthesis. | Gilles Barthe, Juan Manuel Crespo, Sumit Gulwani, Csar Kunz, Mark Marron |
| 2013 | Programming with hardware lock elision. | Yehuda Afek, Amir Levy, Adam Morrison |
| 2013 | Scheduling parallel programs by work stealing with private deques. | Umut A. Acar, Arthur Charguraud, Mike Rainey |
| 2013 | Work-stealing with configurable scheduling strategies. | Martin Wimmer, Daniel Cederman, Jesper Larsson Trff, Philippas Tsigas |
| 2012 | GPU-based NFA implementation for memory efficient high speed regular expression matching. | Yuan Zu, Ming Yang, Zhonghu Xu, Lin Wang, Xin Tian, Kunyang Peng, Qunfeng Dong |
| 2012 | An overview of Medusa: simplified graph processing on GPUs. | Jianlong Zhong, Bingsheng He |
| 2012 | LHlf: lock-free linear hashing (poster paper). | Donghui Zhang, Per-ke Larson |
| 2012 | Shared work list: hacking amorphous data parallelism in UPC. | Shixiong Xu, Li Chen |
| 2012 | RACECAR: a heuristic for automatic function specialization on multi-core heterogeneous systems. | John Robert Wernsing, Greg Stitt |
| 2012 | BDDT: : block-level dynamic dependence analysis for deterministic task-based parallelism. | George Tzenakis, Angelos Papatriantafyllou, John Kesapides, Polyvios Pratikakis, Hans Vandierendonck, Dimitrios S. Nikolopoulos |
| 2012 | Wait-free linked-lists. | Shahar Timnat, Anastasia Braginsky, Alex Kogan, Erez Petrank |