| 2009 | DFTL: a flash translation layer employing demand-based selective caching of page-level address mappings. | Aayush Gupta, Youngjae Kim, Bhuvan Urgaonkar |
| 2009 | An evaluation of the TRIPS computer system. | Mark Gebhart, Bertrand A. Maher, Katherine E. Coons, Jeffrey R. Diamond, Paul Gratz, Mario Marino, Nitya Ranganathan, Behnam Robatmili, Aaron Smith, James H. Burrill, Stephen W. Keckler, Doug Burger, Kathryn S. McKinley |
| 2009 | Accelerating linpack with CUDA on heterogenous clusters. | Massimiliano Fatica |
| 2009 | Per-thread cycle accounting in SMT processors. | Stijn Eyerman, Lieven Eeckhout |
| 2009 | Anomaly-based bug prediction, isolation, and validation: an automated approach for software debugging. | Martin Dimitrov, Huiyang Zhou |
| 2009 | Understanding software approaches for GPGPU reliability. | Martin Dimitrov, Mike Mantor, Huiyang Zhou |
| 2009 | Early experience with a commercial hardware transactional memory implementation. | David Dice, Yossi Lev, Mark Moir, Daniel Nussbaum |
| 2009 | DMP: deterministic shared memory multiprocessing. | Joseph Devietti, Brandon Lucia, Luis Ceze, Mark Oskin |
| 2009 | Gordon: using flash memory to build fast, power-efficient clusters for data-intensive applications. | Adrian M. Caulfield, Laura M. Grupp, Steven Swanson |
| 2009 | Architectural support for SWAR text processing with parallel bit streams: the inductive doubling principle. | Robert D. Cameron, Dan Lin |
| 2009 | Phantom-BTB: a virtualized branch target buffer design. | Ioana Burcea, Andreas Moshovos |
| 2009 | Performance analysis of accelerated image registration using GPGPU. | Peter Bui, Jay B. Brockman |
| 2009 | Leak pruning. | Michael D. Bond, Kathryn S. McKinley |
| 2009 | Commutativity analysis for software parallelization: letting program transformations see the big picture. | Farhana Aleen, Nathan Clark |
| 2008 | Toward molecular programming with DNA. | Erik Winfree |
| 2008 | Adapting to intermittent faults in multicore systems. | Philip M. Wells, Koushik Chakraborty, Gurindar S. Sohi |
| 2008 | Tapping into the fountain of CPUs: on operating system support for programmable devices. | Yaron Weinsberg, Danny Dolev, Tal Anker, Muli Ben-Yehuda, Pete Wyckoff |
| 2008 | The mapping collector: virtual memory support for generational, parallel, and concurrent compaction. | Michal Wegiel, Chandra Krintz |
| 2008 | SoftSig: software-exposed hardware signatures for code analysis and optimization. | James Tuck, Wonsun Ahn, Luis Ceze, Josep Torrellas |
| 2008 | Feedback-driven threading: power-efficient and high-performance execution of multi-threaded workloads on CMPs. | M. Aater Suleman, Moinuddin K. Qureshi, Yale N. Patt |
| 2008 | Adaptive set pinning: managing shared caches in chip multiprocessors. | Shekhar Srikantaiah, Mahmut T. Kandemir, Mary Jane Irwin |
| 2008 | General and efficient locking without blocking. | Yannis Smaragdakis, Anthony Kay, Reimer Behrends, Michal Young |
| 2008 | Hardware counter driven on-the-fly request signatures. | Kai Shen, Ming Zhong, Sandhya Dwarkadas, Chuanpeng Li, Christopher Stewart, Xiao Zhang |
| 2008 | No "power" struggles: coordinated multi-level power management for the data center. | Ramya Raghavendra, Parthasarathy Ranganathan, Vanish Talwar, Zhikui Wang, Xiaoyun Zhu |
| 2008 | Communication optimizations for global multi-threaded instruction scheduling. | Guilherme Ottoni, David I. August |