| 2012 | A GPU-based high-throughput image retrieval algorithm. | Feiwen Zhu, Peng Chen, Donglei Yang, Weihua Zhang, Haibo Chen, Binyu Zang |
| 2012 | JaBEE: framework for object-oriented Java bytecode compilation and execution on graphics processor units. | Wojciech Zaremba, Yuan Lin, Vinod Grover |
| 2012 | Applying transactional memory to concurrency bugs. | Haris Volos, Andres Jaan Tack, Michael M. Swift, Shan Lu |
| 2012 | DejaVu: accelerating resource allocation in virtualized environments. | Nedeljko Vasic, Dejan M. Novakovic, Svetozar Miucin, Dejan Kostic, Ricardo Bianchini |
| 2012 | Architectural support for hypervisor-secure virtualization. | Jakub Szefer, Ruby B. Lee |
| 2012 | Enabling task-level scheduling on heterogeneous platforms. | Enqiang Sun, Dana Schaa, Richard Bagley, Norman Rubin, David R. Kaeli |
| 2012 | An update-aware storage system for low-locality update-intensive workloads. | Dilip Nijagal Simha, Maohua Lu, Tzi-cker Chiueh |
| 2012 | Paragon: collaborative speculative loop execution on GPU and CPU. | Mehrzad Samadi, Amir Hormati, Janghaeng Lee, Scott A. Mahlke |
| 2012 | Full system simulation of many-core heterogeneous SoCs using GPU and QEMU semihosting. | Shivani Raghav, Andrea Marongiu, Christian Pinto, David Atienza, Martino Ruggiero, Luca Benini |
| 2012 | Optimal task assignment in multithreaded processors: a statistical approach. | Petar Radojkovic, Vladimir Cakarevic, Miquel Moret, Javier Verd, Alex Pajuelo, Francisco J. Cazorla, Mario Nemirovsky, Mateo Valero |
| 2012 | SIMD defragmenter: efficient ILP realization on data-parallel architectures. | Yongjun Park, Sangwon Seo, Hyunchul Park, Hyoun Kyu Cho, Scott A. Mahlke |
| 2012 | Chameleon: operating system support for dynamic processors. | Sankaralingam Panneerselvam, Michael M. Swift |
| 2012 | Aikido: accelerating shared data dynamic analyses. | Marek Olszewski, Qin Zhao, David Koh, Jason Ansel, Saman P. Amarasinghe |
| 2012 | Continuous object access profiling and optimizations to overcome the memory wall and bloat. | Rei Odaira, Toshio Nakatani |
| 2012 | High performance 3-D FFT using multiple CUDA GPUs. | Akira Nukada, Yutaka Maruyama, Satoshi Matsuoka |
| 2012 | Introducing 'Bones': a parallelizing source-to-source compiler based on algorithmic skeletons. | Cedric Nugteren, Henk Corporaal |
| 2012 | Whole-system persistence. | Dushyanth Narayanan, Orion Hodson |
| 2012 | FLAT: a GPU programming framework to provide embedded MPI. | Takefumi Miyoshi, Hidetsugu Irie, Keigo Shima, Hiroki Honda, Masaaki Kondo, Tsutomu Yoshinaga |
| 2012 | A distributed data-parallel framework for analysis and visualization algorithm development. | Jeremy S. Meredith, Robert Sisneros, David Pugmire, Sean Ahern |
| 2012 | DreamWeaver: architectural support for deep sleep. | David Meisner, Thomas F. Wenisch |
| 2012 | Path-exploration lifting: hi-fi tests for lo-fi emulators. | Lorenzo Martignoni, Stephen McCamant, Pongsin Poosankam, Dawn Song, Petros Maniatis |
| 2012 | PocketWeb: instant web browsing for mobile devices. | Dimitrios Lymberopoulos, Oriana Riva, Karin Strauss, Akshay Mittal, Alexandros Ntoulas |
| 2012 | Reflex: using low-power processors in smartphones without knowing them. | Felix Xiaozhu Lin, Zhen Wang, Robert LiKamWa, Lin Zhong |
| 2012 | Efficient sequential consistency via conflict ordering. | Changhui Lin, Vijay Nagarajan, Rajiv Gupta, Bharghava Rajaram |
| 2012 | Region scheduling: efficiently using the cache architectures via page-level affinity. | Min Lee, Karsten Schwan |