| 2012 | Network Endpoints for Clusters of SMPs. | Gabriel Ilie Tanase, Gheorghe Almsi, Hanhong Xue, Charles Archer |
| 2012 | Runtime Procedure for Energy Savings in Applications with Point-to-Point Communications. | Vaibhav Sundriyal, Masha Sosonkina, Alexander Gaenko |
| 2012 | Global Data Re-allocation via Communication Aggregation in Chapel. | Alberto Sanz, Rafael Asenjo, Juan Lpez, Rafael Larrosa, Angeles G. Navarro, Vassily Litvinov, Sung-Eun Choi, Bradford L. Chamberlain |
| 2012 | Divergence Analysis with Affine Constraints. | Diogo Sampaio, Rafael Martins de Souza, Caroline Collange, Fernando Magno Quinto Pereira |
| 2012 | Transactional Forwarding: Supporting Highly-Concurrent STM in Asynchronous Distributed Systems. | Mohamed M. Saad, Binoy Ravindran |
| 2012 | Using Heterogeneous Networks to Improve Energy Efficiency in Direct Coherence Protocols for Many-Core CMPs. | Alberto Ros, Ricardo Fernndez-Pascual, Manuel E. Acacio |
| 2012 | The Network Adapter: The Missing Link between MPI Applications and Network Performance. | Germn Rodrguez, Cyriel Minkenberg, Ronald P. Luijten, Ramn Beivide, Patrick Geoffray, Jess Labarta, Mateo Valero, Steve Poole |
| 2012 | Scalable Thread Scheduling in Asymmetric Multicores for Power Efficiency. | Rance Rodrigues, Arunachalam Annamalai, Israel Koren, Sandip Kundu |
| 2012 | Parallelizing Information Set Generation for Game Tree Search Applications. | Mark Richards, Abhishek Gupta, Osman Sarood, Laxmikant V. Kal |
| 2012 | Exploiting Phase-Change Memory in Cooperative Caches. | Luiz E. Ramos, Ricardo Bianchini |
| 2012 | On the Efficiency of Register File versus Broadcast Interconnect for Collective Communications in Data-Parallel Hardware Accelerators. | Ardavan Pedram, Andreas Gerstlauer, Robert A. van de Geijn |
| 2012 | FusedOS: Fusing LWK Performance with FWK Functionality in a Heterogeneous Environment. | Yoonho Park, Eric Van Hensbergen, Marius Hillenbrand, Todd Inglett, Bryan S. Rosenburg, Kyung Dong Ryu, Robert W. Wisniewski |
| 2012 | CSHARP: Coherence and SHaring Aware Cache Replacement Policies for Parallel Applications. | Biswabandan Panda, Shankar Balachandran |
| 2012 | Efficient Sorting on the Tilera Manycore Architecture. | Alessandro Morari, Antonino Tumeo, Oreste Villa, Simone Secchi, Mateo Valero |
| 2012 | Data and Instruction Uniformity in Minimal Multi-threading. | Teo Milanez, Caroline Collange, Fernando Magno Quinto Pereira, Wagner Meira Jr., Renato Ferreira |
| 2012 | Assessing Energy Efficiency of Fault Tolerance Protocols for HPC Systems. | Esteban Meneses, Osman Sarood, Laxmikant V. Kal |
| 2012 | Parallel Exact Inference on Multicore Using MapReduce. | Nam Ma, Yinglong Xia, Viktor K. Prasanna |
| 2012 | Efficiently Handling Memory Accesses to Improve QoS in Multicore Systems under Real-Time Constraints. | Jos Luis March, Salvador Petit, Julio Sahuquillo, Houcine Hassan, Jos Duato |
| 2012 | BTL: A Framework for Measuring and Modeling Energy in Memory Hierarchies. | Ioannis Manousakis, Dimitrios S. Nikolopoulos |
| 2012 | VPC: Scalable, Low Downtime Checkpointing for Virtual Clusters. | Peng Lu, Binoy Ravindran, Changsoo Kim |
| 2012 | Exploiting Concurrent GPU Operations for Efficient Work Stealing on Multi-GPUs. | Joo V. F. Lima, Thierry Gautier, Nicolas Maillard, Vincent Danjean |
| 2012 | Scalable Algorithms for Distributed-Memory Adaptive Mesh Refinement. | Akhil Langer, Jonathan Lifflander, Phil Miller, Kuo-Chuan Pan, Laxmikant V. Kal, Paul M. Ricker |
| 2012 | Low Overhead Instruction-Cache Modeling Using Instruction Reuse Profiles. | Muneeb Khan, Andreas Sembrant, Erik Hagersten |
| 2012 | Compression Speed Enhancements to LZO for Multi-core Systems. | Jason Kane, Qing Yang |
| 2012 | An OS-Hypervisor Infrastructure for Automated OS Crash Diagnosis and Recovery in a Virtualized Environment. | Joefon Jann, R. Sarma Burugula, Ching-Farn Eric Wu, Kaoutar El Maghraoui |