| 2010 | An integer programming framework for optimizing shared memory use on GPUs. | Wenjing Ma, Gagan Agrawal |
| 2010 | A CG-based Poisson solver on a GPU-cluster. | Gnter Knittel |
| 2010 | Efficient algorithms to compute sleep schedules resulting in minimum delay routes in sensor networks. | Ajinkya Kher, Akshay Khatavkar, Shruti Bamb, Vaishali P. Sadaphal |
| 2010 | CUDACL: A tool for CUDA and OpenCL programmers. | Ferosh Jacob, David Whittaker, Sagar Thapaliya, Purushotham V. Bangalore, Marjan Mernik, Jeff Gray |
| 2010 | Approaches for parallelizing reductions on modern GPUs. | Xin Huo, Vignesh T. Ravi, Wenjing Ma, Gagan Agrawal |
| 2010 | Automatic dataflow application tuning for heterogeneous systems. | Timothy D. R. Hartley, Erik Saule, mit V. atalyrek |
| 2010 | A reliable data transport protocol for partitioned actors in Wireless Sensor and Actor Networks. | Nikhil Handigol, Kandasamy Selvaradjou, C. Siva Ram Murthy |
| 2010 | Anomaly detection in large-scale coalition clusters for dependability assurance. | Qiang Guan, Derek Smith, Song Fu |
| 2010 | Balanced stream assignment for service facility. | Rahul Garg, Perwez Shahabuddin, Akshat Verma |
| 2010 | A space-efficient parallel algorithm for computing betweenness centrality in distributed memory. | Nick Edmonds, Torsten Hoefler, Andrew Lumsdaine |
| 2010 | A study of memory-aware scheduling in message driven parallel programs. | Isaac Dooley, Chao Mei, Jonathan Lifflander, Laxmikant V. Kal |
| 2010 | Fair bandwidth allocation in wireless mobile environment using max-flow. | Sourav Kumar Dandapat, Bivas Mitra, Niloy Ganguly, Romit Roy Choudhury |
| 2010 | Diagnosing the root-causes of failures from cluster log files. | Edward Chuah, Shyh-Hao Kuo, Paul Hiew, William-Chandra Tjhi, Gary Kee Khoon Lee, John L. Hammond, Marek T. Michalewicz, Terence Hung, James C. Browne |
| 2010 | Dynamic social grouping based routing in a Mobile Ad-Hoc network. | Roy Cabaniss, Sanjay Madria, George Rush, Abbey Trotta, Srinivasa S. Vulli |
| 2010 | Parallel Sparse Matrix Vector Multiplication using greedy extraction of boxes. | Dhananjay Brahme, Binit Ranjan Mishra, Anup Barve |
| 2010 | Automated mapping of regular communication graphs on mesh interconnects. | Abhinav Bhatele, Gagan Raj Gupta, Laxmikant V. Kal, I-Hsin Chung |
| 2010 | Link-heterogeneity vs. node-heterogeneity in clusters. | Olivier Beaumont, Arnold L. Rosenberg |
| 2010 | Low-overhead diskless checkpoint for hybrid computing systems. | Leonardo Arturo Bautista-Gomez, Akira Nukada, Naoya Maruyama, Franck Cappello, Satoshi Matsuoka |
| 2010 | GRS - GPU radix sort for multifield records. | Shibdas Bandyopadhyay, Sartaj Sahni |
| 2010 | Reducing network load in large-scale, Peer-to-Peer Virtual Environments with 3D Voronoi Diagrams. | Mahathir Almashor, Ibrahim Khalil |
| 2009 | High search performance, small document index: P2P search can have both. | Yingwu Zhu, Haiying Shen |
| 2009 | Speculative p-DFAs for parallel XML parsing. | Ying Zhang, Yinfei Pan, Kenneth Chiu |
| 2009 | Terascale chip multiprocessor memory hierarchy and programming model. | Shoumeng Yan, Xiaocheng Zhou, Ying Gao, Hu Chen, Sai Luo, Peinan Zhang, Naveen Cherukuri, Ronny Ronen, Bratin Saha |
| 2009 | Cache streamization for high performance stream processor. | Nan Wu, Mei Wen, Ju Ren, Yi He, Changqing Xun, Wei Wu, Chunyuan Zhang |
| 2009 | Statistical workload shaping for storage systems. | Hui Wang, Peter J. Varman |