| 2013 | A Branch-and-Bound algorithm using multiple GPU-based LP solvers. | Xavier Meyer, Bastien Chopard, Paul Albuquerque |
| 2013 | Performance and energy consumption analysis of a seismic application for three different architectures intended for oil and gas industry. | Lucas T. Melo, Gilliano Ginno Silva de Menezes, Abel G. Silva-Filho, Manoel Eusbio de Lima |
| 2013 | A new parallel algorithm for connected components in dynamic graphs. | Robert McColl, Oded Green, David A. Bader |
| 2013 | Adding data parallelism to streaming pipelines for throughput optimization. | Peng Li, Kunal Agrawal, Jeremy Buhler, Roger D. Chamberlain |
| 2013 | Parallel branch-and-bound for two-stage stochastic integer optimization. | Akhil Langer, Ramprasad Venkataraman, Udatta S. Palekar, Laxmikant V. Kal |
| 2013 | Accelerating Strassen-Winograd's matrix multiplication algorithm on GPUs. | Pai-Wei Lai, Humayun Arafat, Venmugil Elango, Ponnuswamy Sadayappan |
| 2013 | Speculative dynamic vectorization to assist static vectorization in a HW/SW co-designed environment. | Rakesh Kumar, Alejandro Martnez, Antonio Gonzlez |
| 2013 | GPU-enabled efficient executions of radiation calculations in climate modeling. | Sai Kiran Korwar, Sathish S. Vadhiyar, Ravi S. Nanjundiah |
| 2013 | Performance evaluation of medical imaging algorithms on Intel | Jyotsna Khemka, Mrugesh R. Gajjar, Sharan Vaswani, Naga Vydyanathan, Rama Malladi, Vinutha S. V |
| 2013 | Efficient sparse matrix multiple-vector multiplication using a bitmapped format. | Ramaseshan Kannan |
| 2013 | GAGM: Genome assembly on GPU using mate pairs. | Ashutosh Jain, Anshuj Garg, Kolin Paul |
| 2013 | X10-based distributed and parallel betweenness centrality and its application to social analytics. | Charuwat Houngkaew, Toyotaro Suzumura |
| 2013 | Web-scale entity annotation using MapReduce. | Shashank Gupta, Varun Chandramouli, Soumen Chakrabarti |
| 2013 | Effects of phase imbalance on data center energy management. | Sushil Gupta, Ayan Banerjee, Zahra Abbasi, Sandeep K. S. Gupta |
| 2013 | Evaluation and enhancement of weather application performance on Blue Gene/Q. | Gurbinder Singh Gill, Vaibhav Saxena, Rashmi Mittal, Thomas George, Yogish Sabharwal, Lalit Dagar |
| 2013 | Loop level speculation in a task based programming model. | Rahulkumar Gayatri, Rosa M. Badia, Eduard Ayguad |
| 2013 | A dynamic schema to increase performance in many-core architectures through percolation operations. | Elkin Garcia, Daniel A. Orozco, Rishi Khan, Ioannis E. Venetis, Kelly Livingston, Guang R. Gao |
| 2013 | Exploring energy and performance behaviors of data-intensive scientific workflows on systems with deep memory hierarchies. | Marc Gamell, Ivan Rodero, Manish Parashar, Stephen W. Poole |
| 2013 | Benchmarking MIC architectures with Monte Carlo simulations of spin glass systems. | Alessandro Gabbana, Marcello Pivanti, Sebastiano Fabio Schifano, Raffaele Tripiccione |
| 2013 | iFlatLFS: Performance optimization for accessing massive small files. | Songling Fu, Chenlin Huang, Ligang He, Nadeem Chaudhary, Xiangke Liao, Shazhou Yang, Xiaochuan Wang, Bao Li |
| 2013 | LiPS: A cost-efficient data and task co-scheduler for MapReduce. | Moussa Ehsan, Yao Chen, Hui Kang, Radu Sion, Jennifer L. Wong |
| 2013 | Minimization of cloud task execution length with workload prediction errors. | Sheng Di, Cho-Li Wang |
| 2013 | Can GPUs sort strings efficiently? | Aditya Deshpande, P. J. Narayanan |
| 2013 | MaSiF: Machine learning guided auto-tuning of parallel skeletons. | Alexander Collins, Christian Fensch, Hugh Leather, Murray Cole |
| 2013 | Analyzing the performance impact of authorization constraints and optimizing the authorization methods for workflows. | Nadeem Chaudhary, Ligang He |