| 2001 | Fast parallel in-memory 64-bit sorting. | Daniel Jimnez-Gonzlez, Juan J. Navarro, Josep Llus Larriba-Pey |
| 2001 | Eliminating redundancies in sum-of-product array computations. | Steven J. Deitz, Bradford L. Chamberlain, Lawrence Snyder |
| 2001 | A parallel algorithm for sparse symbolic LU factorization without pivoting on out-of-core matrices. | Michel Cosnard, Laura Grigori |
| 2001 | Improving Gang Scheduling through job performance analysis and malleability. | Julita Corbaln, Xavier Martorell, Jess Labarta |
| 2001 | Loop optimization for a class of memory-constrained computations. | Daniel Cociorva, J. W. Wilkins, Chi-Chung Lam, Gerald Baumgartner, J. Ramanujam, P. Sadayappan |
| 2001 | Optimizing strategies for telescoping languages: procedure strength reduction and procedure vectorization. | Arun Chauhan, Ken Kennedy |
| 2001 | Array language support for parallel sparse computation. | Bradford L. Chamberlain, Lawrence Snyder |
| 2001 | Global optimization techniques for automatic parallelization of hybrid applications. | Dhruva R. Chakrabarti, Prithviraj Banerjee |
| 2001 | Reducing the complexity of the issue logic. | Ramon Canal, Antonio Gonzlez |
| 2001 | A network of cellular automata for a landslide simulation. | Claudia Roberta Calidonna, Claudia Di Napoli, Maurizio Giordano, Mario Mango Furnari, Salvatore Di Gregorio |
| 2001 | Bringing together automatic differentiation and OpenMP. | H. Martin Bcker, Bruno Lang, Dieter an Mey, Christian H. Bischof |
| 2001 | Workload decomposition for particle simulation applications on hierarchical distributed-shared memory parallel systems with integration of HPF and OpenMP. | Sergio Briguglio, Beniamino Di Martino, Gregorio Vlad |
| 2001 | Towards the effective parallel computation of matrix pseudospectra. | Constantine Bekas, Effrosini Kokiopoulou, Ioannis Koutis, Efstratios Gallopoulos |
| 2001 | Evaluating the impact of memory system performance on software prefetching and locality optimizations. | Abdel-Hameed A. Badawy, Aneesh Aggarwal, Donald Yeung, Chau-Wen Tseng |
| 2001 | On the potential of tolerant region reuse for multimedia applications. | Carlos lvarez, Jess Corbal, Esther Salam, Mateo Valero |
| 2001 | Demonstrating the scalability of a molecular dynamics application on a Petaflop computer. | George S. Almsi, Calin Cascaval, Jos G. Castaos, Monty Denneau, Wilm E. Donath, Maria Eleftheriou, Mark Giampapa, C. T. Howard Ho, Derek Lieber, Jos E. Moreira, Dennis M. Newns, Marc Snir, Henry S. Warren Jr. |
| 2000 | Automatic compiler techniques for thread coarsening for multithreaded architectures. | Gary M. Zoppetti, Gagan Agrawal, Lori L. Pollock, Jos Nelson Amaral, Xinan Tang, Guang R. Gao |
| 2000 | A simulation-based study of scheduling mechanisms for a dynamic cluster environment. | Yanyong Zhang, Anand Sivasubramaniam, Jos E. Moreira, Hubertus Franke |
| 2000 | Hardware-only stream prefetching and dynamic access ordering. | Chengqiang Zhang, Sally A. McKee |
| 2000 | Adaptive reduction parallelization techniques. | Hao Yu, Lawrence Rauchwerger |
| 2000 | Push vs. pull: data movement for linked data structures. | Chia-Lin Yang, Alvin R. Lebeck |
| 2000 | Performance evaluation of the IBM SP and the Compaq AlphaServer SC. | Patrick H. Worley |
| 2000 | Performance analysis of distributed applications using automatic classification of communication inefficiencies. | Jeffrey S. Vetter |
| 2000 | A novel application development environment for large-scale scientific computations. | Xiaohui Shen, Wei-keng Liao, Alok N. Choudhary, Gokhan Memik, Mahmut T. Kandemir, Sachin More, George K. Thiruvathukal, Arti Singh |
| 2000 | Table size reduction for data value predictors by exploiting narrow width values. | Toshinori Sato, Itsujiro Arita |