| 1999 | Nonlinear array layouts for hierarchical memory systems. | Siddhartha Chatterjee, Vibhor V. Jain, Alvin R. Lebeck, Shyam Mundhra, Mithuna Thottethodi |
| 1999 | A tile selection algorithm for data locality and cache interference. | Jacqueline Chame, Sungdo Moon |
| 1999 | Problem space promotion and its evaluation as a technique for efficient parallel computation. | Bradford L. Chamberlain, E. Christopher Lewis, Lawrence Snyder |
| 1999 | A graphic parallelizing environment for user-compiler interaction. | Claudia Roberta Calidonna, Maurizio Giordano, Mario Mango Furnari |
| 1999 | Microservers: a new memory semantics for massively parallel computing. | Jay B. Brockman, Peter M. Kogge, Thomas L. Sterling, Vincent W. Freeh, Shannon K. Kuntz |
| 1999 | Performance impact of proxies in data intensive client-server applications. | Michael D. Beynon, Alan Sussman, Joel H. Saltz |
| 1998 | Distributed Data Structure Design for Scientific Computation. | Jan-Jan Wu, Pangfeng Liu |
| 1998 | A Performance Study of Out-of-order Vector Architectures and Short Registers. | Luis Villa, Roger Espasa, Mateo Valero |
| 1998 | Evaluation of High Performance Multicache Parallel Texture Mapping. | Alexis Vartanian, Jean-Luc Bchennec, Nathalie Drach-Temam |
| 1998 | Dependence Driven Execution for Multiprogrammed Multiprocessor. | Suvas Vajracharya, Dirk Grunwald |
| 1998 | Highly Efficient Implementation of MPI Point-to-Point Communication Using Remote Memory Operations. | Osamu Tatebe, Yuetsu Kodama, Satoshi Sekiguchi, Yoshinori Yamaguchi |
| 1998 | Integer Sorting on Shared-Memory Vector Parallel Computers. | Kenji Suehiro, Hitoshi Murai, Yoshiki Seo |
| 1998 | Application Level Scheduling of Gene Sequence Comparison on Metacomputers. | Neil T. Spring, Richard Wolski |
| 1998 | Measuring the Effectiveness of Automatic Parallelization in SUIF. | Byoungro So, Sungdo Moon, Mary W. Hall |
| 1998 | Load Balanced Parallel Radix Sort. | Andrew Sohn, Yuetsu Kodama |
| 1998 | Loop Fusion in High Performance Fortran. | Gerald Roth, Ken Kennedy |
| 1998 | Utilizing Reuse Information in Data Cache Management. | Jude A. Rivers, Edward S. Tam, Gary S. Tyson, Edward S. Davidson, Matthew K. Farrens |
| 1998 | Eliminating Conflict Misses for High Performance Architectures. | Gabriel Rivera, Chau-Wen Tseng |
| 1998 | The Role of Associativity and Commutativity in the Detection and Transformation of Loop-level Parallelism. | William M. Pottenger |
| 1998 | Kernel-level Scheduling for the Nano-threads Programming Model. | Eleftherios D. Polychronopoulos, Xavier Martorell, Dimitrios S. Nikolopoulos, Jess Labarta, Theodore S. Papatheodorou, Nacho Navarro |
| 1998 | The Design and Implementation of Zero Copy MPI Using Commodity Hardware with a High Performance Network. | Francis O'Carroll, Hiroshi Tezuka, Atsushi Hori, Yutaka Ishikawa |
| 1998 | Prefetching on the Cray-T3E. | Matthias M. Mller, Thomas M. Warschko, Walter F. Tichy |
| 1998 | Predicated Array Data-flow Analysis for Run-time Parallelization. | Sungdo Moon, Mary W. Hall, Brian R. Murphy |
| 1998 | MBCF: A Protected and Virtualized High-Speed User-Level Memory-Based Communication Facility. | Takashi Matsumoto, Kei Hiraki |
| 1998 | Speculative Multithreaded Processors. | Pedro Marcuello, Antonio Gonzlez, Jordi Tubella |