| 1999 | A quantitative architectural evaluation of synchronization algorithms and disciplines on ccNUMA systems: the case of the SGI Origin2000. | Dimitrios S. Nikolopoulos, Theodore S. Papatheodorou |
| 1999 | Exploiting SIMD parallelism in DSP and multimedia algorithms using the AltiVec technology. | Huy Nguyen, Lizy Kurian John |
| 1999 | A comparative analysis of four parallelisation schemes. | Nandini Mukherjee, John R. Gurd |
| 1999 | Dynamic removal of redundant computations. | Carlos Molina, Antonio Gonzlez, Jordi Tubella |
| 1999 | High-level semantic optimization of numerical codes. | Vijay Menon, Keshav Pingali |
| 1999 | Improving memory hierarchy performance for irregular applications. | John M. Mellor-Crummey, David B. Whalley, Ken Kennedy |
| 1999 | Thread fork/join techniques for multi-level parallelism exploitation in NUMA multiprocessors. | Xavier Martorell, Eduard Ayguad, Nacho Navarro, Julita Corbaln, Marc Gonzlez, Jess Labarta |
| 1999 | Improving the performance of bristled CC-NUMA systems using virtual channels and adaptivity. | Jos F. Martnez, Josep Torrellas, Jos Duato |
| 1999 | Increasing effective IPC by exploiting distant parallelism. | Ivan Martel, Daniel Ortega, Eduard Ayguad, Mateo Valero |
| 1999 | Clustered speculative multithreaded processors. | Pedro Marcuello, Antonio Gonzlez |
| 1999 | An affine partitioning algorithm to maximize parallelism and minimize communication. | Amy W. Lim, Gerald I. Cheong, Monica S. Lam |
| 1999 | An experimental evaluation of tiling and shackling for memory hierarchy management. | Induprakas Kodukula, Keshav Pingali, Robert Cox, Dror E. Maydan |
| 1999 | Symmetry and performance in consistency protocols. | Peter J. Keleher |
| 1999 | An integer linear programming approach for optimizing cache locality. | Mahmut T. Kandemir, Prithviraj Banerjee, Alok N. Choudhary, J. Ramanujam, Eduard Ayguad |
| 1999 | Communication conscious radix sort. | Daniel Jimnez-Gonzlez, Josep Llus Larriba-Pey, Juan J. Navarro |
| 1999 | Application scaling under shared virtual memory on a cluster of SMPs. | Dongming Jiang, Brian O'Kelley, Xiang Yu, Sanjeev Kumar, Angelos Bilas, Jaswinder Pal Singh |
| 1999 | Comparing the memory system performance of the HP V-class and SGI Origin 2000 multiprocessors using microbenchmarks and scientific applications. | Ravi R. Iyer, Nancy M. Amato, Lawrence Rauchwerger, Laxmi N. Bhuyan |
| 1999 | Shared virtual memory with automatic update support. | Liviu Iftode, Matthias A. Blumrich, Cezary Dubnicki, David L. Oppenheimer, Jaswinder Pal Singh, Kai Li |
| 1999 | A new method to make communication latency uniform: distributed routing balancing. | Daniel Franco, Indhira Garcs, Emilio Luque |
| 1999 | A comparison of two approaches for independent scaling up of processing and communication capacities in multicomputer networks. | A. Ferre-Vilaplana, Jos M. Bernabu-Aubn |
| 1999 | New shape analysis techniques for automatic parallelization of C codes. | Francisco Corbera, Rafael Asenjo, Emilio L. Zapata |
| 1999 | Parallel I/O for scientific applications on heterogeneous clusters: a resource-utilization approach. | Yong E. Cho, Marianne Winslett, Szu-Wen Kuo, Jonghyun Lee, Ying Chen |
| 1999 | Reducing branch misprediction penalties via dynamic control independence detection. | Yuan C. Chou, Jason Fung, John Paul Shen |
| 1999 | Cyclic dependence based data reference prediction. | Chi-Hung Chi, Jun-Li Yuan, Chin-Ming Cheung |
| 1999 | Reorganizing global schedules for register allocation. | Gang Chen, Michael D. Smith |