| 1999 | Access Order and Effective Bandwidth for Streams on a Direct Rambus Memory. | Sung I. Hong, Sally A. McKee, Maximo H. Salinas, Robert H. Klenke, James H. Aylor, William A. Wulf |
| 1999 | Out-of-Order Execution may not be Cost-Effective on Processors Featuring Simultaneous Multithreading. | Sbastien Hily, Andr Seznec |
| 1999 | WildFire: A Scalable Path for SMPs. | Erik Hagersten, Michael Koster |
| 1999 | Memory Hierarchy Considerations for Fast Transpose and Bit-Reversals. | Kang Su Gatlin, Larry Carter |
| 1999 | Distributed Modulo Scheduling. | Marcio Merino Fernandes, Josep Llosa, Nigel P. Topham |
| 1999 | Parallel Dispatch Queue: A Queue-Based Programming Abstraction to Parallelize Fine-Grain Communication Protocols. | Babak Falsafi, David A. Wood |
| 1999 | Comparative Evaluation of Fine- and Coarse-Grain Approaches for Software Distributed Shared Memory. | Sandhya Dwarkadas, Kourosh Gharachorloo, Leonidas I. Kontothanassis, Daniel J. Scales, Michael L. Scott, Robert Stets |
| 1999 | Improving the Accuracy vs. Speed Tradeoff for Simulating Shared-Memory Multiprocessors with ILP Processors. | Murthy Durbhakula, Vijay S. Pai, Sarita V. Adve |
| 1999 | MMR: A High-Performance Multimedia Router - Architecture and Design Trade-Offs. | Jos Duato, Sudhakar Yalamanchili, Mara Blanca Caminero, Damon S. Love, Francisco J. Quiles |
| 1999 | A Performance Comparison of Homeless and Home-Based Lazy Release Consistency Protocols in Software Shared Memory. | Alan L. Cox, Eyal de Lara, Y. Charlie Hu, Willy Zwaenepoel |
| 1999 | Using Lamport Clocks to Reason about Relaxed Memory Models. | Anne Condon, Mark D. Hill, Manoj Plakal, Daniel J. Sorin |
| 1999 | Second Workshop on Computer Architecture Evaluation Using Commercial Workloads. | Russell M. Clapp, Ashwini K. Nanda, Josep Torrellas |
| 1999 | Impulse: Building a Smarter Memory Controller. | John B. Carter, Wilson C. Hsieh, Leigh Stoller, Mark R. Swanson, Lixin Zhang, Erik Brunvand, Al Davis, Chen-Chi Kuo, Ravindra Kuramkote, Michael A. Parker, Lambert Schaelicke, Terry Tateyama |
| 1999 | Dynamically Exploiting Narrow Width Operands to Improve Processor Power and Performance. | David M. Brooks, Margaret Martonosi |
| 1999 | Limits to the Performance of Software Shared Memory: A Layered Approach. | Angelos Bilas, Dongming Jiang, Yuanyuan Zhou, Jaswinder Pal Singh |
| 1998 | Hardware for Speculative Run-Time Parallelization in Distributed Shared-Memory Multiprocessors. | Ye Zhang, Lawrence Rauchwerger, Josep Torrellas |
| 1998 | Partial Sampling with Reverse State Reconstruction: A New Technique for Branch Predictor Performance Estimation. | Darren Erik Vengroff, Guang R. Gao |
| 1998 | The integrated computer engineering design (ICED) curriculum. | Augustus K. Uht |
| 1998 | Control Speculation in Multithreaded Processors through Dynamic Loop Detection. | Jordi Tubella, Antonio Gonzlez |
| 1998 | Performance Study of a Concurrent Multithreaded Processor. | Jenn-Yuan Tsai, Zhenzhen Jiang, Eric Ness, Pen-Chung Yew |
| 1998 | The Potential for Using Thread-Level Data Speculation to Facilitate Automatic Parallelization. | J. Gregory Steffan, Todd C. Mowry |
| 1998 | Using Multicast and Multithreading to Reduce Communication in Software DSM Systems. | Evan Speight, John K. Bennett |
| 1998 | Address Translation Mechanisms In Network Interfaces. | Ioannis Schoinas, Mark D. Hill |
| 1998 | Fine-Grain Software Distributed Shared Memory on SMP Clusters. | Daniel J. Scales, Kourosh Gharachorloo, Anshu Aggarwal |
| 1998 | Home-Based SVM Protocols for SMP Clusters: Design and Performance. | Rudrajit Samanta, Angelos Bilas, Liviu Iftode, Jaswinder Pal Singh |