| 2003 | Enhancing memory level parallelism via recovery-free value prediction. | Huiyang Zhou, Thomas M. Conte |
| 2003 | A compiler approach for reducing data cache energy. | Wei Zhang, Mustafa Karaky, Mahmut T. Kandemir, Guangyu Chen |
| 2003 | Inter-procedural stacked register allocation for itanium® like architecture. | Liu Yang, Sun Chan, Guang R. Gao, Roy Ju, Guei-Yuan Lueh, Zhaoqing Zhang |
| 2003 | Placement of I/O servers to improve parallel I/O performance on switch-based clusters. | Jan-Jan Wu, Da-Wei Wang, Yih-Fang Lin |
| 2003 | Enhancing scalability of parallel structured AMR calculations. | Andrew M. Wissink, David Hysom, Richard D. Hornung |
| 2003 | Profile-guided I/O partitioning. | Yijian Wang, David R. Kaeli |
| 2003 | A new speculation technique to optimize floating-point performance while preserving bit-by-bit reproducibility. | Mikio Takeuchi, Hideaki Komatsu, Toshio Nakatani |
| 2003 | AEGIS: architecture for tamper-evident and tamper-resistant processing. | G. Edward Suh, Dwaine E. Clarke, Blaise Gassend, Marten van Dijk, Srinivas Devadas |
| 2003 | Predictive dynamic thermal management for multimedia applications. | Jayanth Srinivasan, Sarita V. Adve |
| 2003 | Keynote: Is there anything more to learn about high performance processors? | James E. Smith |
| 2003 | PowerHerd: dynamic satisfaction of peak power constraints in interconnection networks. | Li Shang, Li-Shiuan Peh, Niraj K. Jha |
| 2003 | Selecting long atomic traces for high coverage. | Roni Rosner, Micha Moffie, Yiannakis Sazeides, Ronny Ronen |
| 2003 | Inferential queueing and speculative push for reducing critical communication latencies. | Ravi Rajwar, Alain Kgi, James R. Goodman |
| 2003 | Partitioned first-level cache design for clustered microarchitectures. | Paul Racunas, Yale N. Patt |
| 2003 | Modeling and optimization of non-blocking checkpointing for optimistic simulation on myrinet clusters. | Francesco Quaglia, Andrea Santoro |
| 2003 | The impact of data dependence analysis on compilation and program parallelization. | Kleanthis Psarris, Konstantinos Kyriakopoulos |
| 2003 | Dynamic memory instruction bypassing. | Daniel Ortega, Eduard Ayguad, Mateo Valero |
| 2003 | High performance RDMA-based MPI implementation over InfiniBand. | Jiuxing Liu, Jiesheng Wu, Sushmitha P. Kini, Pete Wyckoff, Dhabaleswar K. Panda |
| 2003 | Compiler support for efficient processing of XML datasets. | Xiaogang Li, Renato Ferreira, Gagan Agrawal |
| 2003 | Reducing register ports using delayed write-back queues and operand pre-fetch. | Nam Sung Kim, Trevor N. Mudge |
| 2003 | Untitled record | Xiangmin Jiao, Michael T. Campbell, Michael T. Heath |
| 2003 | A performance analysis of the Berkeley UPC compiler. | Parry Husbands, Costin Iancu, Katherine A. Yelick |
| 2003 | Result checking in global computing systems. | Ccile Germain-Renaud, Nathalie Playez |
| 2003 | Performance characteristics of openMP constructs, and application benchmarks on a large symmetric multiprocessor. | Nathan R. Fredrickson, Ahmad Afsahi, Ying Qian |
| 2003 | Automatic fence insertion for shared memory multiprocessing. | Xing Fang, Jaejin Lee, Samuel P. Midkiff |