| 1994 | MOB forms: a class of multilevel block algorithms for dense linear algebra operations. | Juan J. Navarro, Toni Juan, Toms Lang |
| 1994 | An evaluation of directory protocols for medium-scale shared-memory multiprocessors. | Shubhendu S. Mukherjee, Mark D. Hill |
| 1994 | An evaluation of a compiler optimization for improving the performance of a coherence directory. | Farnaz Mounes-Toussi, David J. Lilja, Zhiyuan Li |
| 1994 | A model for dataflow based vector execution. | William Marcus Miller, Walid A. Najjar, A. P. Wim Bhm |
| 1994 | Optimal local register allocation for a multiple-issue machine. | Waleed Meleis, Edward S. Davidson |
| 1994 | Evaluating automatic parallelization for efficient execution on shared-memory multiprocessors. | Kathryn S. McKinley |
| 1994 | Fast phylogenetic analysis on a massively parallel machine. | Hideo Matsuda, Gary J. Olsen, Ross A. Overbeek, Yukio Kaneda |
| 1994 | An optimal upper bound on the minimal completion time in distributed supercomputing. | Lars Lundberg, Hkan Lennerstad |
| 1994 | Exploiting cache affinity in software cache coherence. | Hui Li, Kenneth C. Sevcik |
| 1994 | A generalized vision of some parallel bidiagonal systems solvers. | Josep Llus Larriba-Pey, Juan J. Navarro, Oriol Roig, Angel Jorba |
| 1994 | Techniques to overlap computation and communication in irregular iterative applications. | Antonio Lain, Prithviraj Banerjee |
| 1994 | Performance of the CM-5 scalable file system. | Thomas T. Kwan, Daniel A. Reed |
| 1994 | On sparse matrix reordering for parallel factorization. | Bharat Kumar, P. Sadayappan, Chua-Huang Huang |
| 1994 | Distributed storage control unit for the Hitachi S-3800 multivector supercomputer. | Katsuyoshi Kitai, Tadaaki Isobe, Tadayuki Sakakibara, Shigeko Yazawa, Yoshiko Tamaki, Teruo Tanaka, Kouichi Ishii |
| 1994 | An approach to communication-efficient data redistribution. | S. D. Kaushik, Chua-Huang Huang, Rodney W. Johnson, P. Sadayappan |
| 1994 | Compilation techniques for block-cyclic distributions. | Seema Hiranandani, Ken Kennedy, John M. Mellor-Crummey, Ajay Sethi |
| 1994 | Performance analysis of a synchronous, circuit-switched interconnection cached network. | Vipul Gupta, Eugen Schenfeld |
| 1994 | Architecture implications of high-speed I/O for distributed-memory computers. | Thomas R. Gross, Peter Steenkiste |
| 1994 | Compiling performance models from parallel programs. | Arjan J. C. van Gemund |
| 1994 | The effectiveness of caches for vector processors. | Jeffrey D. Gee, Alan Jay Smith |
| 1994 | The parallel solution of nonsymmetric sparse linear systems using the H* reordering and an associated factorization. | Kyle A. Gallivan, Bret A. Marsolf, Harry A. G. Wijshoff |
| 1994 | Parallelisation of the SDEM distinct element stress analysis code on the KSR-1. | Gregory K. Egan, Graham D. Riley, J. Mark Bull |
| 1994 | Ultrasonic wave propagation on parallel machines. | Christophe Domain, J.-P. Grgoire, Bernadette Thomas |
| 1994 | Compiler techniques for maximizing fine-grain and coarse-grain parallelism in loops with uniform dependences. | Yeong-Sheng Chen, Sheng-De Wang, Chien-Min Wang |
| 1994 | An efficient approach to computing fixpoints for complex program analysis. | Li-Ling Chen, Williams Ludwell Harrison III |