| 1991 | A parallel genetic algorithm for the graph partitioning problem. | El-Ghazali Talbi, Pierre Bessire |
| 1991 | Accurate modelling of interconnection networks in vector supercomputers. | J. E. Smith, W. R. Taylor |
| 1991 | Sequential description and parallel execution language DFCII dataflow supercomputers. | Satoshi Sekiguchi, Toshio Shimada, Kei Hiraki |
| 1991 | Supercomputers: where are the lost cycles? | Willi Schnauer, Hartmut Hfner |
| 1991 | Optimization of array accesses by collective loop transformations. | Vivek Sarkar, Guang R. Gao |
| 1991 | Uniform techniques for loop optimization. | William W. Pugh |
| 1991 | Extending the I test to direction vectors. | Kleanthis Psarris, Xiangyun Kong, David Klappholz |
| 1991 | The hierarchical task graph and its use in auto-scheduling. | Constantine D. Polychronopoulos |
| 1991 | An algorithm for the generalized singular value decomposition on massively parallel computers. | Haesun Park, Lars-Magnus Ewerbring |
| 1991 | Parallel programs and background load: efficiency studies with the PAR-Bench system. | Wolfgang E. Nagel, Markus A. Linn |
| 1991 | Software and hardware parallelism on the iWarp multi-computer. | Herbert G. Mayer, Brent Baxter |
| 1991 | Parallel program behavioral study on a shared-memory multiprocessor. | Hock-Beng Lim, Pen-Chung Yew |
| 1991 | Combining hardware and software cache coherence strategies. | David J. Lilja, Pen-Chung Yew |
| 1991 | Compiler algorithms for event variable synchronization. | Zhiyuan Li |
| 1991 | Analysis of scalability of parallel algorithms and architectures: a survey. | Vipin Kumar, Anshul Gupta |
| 1991 | Loop partitioning for distributed memory multiprocessors as unimodular transformations. | Dattatraya Kulkarni, Kamlesh G. Kumar, Anupam Basu, Arogyaswami Paulraj |
| 1991 | A space-efficient parallel garbage compaction algorithm. | Wolfgang Kchlin |
| 1991 | GRACIA: a software environment for graphical specification, automatic configuration and animation of parallel programs. | Ottmar Krmer-Fuhrmann, T. Brandes |
| 1991 | Programming data parallel algorithms on distributed memory using Kali. | Charles Koelbel, Piyush Mehrotra |
| 1991 | A scheme to extract run-time parallelism form sequential loops. | Chinhyun Kim, Jean-Luc Gaudiot |
| 1991 | Analysis and transformation in the ParaScope editor. | Ken Kennedy, Kathryn S. McKinley, Chau-Wen Tseng |
| 1991 | A performance bound analysis of multistage combining networks using a probabilistic model. | Byung-Chang Kang, Gyungho Lee, Richard Y. Kain |
| 1991 | Parallel computer ADENART - its architecture and application. | Hiroshi Kadota, Katsuyuki Kaneko, Ichiro Okabayashi, Tadashi Okamoto, T. Mimura, Yasuhiro Nakakura, Akiyoshi Wakatani, Masaitsu Nakajima, Junji Nishikawa, K. Zaiki, Tatsuo Nogi |
| 1991 | Semantical interprocedural parallelization: an overview of the PIPS project. | Franois Irigoin, Pierre Jouvelot, Rmi Triolet |
| 1991 | Beyond loop partitioning: data assignment and overlap to reduce communication overhead. | David E. Hudak, Santosh G. Abraham |