| 2022 | Relative Performance Projection on Arm Architectures. | Clment Gavoille, Hugo Taboada, Patrick Carribault, Fabrice Dupros, Brice Goglin, Emmanuel Jeannot |
| 2022 | Implementation and Performance Evaluation of Memory System Using Addressable Cache for HPC Applications on HBM2 Equipped FPGAs. | Norihisa Fujita, Ryohei Kobayashi, Yoshiki Yamaguchi, Taisuke Boku |
| 2022 | Accelerating Parallel Operation for Compacting Selected Elements on GPUs. | Johannes Fett, Urs Kober, Christian Schwarz, Dirk Habich, Wolfgang Lehner |
| 2022 | Programming Heterogeneous Architectures Using Hierarchical Tasks. | Mathieu Faverge, Nathalie Furmento, Abdou Guermouche, Gwenol Lucas, Raymond Namyst, Samuel Thibault, Pierre-Andr Wacrenier |
| 2022 | Performance Portability Assessment: Non-negative Matrix Factorization as a Case Study. | Youssef Faqir-Rhazoui, Carlos Garca, Francisco Tirado |
| 2022 | Two-Agent Scheduling with Resource Augmentation on Multiple Machines. | Vincent Fagnon, Giorgio Lucarelli, Clment Mommessin, Denis Trystram |
| 2022 | mCAP: Memory-Centric Partitioning for Large-Scale Pipeline-Parallel DNN Training. | Henk Dreuning, Henri E. Bal, Rob V. van Nieuwpoort |
| 2022 | Exploring the Suitability of the Cerebras Wafer Scale Engine for Stencil-Based Computation Codes. | Nick Brown, Brandon Echols, Justs Zarins, Tobias Grosser |
| 2022 | Rapid Development of OS Support with PMCSched for Scheduling on Asymmetric Multicore Systems. | Carlos Bilbao, Juan Carlos Saez, Manuel Prieto-Matas |
| 2022 | Accurate Fork-Join Profiling on the Java Virtual Machine. | Matteo Basso, Eduardo Rosales, Filippo Schiavio, Andrea Ros, Walter Binder |
| 2022 | Preliminary Study of Resource Allocation in Wireless Communications. | Peace Ayegba, Sofiat Olaosebikan |
| 2022 | A Bi-Criteria FPTAS for Scheduling with Memory Constraints on Graphs with Bounded Tree-Width. | Eric Angel, Sbastien Morais, Damien Regnault |
| 2021 | Optimized Implementation of the HPCG Benchmark on Reconfigurable Hardware. | Alberto Zeni, Kenneth O'Brien, Michaela Blott, Marco D. Santambrogio |
| 2021 | A Low Overhead Tasking Model for OpenMP. | Chenle Yu, Sara Royuela, Eduardo Quiones |
| 2021 | Efficient and Systematic Partitioning of Large and Deep Neural Networks for Parallelization. | Haoran Wang, Chong Li, Thibaut Tachon, Hongxing Wang, Sheng Yang, Sbastien Limet, Sophie Robert |
| 2021 | Accelerating FFT Using NEC SX-Aurora Vector Engine. | Pablo Vizcaino, Filippo Mantovani, Jess Labarta |
| 2021 | Exploring the Impact of Node Failures on the Resource Allocation for Parallel Jobs. | Ioannis Vardas, Manolis Ploumidis, Manolis Marazakis |
| 2021 | OpenMP Target Task: Tasking and Target Offloading on Heterogeneous Systems. | Pedro Valero-Lara, Jungwon Kim, Oscar R. Hernandez, Jeffrey S. Vetter |
| 2021 | Porting Sparse Linear Algebra to Intel GPUs. | Yuhsiang M. Tsai, Terry Cojean, Hartwig Anzt |
| 2021 | Communication Overlapping Pipelined Conjugate Gradients for Distributed Memory Systems and Heterogeneous Architectures. | Manasi Tiwari, Sathish Vadhiyar |
| 2021 | A Fixed-Parameter Algorithm for Scheduling Unit Dependent Tasks with Unit Communication Delays. | Ning Tang, Alix Munier Kordon |
| 2021 | Interferences Between Communications and Computations in Distributed HPC Systems. | Philippe Swartvagher |
| 2021 | Kernel Fusion in OpenCL. | John A. Stratton, Jyothi Krishna V. S, Jeevitha Palanisamy, Karthikadevi Chinnaraju |
| 2021 | Monitoring Collective Communication Among GPUs. | Muhammet Abdullah Soytrk, Palwisha Akhtar, Erhan Tezcan, Didem Unat |
| 2021 | GPU-Accelerated Mahalanobis-Average Hierarchical Clustering Analysis. | Adam Smelko, Miroslav Kratochvl, Martin Krulis, Toms Sieger |