| 2024 | Understanding Mixed Precision GEMM with MPGemmFI: Insights into Fault Resilience. | Bo Fang, Xinyi Li, Harvey Dam, Cheng Tan, Siva Kumar Sastry Hari, Timothy Tsai, Ignacio Laguna, Dingwen Tao, Ganesh Gopalakrishnan, Prashant J. Nair, Kevin J. Barker, Ang Li |
| 2024 | Welcome Message from the IEEE Cluster 2024 Workshop chair. | Yohei Miki |
| 2023 | FullRepair: Towards Optimal Repair Pipelining in Erasure-Coded Clustered Storage Systems. | Yuzuo Zhang, Xinyuan Tu, Lin Wang, Yuchong Hu, Fang Wang, Ye Wang |
| 2023 | DoW-KV: A DPU-offloaded and Write-optimized Key-Value Store on Disaggregated Persistent Memory. | Yiwen Zhang, Guokuan Li, Jiguang Wan, Junyue Wang, Jun Li, Ting Yao, Huatao Wu, Daohui Wang |
| 2023 | A Finite-Difference Time-Domain (FDTD) solver with linearly scalable performance in an FPGA cluster. | Zhenyu Xu, Miaoxiang Yu, Jillian Cai, Qing Yang, Tao Wei |
| 2023 | Generalized Collective Algorithms for the Exascale Era. | Michael Wilkins, Hanming Wang, Peizhi Liu, Bangyen Pham, Yanfei Guo, Rajeev Thakur, Peter A. Dinda, Nikos Hardavellas |
| 2023 | Exact Distributed Stochastic Block Partitioning. | Frank Wanye, Vitaliy Gleyzer, Edward K. Kao, Wu-Chun Feng |
| 2023 | On the Multi-Dimensional Acceleration of Stochastic Blockmodeling for Community Detection. | Frank Wanye, Wu-Chun Feng |
| 2023 | SciLance: Mitigate Load Imbalance for Parallel Scientific Applications in Cloud Environments. | Xinying Wang, Lipeng Wan, Scott Klasky, Dongfang Zhao, Feng Yan |
| 2023 | Prophet: Fine-grained Load Balancing for Parallel Training of Large-scale MoE Models. | Wei Wang, Zhiquan Lai, Shengwei Li, Weijie Liu, Keshi Ge, Yujie Liu, Ao Shen, Dongsheng Li |
| 2023 | OpenMP Offloading to DPU. | Muhammad Usman, Sergio Iserte, Roger Ferrer, Antonio J. Pea |
| 2023 | Accelerating Distributed ML Training via Selective Synchronization (Poster Abstract). | Sahil Tyagi, Martin Swany |
| 2023 | Accelerating Distributed ML Training via Selective Synchronization. | Sahil Tyagi, Martin Swany |
| 2023 | Uniform Algorithms for Reduce-scatter and (most) other Collectives for MPI. | Jesper Larsson Trff, Sascha Hunold, Ioannis Vardas, Nikolaus Manes Funk |
| 2023 | A Dynamic Network-Native MPI Partitioned Aggregation Over InfiniBand Verbs. | Yiltan Hassan Temuin, Scott Levy, Whit Schonbein, Ryan E. Grant, Ahmad Afsahi |
| 2023 | TopoCommit: A Topological Commit Protocol for Cross-Ledger Transactions in Scientific Computing. | Olamide Timothy Tawose, Lei Yang, Dongfang Zhao |
| 2023 | I/O-Aware Flushing for HPC Caching Filesystem. | Osamu Tatebe, Kohei Hiraga, Hiroki Ohtsuji |
| 2023 | Efficient Particle Tracing for Scalable Kinetic Plasma Simulation Analysis. | Nigel Tan, Scott V. Luedtke, Michela Taufer, Brian J. Albright |
| 2023 | Communication-Avoiding Recursive Aggregation. | Yihao Sun, Sidharth Kumar, Thomas Gilray, Kristopher K. Micinski |
| 2023 | A Lightweight Network Traffic Prediction Method for SmartNICs. | Whit Schonbein, Tinotenda Matsika, Ryan E. Grant |
| 2023 | Hierarchical Resource Partitioning on Modern GPUs: A Reinforcement Learning Approach. | Urvij Saroliya, Eishi Arima, Dai Liu, Martin Schulz |
| 2023 | Mappings and patterns to improve the triangular matrix product on distributed systems. | Inmaculada Santamaria-Valenzuela, Roco Carratal-Sez, Yuri Torres, Diego R. Llanos, Arturo Gonzlez-Escribano |
| 2023 | Performance improvement by enhancing spatial parallelism on FPGA for HPC applications. | Yuka Sano, Taisuke Boku, Mitsuhisa Sato, Miwako Tsuji, Norihisa Fujita, Ryohei Kobayashi |
| 2023 | ProvLight: Efficient Workflow Provenance Capture on the Edge-to-Cloud Continuum. | Daniel Rosendo, Marta Mattoso, Alexandru Costan, Renan Souza, Dbora B. Pina, Patrick Valduriez, Gabriel Antoniu |
| 2023 | Latency and Bandwidth Microbenchmarks of Six US Department of Energy Systems in the Top500. | Carl Pearson, Christopher M. Siefert, Stephen L. Olivier, Andrey Prokopenko, Timothy J. Fuller, Jonathan J. Hu |