| 2025 | Reproducibility Report for SC25 Paper XaaS Containers: Performance-Portable Representation With Source and IR Containers. | Joao Vicente Ferreira Lima |
| 2025 | SlimPipe: Memory-Thrifty and Efficient Pipeline Parallelism for Long-Context LLM Training. | Zhouyang Li, Yuliang Liu, Wei Zhang, Tailing Yuan, Bin Chen, Chengru Song |
| 2025 | Parallel Data Object Creation: Scalable Metadata Management in Parallel I/O Library. | Youjia Li, Robert Latham, Robert B. Ross, Ankit Agrawal, Alok N. Choudhary, Wei-keng Liao |
| 2025 | HP-MDR: High-performance and Portable Data Refactoring and Progressive Retrieval with Advanced GPUs. | Yanliang Li, Wenbo Li, Qian Gong, Qing Liu, Norbert Podhorszki, Scott Klasky, Xin Liang, Jieyang Chen |
| 2025 | Diff-MoE: Efficient Batched MoE Inference with Priority-Driven Differential Expert Caching. | Kexin Li, Wenkan Huang, Qinggang Wang, Long Zheng, Xiaofei Liao, Hai Jin, Jingling Xue |
| 2025 | MaverIQ: Fingerprint-Guided Extrapolation and Fragmentation-Aware Layering for Intent-Based LLM Serving. | Dimitrios Liakopoulos, Prasoon Sinha, Tianrui Hu, Myungjin Lee, Neeraja J. Yadwadkar |
| 2025 | Reproducibility Report for SC25 Paper Sparsified Preconditioned Conjugate Gradient Solver on GPUs. | Sixu Li |
| 2025 | SparStencil: Retargeting Sparse Tensor Cores to Scientific Stencil Computations via Structured Sparsity Transformation. | Qi Li, Kun Li, Haozhi Han, Liang Yuan, Yunquan Zhang, Yifeng Chen, Junshi Chen, Hong An, Ting Cao, Mao Yang |
| 2025 | Fine-grained Automated Failure Management for Extreme-Scale GPU Accelerated Systems. | Yonatan Levitt, Richard Barella, Sam Zeltner, Thomas Musta, Lance Cheney, Gustavo Espinosa, Olivier Franza, Balazs Gerofi |
| 2025 | Breaking the System Noise Barrier at Exascale. | Edgar A. Len, Joseph Glenski, Mark J. Stock, Kim H. McMahon, William Loewe, Clark Snyder, Larry Kaplan, Srinath Vadlamani, Timothy I. Mattox, Trent D'Hooge, Brian Behlendorf, Nathan Hanford, Ramesh Pankajakshan, Matthew L. Leininger |
| 2025 | A GPU FFT Wrapper to Co-optimize Floating-Point Precision and Library Selection via Predictive Error Modeling. | Julius Lehner, Eishi Arima, Martin Schulz |
| 2025 | SlimIO: Lightweight I/O Path Design for Write Isolation in FDP-backed In-Memory Databases. | Sangyun Lee, Sungjin Byeon, Soon Hwang, Jaewan Park, Jooyoung Hwang, Junyoung Han, Javier Gonzlez, Awais Khan, Youngjae Kim |
| 2025 | Assessing a RISC-V Accelerator for Cross-Section Lookup in Chipyard. | Andrew Ledbetter, Kazutomo Yoshii, John R. Tramm |
| 2025 | Fast Linear Solvers via AI-Tuned Markov Chain Monte Carlo-based Matrix Inversion. | Anton Lebedev, Won Kyung Lee, Soumyadip Ghosh, Olha Ivanyshyn Yaman, Vassilis Kalantzis, Yingdong Lu, Tomasz Nowicki, Shashanka Ubaru, Lior Horesh, Vassil Alexandrov |
| 2025 | Reproducibility Report for SC25 Paper Bine Trees: Enhancing Collective Operations by Optimizing Communication Locality. | Jan Laukemann |
| 2025 | Reproducibility Report for SC25 Paper Addressing Reproducibility Challenges in HPC with Continuous Integration. | Ruben Laso |
| 2025 | Reproducibility Report for SC25 Paper RAPTOR: Practical Numerical Profiling of Scientific Applications. | Ruben Laso |
| 2025 | Modelling Load Imbalance In Shared Memory Multicore Systems. | Johannes Langguth, James D. Trotter, Xing Cai |
| 2025 | A RISC-V Vector Extension for Multi-word Arithmetic. | Yunhao Lan, Larry Tang, Naifeng Zhang, Youngjin Eum, James C. Hoe, Franz Franchetti |
| 2025 | RISC-V Vectorization Coverage for HPC: A TSVC-Based Analysis. | Hung-Ming Lai, Pei-Hung Lin, Maya B. Gokhale, Ivy Peng, Hiren D. Patel, Jenq-Kuen Lee |
| 2025 | Microscopic-Level Mouse Whole Cortex Simulation Composed of 9 Million Biophysical Neurons and 26 Billion Synapses on the Supercomputer Fugaku. | Rin Kuriyama, Kaaya Akira, Laura Green, Beatriz Herrera, Kael Dai, Mari Iura, Gilles Gouaillardet, Asako Terasawa, Taira Kobayashi, Jun Igarashi, Anton Arkhipov, Tadashi Yamazaki |
| 2025 | xGFabric: Coupling Sensor Networks and HPC Facilities with Private 5G Wireless Networks for Real-Time Digital Agriculture. | Liubov Kurafeeva, Alan Subedi, Ryan Hartung, Michael Fay, Avhishek Biswas, Shantenu Jha, Ozgur O. Kilic, Chandra Krintz, Andr Merzky, Douglas Thain, Mehmet C. Vuran, Rich Wolski |
| 2025 | TT-LoRA MoE: Using Parameter-Efficient Fine-Tuning and Sparse Mixture-Of-Experts. | Pradip Kunwar, Minh N. Vu, Maanak Gupta, Mahmoud Abdelsalam, Manish Bhattarai |
| 2025 | MPI Communication Performance on AMD MI300A: Microbenchmarks and Applications. | Goutham Kalikrishna Reddy Kuncham, Siyuan Zhang, Shoaib Mohammad, Chen-Chun Chen, Dhabaleswar K. Panda |
| 2025 | Teaching Task-Based Parallel Programming with a Runtime Systems-Aware Perspective. | Vivek Kumar |