| 2025 | Fast on-demand Memory Mapping for Shared Memory and Disaggregated Systems. | Yuang Yan, Ryan E. Grant |
| 2025 | Applying Surrogate Modeling to Decouple Data Collection and Analysis from Simulation for Accelerated In-Situ Analysis. | Kewei Yan, Yonghong Yan |
| 2025 | Graphago: Accelerating SSD-based Graph Processing via Activity-Aware Graph Preprocessing. | Xianghao Xu, Yucheng Zhang, Gongxuan Zhang, Yongli Cheng, Fang Wang |
| 2025 | Kilometer-Scale AI-Powered and Performance-Portable Earth System Model (AP3ESM) to Achieve Year-Scale Simulation Speed on Heterogeneous Supercomputers. | Kai Xu, Maoxue Yu, Yuhu Chen, Jie Gao, Shuang Wang, Jiaying Song, Xiaohui Duan, Junwei Wei, Jiangfeng Yu, Hailong Liu, Jinrong Jiang, Yi Zhang, Pengfei Lin, Tianyi Wang, Pengfei Wang, Weipeng Zheng, Jingwei Xie, Jiakang Zhang, Zilu Liu, Xiaoyu Jin, Jilin Wei, Qixin Chang, Qingxia Lin, Yanzhi Zhou, Weiguo Liu, Wei Xue, Yiwen Li, Haohuan Fu, Yue Yu, Xuebin Chi, Lixin Wu |
| 2025 | Optimizing Quantum Circuit Mapping to Reduce Inter-Module Communications in Distributed Architectures. | Longshan Xu, Edwin Hsing-Mean Sha, Xiulin Cui, Qingfeng Zhuge |
| 2025 | Effective Node-Level Anomaly Detection in HPC Systems via Coarse-Grained Clustering and Fine-Grained Model Sharing. | Sibo Xia, Yongqian Sun, Xijie Pan, Yuan Yuan, Shenglin Zhang, Shaoyu Hu, Lei Tao, Yuqi Li, Jinghua Feng |
| 2025 | Reproducibility Report for SC25 Paper CPU- and GPU-initiated Communication Strategies for Conjugate Gradient Methods on Large GPU Clusters. | Brian J. N. Wylie |
| 2025 | "Offloading" Undergraduate Research to the Graphics Processing Unit for Acceleration. | Bryant M. Wyatt, Mason Bane |
| 2025 | TurboFNO: High-Performance Fourier Neural Operator with Fused FFT-GEMM-iFFT on GPU. | Shixun Wu, Yujia Zhai, Huangliang Dai, Yue Zhu, Haiyang Hu, Zizhong Chen |
| 2025 | Boosting Scientific Error-Bounded Lossy Compression through Optimized Synergistic Lossy-Lossless Orchestration. | Shixun Wu, Jinwen Pan, Jinyang Liu, Jiannan Tian, Ziwei Qiu, Jiajun Huang, Kai Zhao, Xin Liang, Sheng Di, Zizhong Chen, Franck Cappello |
| 2025 | ACTINA: Adapting Circuit-Switching Techniques for AI Networking Architectures. | Zhenguo Wu, Benjamin Klenk, Larry Dennison, Keren Bergman |
| 2025 | Network Replay and Consistency Across Testbeds. | Alexander Wolosewicz, Vinod Yegneswaran, Ashish Gehani, Nik Sultana |
| 2025 | Simulating many-engine spacecraft: Exceeding 1 quadrillion degrees of freedom via information geometric regularization. | Benjamin Wilfong, Anand Radhakrishnan, Henry Le Berre, Daniel Vickers, Tanush Prathi, Nikolaos Tselepidis, Benedikt Dorschner, Reuben D. Budiardja, Brian Cornille, Stephen Abbott, Florian Schfer, Spencer H. Bryngelson |
| 2025 | Testing and Benchmarking Emerging Supercomputers via the MFC Flow Solver. | Benjamin Wilfong, Anand Radhakrishnan, Henry Le Berre, Tanush Prathi, Stephen Abbott, Spencer H. Bryngelson |
| 2025 | Reproducibility Report for SC25 Paper Demystifying the Resilience of Large Language Model Inference: An End-to-End Perspective. | Sandra Wienke |
| 2025 | Large-Message All-to-All Communication at Frontier Scale. | James Buford White |
| 2025 | Towards a user-centric HPC-QC environment. | Aleksander Wennersteen, Matthieu Moreau, Aurelien Nober, Mourad Beji |
| 2025 | Rapid Quantum Network Simulation Design with a Path to Scalable Execution. | Aaron Welch, Joel A. Dawson, Mariam Kiran |
| 2025 | HPC-R1: Characterizing R1-like Large Reasoning Models on HPC. | Adam Weingram, Duo Zhang, Zhonghao Chen, Hao Qi, Xiaoyi Lu |
| 2025 | Reproducibility Report for SC25 Paper RedSan: A Redundant Memory Instruction Sanitizer for GPU Programs. | Volker Weinberg |
| 2025 | An Optimized Generalized Multi-Color Point Implicit Solver for Intel GPUs using OneAPI ESIMD. | Joseph Wassell, Mohammad Zubair, Aaron Walden, Gabriel Nastac, Eric J. Nielsen, Timothe Ewart |
| 2025 | Towards Efficient LLM Inference via Collective and Adaptive Speculative Decoding. | Siqi Wang, Hailong Yang, Xuezhu Wang, Tongxuan Liu, Pengbo Wang, Yufan Xu, Xuning Liang, Kejie Ma, Tianyu Feng, Xin You, Ruihao Gong, Rui Wang, Zhongzhi Luan, Yi Liu, Depei Qian |
| 2025 | MXBLAS: Accelerating 8-bit Deep Learning with a Unified Micro-Scaled GEMM Library. | Weihu Wang, Yaqi Xia, Donglin Yang, Xiaobo Zhou, Dazhao Cheng |
| 2025 | WAGES: Workload-Aware GPU Sharing System for Energy-Efficient Serverless LLM Serving. | Tianyu Wang, Gourav Rattihalli, Aditya Dhakal, Xulong Tang, Dejan S. Milojicic |
| 2025 | LLM4FP: LLM-Based Program Generation for Triggering Floating-Point Inconsistencies Across Compilers. | Yutong Wang, Cindy Rubio-Gonzlez |