Skip to content

Samuel Williams

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

47

Venues

15

Active years

2007–2025

Best venue rank

A*

Where they publish

Papers

47 indexed papers, newest first.

YearVenueTitleAuthors
2025CAVAutomatic Synthesis of Smooth Infinite Horizon Paths Satisfying Linear Temporal Logic Specifications.Samuel Williams, Jyotirmoy Deshmukh
2025SCRoofline Analysis of Tightly-Coupled CPU-GPU Superchips: A Study on MI300A and GH200.Oscar Antepara, Leonid Oliker, Samuel Williams
2025SCBenchmark-driven Models for Energy Analysis and Attribution of GPU-Accelerated Supercomputing.Oscar Antepara, Zhengji Zhao, Brian Austin, Nan Ding, Leonid Oliker, Nicholas J. Wright, Samuel Williams
2024HCIA Relative Pitch Based Approach to Non-verbal Vocal Interaction as a Continuous and One-Dimensional Controller.Samuel Williams, Denis Gracanin
2024ICPPBrickDL: Graph-Level Optimizations for DNNs with Fine-Grained Data Blocking on GPUs.Mahesh Lakshminarasimhan, Mary W. Hall, Samuel Williams, Oscar Antepara
2024SCA Workflow Roofline Model for End-to-End Workflow Performance Analysis.Nan Ding, Brian Austin, Yang Liu, Neil Mehta, Steven Farrell, Johannes P. Blaschke, Leonid Oliker, Hai Ah Nam, Nicholas J. Wright, Samuel Williams
2024SCPerformance Portable Optimizations of an Ice-sheet Modeling Code on GPU-supercomputers.Oscar Antepara, Samuel Williams, Max Carlson, Jerry Watkins
2024SCHigh-Performance, Scalable Geometric Multigrid via Fine-Grain Data Blocking for GPUs.Oscar Antepara, Samuel Williams, Hans Johansen, Mary W. Hall
2024SCSystem-Wide Roofline Profiling -a Case Study on NERSC's Perlmutter Supercomputer.Brian Austin, Dhruva Kulkarni, Brandon Cook, Samuel Williams, Nicholas J. Wright
2024SCComprehensive Performance Modeling and System Design Insights for Foundation Models.Shashank Subramanian, Ermal Rrapaj, Peter Harrington, Smeet Chheda, Steven Farrell, Brian Austin, Samuel Williams, Nicholas J. Wright, Wahid Bhimji
2024VRAn Approach to Pitch Based Implementation of Non-verbal Vocal Interaction (NVVI).Samuel Williams, Denis Gracanin
2023SCEvaluating the Performance of One-sided Communication on CPUs and GPUs.Nan Ding, Muhammad Haseeb, Taylor L. Groves, Samuel Williams
2023SCUnified Communication Optimization Strategies for Sparse Triangular Solver on CPU and GPU Clusters.Yang Liu, Nan Ding, Piyush Sao, Samuel Williams, Xiaoye Sherry Li
2023SCPerformance Portability Evaluation of Blocked Stencil Computations on GPUs.Oscar Antepara, Samuel Williams, Hans Johansen, Tuowen Zhao, Samantha Hirsch, Priya Goyal, Mary W. Hall
2023SCPerformance-Portable GPU Acceleration of the EFIT Tokamak Plasma Equilibrium Reconstruction Code.Oscar Antepara, Samuel Williams, Scott Kruger, Torrin Bechtel, Joseph McClenaghan, Lang Lao
2021HCIRedefining the Digital Paradigm for Virtual Museums - Towards Interactive and Engaging Experiences in the Post-pandemic Era.Archi Dasgupta, Samuel Williams, Gunnar Nelson, Mark Manuel, Shaoli Dasgupta, Denis Gracanin
2021iLRNImmersive Technology in the Public School Classroom: When a Class Meets.Samuel Williams, Rowena Enatsky, Holly Gillcash, James J. Murphy, Denis Gracanin
2021PPoPPImproving communication by optimizing on-node data movement with data layout.Tuowen Zhao, Mary W. Hall, Hans Johansen, Samuel Williams
2020ICCADA CAD-based methodology to optimize HLS code via the Roofline model.Marco Siracusa, Marco Rabozzi, Emanuele Del Sozzo, Lorenzo Di Tucci, Samuel Williams, Marco D. Santambrogio
2020SCTime-Based Roofline for Deep Learning Performance Analysis.Yunsong Wang, Charlene Yang, Steven Farrell, Yan Zhang, Thorsten Kurth, Samuel Williams
2019SCAn Instruction Roofline Model for GPUs.Nan Ding, Samuel Williams
2019SCExploiting reuse and vectorization in blocked stencil computations on CPUs and GPUs.Tuowen Zhao, Protonu Basu, Samuel Williams, Mary W. Hall, Hans Johansen
2018PPoPPSIMD code generation for stencils on brick decompositions.Tuowen Zhao, Mary W. Hall, Protonu Basu, Samuel Williams, Hans Johansen
2018SCImproving MPI Reduction Performance for Manycore Architectures with OpenMP and Data Compression.Hongzhang Shan, Samuel Williams, Calvin W. Johnson
2017ICSEngDeveloping an Air Hockey Game in LabVIEW.Matthew Bara, Sneha Gollamudi, Samuel Williams
2017OOPSLAPerformance analysis and optimization of the RAMPAGE metal alloy potential generation software.Philip C. Roth, Hongzhang Shan, David Riegner, Nikolas Antolin, Sarat Sreepathi, Leonid Oliker, Samuel Williams, Shirley Moore, Wolfgang Windl
2016SCEvaluating and Optimizing the NERSC Workload on Knights Landing.Taylor Barnes, Brandon Cook, Jack Deslippe, Douglas Doerfler, Brian Friesen, Yun (Helen) He, Thorsten Kurth, Tuomas Koskela, Mathieu Lobet, Tareq M. Malas, Leonid Oliker, Andrey Ovsyannikov, Abhinav Sarje, Jean-Luc Vay, Henri Vincenti, Samuel Williams, Pierre Carrier, Nathan Wichmann, Marcus Wagner, Paul R. C. Kent, Christopher Kerr, John M. Dennis
2016SCExperiences of Applying One-Sided Communication to Nearest-Neighbor Communication.Hongzhang Shan, Samuel Williams, Yili Zheng, Weiqun Zhang, Bei Wang, Stphane Ethier, Zhengji Zhao
2016SCExtreme scale plasma turbulence simulations on top supercomputers worldwide.William M. Tang, Bei Wang, Stphane Ethier, Grzegorz Kwasniewski, Torsten Hoefler, Khaled Z. Ibrahim, Kamesh Madduri, Samuel Williams, Leonid Oliker, Carlos Rosales-Fernandez, Timothy J. Williams
2015ICCSParallel Performance Optimizations on Unstructured Mesh-based Simulations.Abhinav Sarje, Sukhyun Song, Douglas Jacobsen, Kevin A. Huck, Jeffrey K. Hollingsworth, Allen D. Malony, Samuel Williams, Leonid Oliker
2015PPAMComparative Performance Analysis of Coarse Solvers for Algebraic Multigrid on Multicore and Manycore Architectures.Alex Druinsky, Pieter Ghysels, Xiaoye S. Li, Osni Marques, Samuel Williams, Andrew T. Barker, Delyan Kalchev, Panayot S. Vassilevski
2015PPoPPExploiting communication concurrency on high performance computing systems.Nicholas Chaimov, Khaled Z. Ibrahim, Samuel Williams, Costin Iancu
2015PPoPPThread-level parallelization and optimization of NWChem for the Intel MIC architecture.Hongzhang Shan, Samuel Williams, Wibe de Jong, Leonid Oliker
2015SCParallel implementation and performance optimization of the configuration-interaction method.Hongzhang Shan, Samuel Williams, Calvin W. Johnson, Kenneth S. McElvain, W. Erich Ormand
2014ICSCollective memory transfers for multi-core chips.George Michelogiannakis, Alexander Williams, Samuel Williams, John Shalf
2014SCRoofline Model Toolkit: A Practical Tool for Architectural and Program Analysis.Yu Jung Lo, Samuel Williams, Brian van Straalen, Terry J. Ligocki, Matthew J. Cordery, Nicholas J. Wright, Mary W. Hall, Leonid Oliker
2013HiPCCompiler generation and autotuning of communication-avoiding operators for geometric multigrid.Protonu Basu, Anand Venkat, Mary W. Hall, Samuel Williams, Brian van Straalen, Leonid Oliker
2013IVCNZTransform flow: A mobile augmented reality visualisation and evaluation toolkit.Samuel Williams, Richard D. Green, Mark Billinghurst
2013SCKinetic turbulence simulations at extreme scale on leadership-class systems.Bei Wang, Stphane Ethier, William M. Tang, Timothy J. Williams, Khaled Z. Ibrahim, Kamesh Madduri, Samuel Williams, Leonid Oliker
2012SCOptimization of geometric multigrid for emerging multi- and manycore processors.Samuel Williams, Dhiraj D. Kalamkar, Amik Singh, Anand M. Deshpande, Brian van Straalen, Mikhail Smelyanskiy, Ann S. Almgren, Pradeep Dubey, John Shalf, Leonid Oliker
2011SCHardware/software co-design for energy-efficient seismic modeling.Jens Krueger, David Donofrio, John Shalf, Marghoob Mohiyuddin, Samuel Williams, Leonid Oliker, Franz-Josef Pfreundt
2011SCGyrokinetic toroidal simulations on leading multi- and manycore HPC systems.Kamesh Madduri, Khaled Z. Ibrahim, Samuel Williams, Eun-Jin Im, Stphane Ethier, John Shalf, Leonid Oliker
2011SCExtracting ultra-scale Lattice Boltzmann performance via hierarchical and distributed auto-tuning.Samuel Williams, Leonid Oliker, Jonathan Carter, John Shalf
2009SCMemory-efficient optimization of Gyrokinetic particle-to-grid interpolation for multicore processors.Kamesh Madduri, Samuel Williams, Stphane Ethier, Leonid Oliker, John Shalf, Erich Strohmaier, Katherine A. Yelick
2009SCA design methodology for domain-optimized power-efficient supercomputing.Marghoob Mohiyuddin, Mark Murphy, Leonid Oliker, John Shalf, John Wawrzynek, Samuel Williams
2008SCStencil computation optimization and auto-tuning on state-of-the-art multicore architectures.Kaushik Datta, Mark Murphy, Vasily Volkov, Samuel Williams, Jonathan Carter, Leonid Oliker, David A. Patterson, John Shalf, Katherine A. Yelick
2007SCOptimization of sparse matrix-vector multiplication on emerging multicore platforms.Samuel Williams, Leonid Oliker, Richard W. Vuduc, John Shalf, Katherine A. Yelick, James Demmel