Skip to content

Azzam Haidar

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

28

Venues

9

Active years

2011–2023

Best venue rank

A

Where they publish

Papers

28 indexed papers, newest first.

YearVenueTitleAuthors
2023QCEcuQuantum SDK: A High-Performance Library for Accelerating Quantum Science.Harun Bayraktar, Ali Charara, David Clark, Saul Cohen, Timothy B. Costa, Yao-Lung L. Fang, Yang Gao, Jack Guan, John A. Gunnels, Azzam Haidar, Andreas Hehn, Markus Hhnerbach, Matthew Jones, Tom Lubowe, Dmitry I. Lyakh, Shinya Morino, Paul Springer, Sam Stanwyck, Igor Terentyev, Satya Varadhan, Jonathan Wong, Takuma Yamaguchi
2020ICCSheFFTe: Highly Efficient FFT for Exascale.Alan Ayala, Stanimire Tomov, Azzam Haidar, Jack J. Dongarra
2018ICCSThe Design of Fast and Energy-Efficient Linear Solvers: On the Potential of Half-Precision Arithmetic and Iterative Refinement Techniques.Azzam Haidar, Ahmad Abdelfattah, Mawussi Zounon, Panruo Wu, Srikara Pranesh, Stanimire Tomov, Jack J. Dongarra
2018SCHarnessing GPU tensor cores for fast FP16 arithmetic to speed up mixed-precision iterative refinement solvers.Azzam Haidar, Stanimire Tomov, Jack J. Dongarra, Nicholas J. Higham
2017ICCSFactorization and Inversion of a Million Matrices using GPUs: Challenges and Countermeasures.Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2017ICCSOptimizing the SVD Bidiagonalization Process for a Batch of Small Matrices.Tingxing Dong, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2017ICSNovel HPC techniques to batch execution of many variable size BLAS computations on GPUs.Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2017PPoPPHigh-performance Cholesky factorization for GPU-only execution.Azzam Haidar, Ahmad Abdelfattah, Stanimire Tomov, Jack J. Dongarra
2017SCInvestigating half precision arithmetic to accelerate dense linear system solvers.Azzam Haidar, Panruo Wu, Stanimire Tomov, Jack J. Dongarra
2016EuroParHigh-Performance Matrix-Matrix Multiplications of Very Small Matrices.Ian Masliah, Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Marc Baboulin, Jol Falcou, Jack J. Dongarra
2016ICCSHigh-Performance Tensor Contractions for GPUs.Ahmad Abdelfattah, Marc Baboulin, Veselin Dobrev, Jack J. Dongarra, Christopher W. Earl, Joel Falcou, Azzam Haidar, Ian Karlin, Tzanio V. Kolev, Ian Masliah, Stanimire Tomov
2016ICCSPerformance Tuning and Optimization Techniques of Fixed and Variable Size Batched Cholesky Factorization on GPUs.Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2016SCTowards Achieving Performance Portability Using Directives for Accelerators.M. Graham Lopez, Vernica G. Vergara Larrea, Wayne Joubert, Oscar R. Hernandez, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2015HPCCFlexible Linear Algebra Development and Scheduling with Cholesky Factorization.Azzam Haidar, Asim YarKhan, Chongxiao Cao, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra
2015ICCSPerformance Analysis and Optimisation of Two-sided Factorization Algorithms for Heterogeneous Platform.Khairul Kabir, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2015PPoPPTowards batched linear solvers on accelerated hardware platforms.Azzam Haidar, Tingxing Dong, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra
2015PPoPPOptimization for performance and energy for batched matrix computations on GPUs.Azzam Haidar, Tingxing Dong, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra
2015SCWeighted dynamic scheduling with many parallelism grains for offloading of numerical workloads to multiple varied accelerators.Azzam Haidar, Yulu Jia, Piotr Luszczek, Stanimire Tomov, Asim YarKhan, Jack J. Dongarra
2015SCEfficient implementation of quantum materials simulations on distributed CPU-GPU systems.Raffaele Solc, Anton Kozhevnikov, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra, Thomas C. Schulthess
2014HPCCLU Factorization of Small Matrices: Accelerating Batched DGETRF on the GPU.Tingxing Dong, Azzam Haidar, Piotr Luszczek, James Austin Harris, Stanimire Tomov, Jack J. Dongarra
2014ICPPA Fast Batched Cholesky Factorization on a GPU.Tingxing Dong, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra
2014SCPerformance and portability with OpenCL for throughput-oriented HPC workloads across accelerators, coprocessors, and multicore processors.Chongxiao Cao, Mark Gates, Azzam Haidar, Piotr Luszczek, Stanimire Tomov, Ichitaro Yamazaki, Jack J. Dongarra
2013ICSToward a scalable multi-GPU eigensolver via compute-intensive kernels and efficient communication.Azzam Haidar, Mark Gates, Stanimire Tomov, Jack J. Dongarra
2013PPAMPortable HPC Programming on Intel Many-Integrated-Core Hardware with MAGMA Port to Xeon Phi.Jack J. Dongarra, Mark Gates, Azzam Haidar, Yulu Jia, Khairul Kabir, Piotr Luszczek, Stanimire Tomov
2013SCAn improved parallel singular value algorithm and its implementation for multicore hardware.Azzam Haidar, Jakub Kurzak, Piotr Luszczek
2012SCAbstract: A Novel Hybrid CPU-GPU Generalized Eigensolver for Electronic Structure Calculations Based on Fine Grained Memory Aware Tasks.Raffaele Solc, Azzam Haidar, Stanimire Tomov, Thomas C. Schulthess, Jack J. Dongarra
2012SCPoster: A Novel Hybrid CPU-GPU Generalized Eigensolver for Electronic Structure Calculations Based on Fine Grained Memory Aware Tasks.Raffaele Solc, Azzam Haidar, Stanimire Tomov, Thomas C. Schulthess, Jack J. Dongarra
2011SCParallel reduction to condensed forms for symmetric eigenvalue problems using aggregated fine-grained and memory-aware kernels.Azzam Haidar, Hatem Ltaief, Jack J. Dongarra