Azzam Haidar
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
28
Venues
9
Active years
2011–2023
Best venue rank
A
Where they publish
Papers
28 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | QCE | cuQuantum SDK: A High-Performance Library for Accelerating Quantum Science. | Harun Bayraktar, Ali Charara, David Clark, Saul Cohen, Timothy B. Costa, Yao-Lung L. Fang, Yang Gao, Jack Guan, John A. Gunnels, Azzam Haidar, Andreas Hehn, Markus Hhnerbach, Matthew Jones, Tom Lubowe, Dmitry I. Lyakh, Shinya Morino, Paul Springer, Sam Stanwyck, Igor Terentyev, Satya Varadhan, Jonathan Wong, Takuma Yamaguchi |
| 2020 | ICCS | heFFTe: Highly Efficient FFT for Exascale. | Alan Ayala, Stanimire Tomov, Azzam Haidar, Jack J. Dongarra |
| 2018 | ICCS | The Design of Fast and Energy-Efficient Linear Solvers: On the Potential of Half-Precision Arithmetic and Iterative Refinement Techniques. | Azzam Haidar, Ahmad Abdelfattah, Mawussi Zounon, Panruo Wu, Srikara Pranesh, Stanimire Tomov, Jack J. Dongarra |
| 2018 | SC | Harnessing GPU tensor cores for fast FP16 arithmetic to speed up mixed-precision iterative refinement solvers. | Azzam Haidar, Stanimire Tomov, Jack J. Dongarra, Nicholas J. Higham |
| 2017 | ICCS | Factorization and Inversion of a Million Matrices using GPUs: Challenges and Countermeasures. | Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2017 | ICCS | Optimizing the SVD Bidiagonalization Process for a Batch of Small Matrices. | Tingxing Dong, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2017 | ICS | Novel HPC techniques to batch execution of many variable size BLAS computations on GPUs. | Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2017 | PPoPP | High-performance Cholesky factorization for GPU-only execution. | Azzam Haidar, Ahmad Abdelfattah, Stanimire Tomov, Jack J. Dongarra |
| 2017 | SC | Investigating half precision arithmetic to accelerate dense linear system solvers. | Azzam Haidar, Panruo Wu, Stanimire Tomov, Jack J. Dongarra |
| 2016 | EuroPar | High-Performance Matrix-Matrix Multiplications of Very Small Matrices. | Ian Masliah, Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Marc Baboulin, Jol Falcou, Jack J. Dongarra |
| 2016 | ICCS | High-Performance Tensor Contractions for GPUs. | Ahmad Abdelfattah, Marc Baboulin, Veselin Dobrev, Jack J. Dongarra, Christopher W. Earl, Joel Falcou, Azzam Haidar, Ian Karlin, Tzanio V. Kolev, Ian Masliah, Stanimire Tomov |
| 2016 | ICCS | Performance Tuning and Optimization Techniques of Fixed and Variable Size Batched Cholesky Factorization on GPUs. | Ahmad Abdelfattah, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2016 | SC | Towards Achieving Performance Portability Using Directives for Accelerators. | M. Graham Lopez, Vernica G. Vergara Larrea, Wayne Joubert, Oscar R. Hernandez, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2015 | HPCC | Flexible Linear Algebra Development and Scheduling with Cholesky Factorization. | Azzam Haidar, Asim YarKhan, Chongxiao Cao, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra |
| 2015 | ICCS | Performance Analysis and Optimisation of Two-sided Factorization Algorithms for Heterogeneous Platform. | Khairul Kabir, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2015 | PPoPP | Towards batched linear solvers on accelerated hardware platforms. | Azzam Haidar, Tingxing Dong, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra |
| 2015 | PPoPP | Optimization for performance and energy for batched matrix computations on GPUs. | Azzam Haidar, Tingxing Dong, Piotr Luszczek, Stanimire Tomov, Jack J. Dongarra |
| 2015 | SC | Weighted dynamic scheduling with many parallelism grains for offloading of numerical workloads to multiple varied accelerators. | Azzam Haidar, Yulu Jia, Piotr Luszczek, Stanimire Tomov, Asim YarKhan, Jack J. Dongarra |
| 2015 | SC | Efficient implementation of quantum materials simulations on distributed CPU-GPU systems. | Raffaele Solc, Anton Kozhevnikov, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra, Thomas C. Schulthess |
| 2014 | HPCC | LU Factorization of Small Matrices: Accelerating Batched DGETRF on the GPU. | Tingxing Dong, Azzam Haidar, Piotr Luszczek, James Austin Harris, Stanimire Tomov, Jack J. Dongarra |
| 2014 | ICPP | A Fast Batched Cholesky Factorization on a GPU. | Tingxing Dong, Azzam Haidar, Stanimire Tomov, Jack J. Dongarra |
| 2014 | SC | Performance and portability with OpenCL for throughput-oriented HPC workloads across accelerators, coprocessors, and multicore processors. | Chongxiao Cao, Mark Gates, Azzam Haidar, Piotr Luszczek, Stanimire Tomov, Ichitaro Yamazaki, Jack J. Dongarra |
| 2013 | ICS | Toward a scalable multi-GPU eigensolver via compute-intensive kernels and efficient communication. | Azzam Haidar, Mark Gates, Stanimire Tomov, Jack J. Dongarra |
| 2013 | PPAM | Portable HPC Programming on Intel Many-Integrated-Core Hardware with MAGMA Port to Xeon Phi. | Jack J. Dongarra, Mark Gates, Azzam Haidar, Yulu Jia, Khairul Kabir, Piotr Luszczek, Stanimire Tomov |
| 2013 | SC | An improved parallel singular value algorithm and its implementation for multicore hardware. | Azzam Haidar, Jakub Kurzak, Piotr Luszczek |
| 2012 | SC | Abstract: A Novel Hybrid CPU-GPU Generalized Eigensolver for Electronic Structure Calculations Based on Fine Grained Memory Aware Tasks. | Raffaele Solc, Azzam Haidar, Stanimire Tomov, Thomas C. Schulthess, Jack J. Dongarra |
| 2012 | SC | Poster: A Novel Hybrid CPU-GPU Generalized Eigensolver for Electronic Structure Calculations Based on Fine Grained Memory Aware Tasks. | Raffaele Solc, Azzam Haidar, Stanimire Tomov, Thomas C. Schulthess, Jack J. Dongarra |
| 2011 | SC | Parallel reduction to condensed forms for symmetric eigenvalue problems using aggregated fine-grained and memory-aware kernels. | Azzam Haidar, Hatem Ltaief, Jack J. Dongarra |