Pedro Valero-Lara
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
35
Venues
16
Active years
2011–2025
Best venue rank
A*
Where they publish
Papers
35 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | SC | Mojo: MLIR-based Performance-Portable HPC Science Kernels on GPUs for the Python Ecosystem. | William F. Godoy, Tatiana Melnichenko, Pedro Valero-Lara, Wael R. Elwasif, Philip W. Fackler, Rafael Ferreira da Silva, Keita Teranishi, Jeffrey S. Vetter |
| 2025 | SC | ChatHPC: Building the Foundations for a Productive and Trustworthy AI-Assisted HPC Ecosystem. | Pedro Valero-Lara, Aaron R. Young, Jeffrey S. Vetter, Zheming Jin, Swaroop Pophale, Mohammad Alaul Haque Monil, Keita Teranishi, William F. Godoy |
| 2024 | SC | Integrating ORNL's HPC and Neutron Facilities with a Performance-Portable CPU/GPU Ecosystem. | Steven E. Hahn, Philip W. Fackler, William F. Godoy, Ketan Maheshwari, Zachary Morgan, Andrei T. Savici, Christina M. Hoffmann, Pedro Valero-Lara, Jeffrey S. Vetter, Rafael Ferreira da Silva |
| 2024 | SC | JACC: Leveraging HPC Meta-Programming and Performance Portability with the Just-in-Time and LLVM-based Julia Language. | Pedro Valero-Lara, William F. Godoy, Het Mankad, Keita Teranishi, Jeffrey S. Vetter, Johannes P. Blaschke, Michel Schanen |
| 2024 | SC | ChatBLAS: The First AI-Generated and Portable BLAS Library. | Pedro Valero-Lara, William F. Godoy, Keita Teranishi, Prasanna Balaprakash, Jeffrey S. Vetter |
| 2023 | PLDI | A MultiGPU Performance-Portable Solution for Array Programming Based on Kokkos. | Pedro Valero-Lara, Jeffrey S. Vetter |
| 2023 | PPoPP | Tiling Framework for Heterogeneous Computing of Matrix based Tiled Algorithms. | Narasinga Rao Miniskar, Mohammad Alaul Haque Monil, Pedro Valero-Lara, Frank Liu, Jeffrey S. Vetter |
| 2023 | SC | Julia as a unifying end-to-end workflow language on the Frontier exascale system. | William F. Godoy, Pedro Valero-Lara, Caira Anderson, Katrina W. Lee, Ana Gainaru, Rafael Ferreira da Silva, Jeffrey S. Vetter |
| 2023 | SC | MatRIS: Multi-level Math Library Abstraction for Heterogeneity and Performance Portability using IRIS Runtime. | Mohammad Alaul Haque Monil, Narasinga Rao Miniskar, Keita Teranishi, Jeffrey S. Vetter, Pedro Valero-Lara |
| 2023 | SC | Mixed-Precision S/DGEMM Using the TF32 and TF64 Frameworks on Low-Precision AI Tensor Cores. | Pedro Valero-Lara, Ian Jorquera, Frank Liu, Jeffrey S. Vetter |
| 2023 | SC | Moment Representation of Regularized Lattice Boltzmann Methods on NVIDIA and AMD GPUs. | Pedro Valero-Lara, Jeffrey S. Vetter, John Gounley, Amanda Randles |
| 2022 | EuroPar | A Portable and Heterogeneous LU Factorization on IRIS. | Pedro Valero-Lara, Jungwon Kim, Jeffrey S. Vetter |
| 2022 | HiPC | IRIS-BLAS: Towards a Performance Portable and Heterogeneous BLAS Library. | Narasinga Rao Miniskar, Mohammad Alaul Haque Monil, Pedro Valero-Lara, Frank Liu, Jeffrey S. Vetter |
| 2022 | SC | LaRIS: Targeting Portability and Productivity for LAPACK Codes on Extreme Heterogeneous Systems by Using IRIS. | Mohammad Alaul Haque Monil, Narasinga Rao Miniskar, Frank Y. Liu, Jeffrey S. Vetter, Pedro Valero-Lara |
| 2021 | EuroPar | OpenMP Target Task: Tasking and Target Offloading on Heterogeneous Systems. | Pedro Valero-Lara, Jungwon Kim, Oscar R. Hernandez, Jeffrey S. Vetter |
| 2021 | HiPC | Static Graphs for Coding Productivity in OpenACC. | Leonel Toledo, Pedro Valero-Lara, Jeffrey S. Vetter, Antonio J. Pea |
| 2019 | PDCAT | Accelerating Conjugate Gradient using OmpSs. | Sandra Cataln, Xavier Martorell, Jess Labarta, Tetsuzo Usui, Leonel Antonio Toledo Daz, Pedro Valero-Lara |
| 2019 | PDCAT | Tasking in Accelerators: Performance Evaluation. | Leonel Toledo, Antonio J. Pea, Sandra Cataln, Pedro Valero-Lara |
| 2019 | PDP | BLAS-3 Optimized by OmpSs Regions (LASs Library). | Pedro Valero-Lara, Sandra Cataln, Xavier Martorell, Jess Labarta |
| 2018 | PDP | Variable Batched DGEMM. | Pedro Valero-Lara, Ivan Martnez-Prez, Sergi Mateo, Ral Sirvent, Vicen Beltran, Xavier Martorell, Jess Labarta |
| 2017 | ICCS | The Design and Performance of Batched BLAS on Modern High-Performance Computing Systems. | Jack J. Dongarra, Sven Hammarling, Nicholas J. Higham, Samuel D. Relton, Pedro Valero-Lara, Mawussi Zounon |
| 2017 | ICCS | cuHinesBatch: Solving Multiple Hines systems on GPUs Human Brain Project | Pedro Valero-Lara, Ivan Martnez-Prez, Antonio J. Pea, Xavier Martorell, Ral Sirvent, Jess Labarta |
| 2017 | IWANN | Heuristics for ROSA's LTS Searching. | Fernando Lpez Pelayo, Fernando Cuartero Gmez, Diego Cazorla, Pedro Valero-Lara, Mercedes G. Merayo |
| 2017 | PPAM | NVIDIA GPUs Scalability to Solve Multiple (Batch) Tridiagonal Systems Implementation of cuThomasBatch. | Pedro Valero-Lara, Ivan Martnez-Prez, Ral Sirvent, Xavier Martorell, Antonio J. Pea |
| 2016 | ICA3PP | Leveraging the Performance of LBM-HPC for Large Sizes on GPUs Using Ghost Cells. | Pedro Valero-Lara |
| 2015 | CLUSTER | LBM-HPC - An Open-Source Tool for Fluid Simulations. Case Study: Unified Parallel C (UPC-PGAS). | Pedro Valero-Lara, Johan Jansson |
| 2015 | ICCS | A Non-uniform Staggered Cartesian Grid Approach for Lattice-boltzmann Method. | Pedro Valero-Lara, Johan Jansson |
| 2014 | CLUSTER | Multi-GPU acceleration of DARTEL (early detection of Alzheimer). | Pedro Valero-Lara |
| 2014 | ICCS | Accelerating Solid-fluid Interaction using Lattice-boltzmann and Immersed Boundary Coupled Simulations on Heterogeneous Platforms. | Pedro Valero-Lara, Alfredo Pinelli, Manuel Prieto-Matas |
| 2014 | ICCSA | hLCS. A Hybrid GPGPU Approach for Solving Multiple Short and Unbalanced LCS Problems. | Pedro Valero-Lara |
| 2013 | ICPP | GPU Powered ROSA Analyzer. | Ral Pardo, Fernando L. Pelayo, Pedro Valero-Lara |
| 2012 | DEXA | Improving the Performance for the Range Search on Metric Spaces Using a Multi-GPU Platform. | Roberto Uribe Paredes, Enrique Arias, Jos L. Snchez, Diego Cazorla, Pedro Valero-Lara |
| 2012 | ISPA | Block Tridiagonal Solvers on Heterogeneous Architectures. | Pedro Valero-Lara, Alfredo Pinelli, Julien Favier, Manuel Prieto-Matas |
| 2011 | ICCSA | A GPU-Based Implementation for Range Queries on Spaghettis Data Structure. | Roberto Uribe Paredes, Pedro Valero-Lara, Enrique Arias, Jos L. Snchez, Diego Cazorla |
| 2011 | ICCSA | Towards a More Efficient Use of GPUs. | Pedro Valero-Lara, Fernando L. Pelayo |