| 2026 | CCGRID | Scalable, Topology- and Multi-HCA-Aware Hierarchical GPU Allgather Using Parallel Rings. | Amirreza Barati Sedeh, Ryan Grant, Ahmad Afsahi |
| 2025 | CLUSTER | Cascade: a Collaborative Algorithm for Scalable and Efficient Neighborhood Allgather. | Hamed Sharifian, Amir Hossein Sojoodi, Ahmad Afsahi |
| 2025 | PDP | Utilizing Network Hardware Parallelism for MPI Partitioned Collective Communication. | Yiltan Hassan Temuin, Amirreza Barati Sedeh, Whit Schonbein, Ryan E. Grant, Ahmad Afsahi |
| 2025 | SC | Accelerating Intra-Node GPU Communication: A Performance Model for Multi-Path Transfers. | Amir Hossein Sojoodi, Mohammad Akbari, Hamed Sharifian, Ali Farazdaghi, Ryan E. Grant, Ahmad Afsahi |
| 2024 | CLUSTER | A Topology- and Load-Aware Design for Neighborhood Allgather. | Hamed Sharifian, Amir Hossein Sojoodi, Ahmad Afsahi |
| 2024 | SC | Design and Implementation of MPI-Native GPU-Initiated MPI Partitioned Communication. | Yiltan Hassan Temuin, Whit Schonbein, Scott Levy, Amir Hossein Sojoodi, Ryan E. Grant, Ahmad Afsahi |
| 2023 | CLUSTER | A Dynamic Network-Native MPI Partitioned Aggregation Over InfiniBand Verbs. | Yiltan Hassan Temuin, Scott Levy, Whit Schonbein, Ryan E. Grant, Ahmad Afsahi |
| 2022 | ICPP | Micro-Benchmarking MPI Partitioned Point-to-Point Communication. | Yiltan Hassan Temuin, Ryan E. Grant, Ahmad Afsahi |
| 2021 | HOTI | Efficient Multi-Path NVLink/PCIe-Aware UCX based Collective Communication for Deep Learning. | Yiltan Hassan Temuin, Amir Hossein Sojoodi, Pedram Alizadeh, Ahmad Afsahi |
| 2019 | CCGRID | Fuzzy Matching: Hardware Accelerated MPI Communication Middleware. | Matthew G. F. Dosanjh, Whit Schonbein, Ryan E. Grant, Patrick G. Bridges, S. Mahdieh Ghazimirsaeed, Ahmad Afsahi |
| 2018 | ICPP | The Case for Semi-Permanent Cache Occupancy: Understanding the Impact of Data Locality on Network Processing. | Matthew G. F. Dosanjh, S. Mahdieh Ghazimirsaeed, Ryan E. Grant, Whit Schonbein, Michael J. Levenhagen, Patrick G. Bridges, Ahmad Afsahi |
| 2017 | HiPC | Exploiting Common Neighborhoods to Optimize MPI Neighborhood Collectives. | Seyed Hessam Mirsadeghi, Jesper Larsson Trff, Pavan Balaji, Ahmad Afsahi |
| 2016 | SBAC-PAD | MAGC: A Mapping Approach for GPU Clusters. | Seyed Hessam Mirsadeghi, Iman Faraji, Ahmad Afsahi |
| 2015 | SC | Hyper-Q aware intranode MPI collectives on the GPU. | Iman Faraji, Ahmad Afsahi |
| 2014 | SC | Nonblocking Epochs in MPI One-Sided Communication. | Judicael A. Zounmevo, Xin Zhao, Pavan Balaji, William Gropp, Ahmad Afsahi |
| 2013 | CCGRID | Toward Asynchronous and MPI-Interoperable Active Messages. | Xin Zhao, Darius Buntinas, Judicael A. Zounmevo, James Dinan, David Goodell, Pavan Balaji, Rajeev Thakur, Ahmad Afsahi, William Gropp |
| 2013 | CLUSTER | Mercury: Enabling remote procedure call for high-performance computing. | Jrome Soumagne, Dries Kimpe, Judicael A. Zounmevo, Mohamad Chaarawi, Quincey Koziol, Ahmad Afsahi, Robert B. Ross |
| 2012 | CLUSTER | Designing an Offloaded Nonblocking MPI_Allgather Collective Using CORE-Direct. | Grigori Inozemtsev, Ahmad Afsahi |
| 2012 | ICPADS | An Efficient MPI Message Queue Mechanism for Large-scale Jobs. | Judicael A. Zounmevo, Ahmad Afsahi |
| 2011 | CLUSTER | Investigating Scenario-Conscious Asynchronous Rendezvous over RDMA. | Judicael A. Zounmevo, Ahmad Afsahi |
| 2010 | HiPC | iWARP redefined: Scalable connectionless communication over high-speed Ethernet. | Mohammad J. Rashti, Ryan E. Grant, Ahmad Afsahi, Pavan Balaji |
| 2010 | ISPASS | A study of hardware assisted IP over InfiniBand and its impact on enterprise data center performance. | Ryan E. Grant, Pavan Balaji, Ahmad Afsahi |
| 2009 | ICPADS | Evaluation of ConnectX Virtual Protocol Interconnect for Data Centers. | Ryan E. Grant, Ahmad Afsahi, Pavan Balaji |
| 2007 | CLUSTER | Improving system efficiency through scheduling and power management. | Ryan E. Grant, Ahmad Afsahi |
| 2007 | CLUSTER | A feasibility analysis of power-awareness and energy minimization in modern interconnects for high-performance computing. | Reza Zamani, Ahmad Afsahi, Ying Qian, V. Carl Hamacher |
| 2007 | HOTI | Assessing the Ability of Computation/Communication Overlap and Communication Progress in Modern Interconnects. | Mohammad J. Rashti, Ahmad Afsahi |
| 2007 | ICPP | RDMA-based and SMP-aware Multi-port All-Gather on Multi-rail QsNet^II SMP Clusters. | Ying Qian, Ahmad Afsahi |
| 2004 | NCA | Myrinet Networks: A Performance Study. | Ying Qian, Ahmad Afsahi, Reza Zamani |
| 2003 | ICS | Performance characteristics of openMP constructs, and application benchmarks on a large symmetric multiprocessor. | Nathan R. Fredrickson, Ahmad Afsahi, Ying Qian |
| 1997 | OPODIS | Collective Communications on a Reconfigurable Optical Interconnect. | Ahmad Afsahi, Nikitas J. Dimopoulos |