| 2025 | ICS | Fused3S: Fast Sparse Attention on Tensor Cores. | Zitong Li, Aparna Chandramowlishwaran |
| 2023 | ICPP | ADARNet: Deep Learning Predicts Adaptive Mesh Refinement. | Octavi Obiols-Sales, Abhinav Vishnu, Nicholas Malaya, Aparna Chandramowlishwaran |
| 2023 | SC | Breaking Boundaries: Distributed Domain Decomposition with Scalable Physics-Informed Neural PDE Solvers. | Arthur Feeney, Zitong Li, Ramin Bostanabad, Aparna Chandramowlishwaran |
| 2022 | SC | Lessons Learned on MPI+Threads Communication. | Rohit Zambre, Aparna Chandramowlishwaran |
| 2021 | SIGMETRICS | adPerf: Characterizing the Performance of Third-party Ads. | Behnam Pourghassemi, Jordan Bonecutter, Zhou Li, Aparna Chandramowlishwaran |
| 2020 | ICS | CFDNet: a deep learning-based accelerator for fluid simulations. | Octavi Obiols-Sales, Abhinav Vishnu, Nicholas Malaya, Aparna Chandramowlishwaran |
| 2020 | ICS | How I learned to stop worrying about user-visible endpoints and love MPI. | Rohit Zambre, Aparna Chandramowlishwaran, Pavan Balaji |
| 2020 | SC | Pencil: a pipelined algorithm for distributed stencils. | Hengjie Wang, Aparna Chandramowlishwaran |
| 2020 | SPAA | On the Limits of Parallelizing Convolutional Neural Networks on GPUs. | Behnam Pourghassemi, Chenghao Zhang, Joo Hwan Lee, Aparna Chandramowlishwaran |
| 2019 | EuroPar | Towards Portable Online Prediction of Network Utilization Using MPI-Level Monitoring. | Shu-Mei Tseng, Bogdan Nicolae, George Bosilca, Emmanuel Jeannot, Aparna Chandramowlishwaran, Franck Cappello |
| 2019 | ICPP | Breaking Band: A Breakdown of High-performance Communication. | Rohit Zambre, Megan Grodowitz, Aparna Chandramowlishwaran, Pavel Shamis |
| 2019 | ICS | Multi-criteria partitioning of multi-block structured grids. | Hengjie Wang, Aparna Chandramowlishwaran |
| 2019 | SIGMETRICS | What-If Analysis of Page Load Time in Web Browsers Using Causal Profiling. | Behnam Pourghassemi, Ardalan Amiri Sani, Aparna Chandramowlishwaran |
| 2018 | ASPLOS | Sugar: Secure GPU Acceleration in Web Browsers. | Zhihao Yao, Zongheng Ma, Yingtong Liu, Ardalan Amiri Sani, Aparna Chandramowlishwaran |
| 2018 | ICPADS | Scalable Communication Endpoints for MPI+Threads Applications. | Rohit Zambre, Aparna Chandramowlishwaran, Pavan Balaji |
| 2017 | CLUSTER | cudaCR: An In-Kernel Application-Level Checkpoint/Restart Scheme for CUDA-Enabled GPUs. | Behnam Pourghassemi, Aparna Chandramowlishwaran |
| 2017 | EuroPar | PASCAL: A Parallel Algorithmic SCALable Framework for N-body Problems. | Laleh Aghababaie Beni, Aparna Chandramowlishwaran |
| 2016 | HiPC | Parallel Performance-Energy Predictive Modeling of Browsers: Case Study of Servo. | Rohit Zambre, Lars Bergstrom, Laleh Aghababaie Beni, Aparna Chandramowlishwaran |
| 2014 | ASPLOS | A CPU: GPU Hybrid Implementation and Model-Driven Scheduling of the Fast Multipole Method. | JeeWhan Choi, Aparna Chandramowlishwaran, Kamesh Madduri, Richard W. Vuduc |
| 2012 | SPAA | Brief announcement: towards a communication optimal fast multipole method and its implications at exascale. | Aparna Chandramowlishwaran, JeeWhan Choi, Kamesh Madduri, Richard W. Vuduc |
| 2010 | PPoPP | Applying the concurrent collections programming model to asynchronous parallel dense linear algebra. | Aparna Chandramowlishwaran, Kathleen Knobe, Richard W. Vuduc |
| 2010 | SC | Diagnosis, Tuning, and Redesign for Multicore Performance: A Case Study of the Fast Multipole Method. | Aparna Chandramowlishwaran, Kamesh Madduri, Richard W. Vuduc |
| 2010 | SC | Petascale Direct Numerical Simulation of Blood Flow on 200K Cores and Heterogeneous Architectures. | Abtin Rahimian, Ilya Lashuk, Shravan K. Veerapaneni, Aparna Chandramowlishwaran, Dhairya Malhotra, Logan Moon, Rahul S. Sampath, Aashay Shringarpure, Jeffrey S. Vetter, Richard W. Vuduc, Denis Zorin, George Biros |
| 2009 | POPL | Declarative aspects of memory management in the concurrent collections parallel programming model. | Zoran Budimlic, Aparna Chandramowlishwaran, Kathleen Knobe, Geoff N. Lowney, Vivek Sarkar, Leo Treggiari |
| 2009 | SC | A massively parallel adaptive fast-multipole method on heterogeneous architectures. | Ilya Lashuk, Aparna Chandramowlishwaran, Harper Langston, Tuan-Anh Nguyen, Rahul S. Sampath, Aashay Shringarpure, Richard W. Vuduc, Lexing Ying, Denis Zorin, George Biros |
| 2008 | ICPP | On the Design of Fast Pseudo-Random Number Generators for the Cell Broadband Engine and an Application to Risk Analysis. | David A. Bader, Aparna Chandramowlishwaran, Virat Agarwal |