| 2026 | ACL | EQUIP: EQUivariant preserving In-Place updates for Efficient Token Pruning. | Arun Ramachandran, R. Govindarajan, Murali Annavaram, Prakash Sathyanath Raghavendra |
| 2025 | HiPC | gSubCSR: Dynamic Graph Algorithms for Out-of-GPU-Memory Computation. | Ullas A, R. Govindarajan, Rupesh Nasre |
| 2025 | HiPC | IMBPS - Iterative MLP Blocks with Parameter Splits for Improving LLM Inference. | Ganesh Prasad Nagaraja, Arun Ramachandran, Shubhendu Sharma, R. Govindarajan, Prakash Sathyanath Raghavendra |
| 2025 | ICS | UJOpt: Heuristic Approach for Applying Unroll-and-Jam Optimization and Loop Order Selection. | Shilpa Babalad, Shirish K. Shevade, Matthew Jacob Thazhuthaveetil, R. Govindarajan |
| 2024 | ICS | Tile Size and Loop Order Selection using Machine Learning for Multi-/Many-Core Architectures. | Shilpa Babalad, Shirish K. Shevade, Matthew Jacob Thazhuthaveetil, R. Govindarajan |
| 2024 | SOSP | SilvanForge: A Schedule-Guided Retargetable Compiler for Decision Tree Inference. | Ashwin Prasad, Sampath Rajendra, Kaushik Rajan, R. Govindarajan, Uday Bondhugula |
| 2023 | HiPC | Reduce, Reuse, and Adapt: Accelerating Graph Processing on GPUs. | Ullas A, Rupesh Nasre, R. Govindarajan |
| 2022 | MICRO | Treebeard: An Optimizing Compiler for Decision Tree Based ML Inference. | Ashwin Prasad, Sampath Rajendra, Kaushik Rajan, R. Govindarajan, Uday Bondhugula |
| 2017 | CGO | Taming warp divergence. | Jayvant Anantpur, R. Govindarajan |
| 2015 | CGO | Approximating flow-sensitive pointer analysis using frequent itemset mining. | Vaivaswatha Nagaraj, R. Govindarajan |
| 2014 | CC | Taming Control Divergence in GPUs through Control Flow Linearization. | Jayvant Anantpur, R. Govindarajan |
| 2014 | CGO | Fluidic Kernels: Cooperative Execution of OpenCL Programs on Multiple Heterogeneous Devices. | Prasanna Pandit, R. Govindarajan |
| 2014 | MICRO | Bi-Modal DRAM Cache: Improving Hit Rate, Hit Latency and Bandwidth. | Nagendra Dwarakanath Gulur, Mahesh Mehendale, R. Manikantan, R. Govindarajan |
| 2013 | ASPLOS | Improving GPGPU concurrency with elastic kernels. | Sreepathi Pai, Matthew J. Thazhuthaveetil, R. Govindarajan |
| 2013 | CGO | Runtime dependence computation and execution of loops on heterogeneous systems. | Jayvant Anantpur, R. Govindarajan |
| 2012 | CGO | Reconciling transactional conflicts with compiler's help. | Sandya S. Mannarswamy, R. Govindarajan |
| 2012 | EuroPar | CUDA-For-Clusters: A System for Efficient Execution of CUDA Kernels on Multi-core Clusters. | Raghu Prabhakar, R. Govindarajan, Matthew J. Thazhuthaveetil |
| 2012 | ISCA | Probabilistic Shared Cache Management (PriSM). | R. Manikantan, Kaushik Rajan, R. Govindarajan |
| 2012 | ICS | Multiple sub-row buffers in DRAM: unlocking performance and energy improvement opportunities. | Nagendra Dwarakanath Gulur, R. Manikantan, Mahesh Mehendale, R. Govindarajan |
| 2011 | HPCA | NUcache: An efficient multicore cache organization based on Next-Use distance. | R. Manikantan, Kaushik Rajan, R. Govindarajan |
| 2011 | PLDI | Automatic compilation of MATLAB programs for synergistic execution on heterogeneous processors. | Ashwin Prasad, Jayvant Anantpur, R. Govindarajan |
| 2010 | ICPP | Handling Conflicts with Compiler's Help in Software Transactional Memory Systems. | Sandya Mannarswamy, R. Govindarajan |
| 2009 | CGO | Software Pipelined Execution of Stream Programs on GPUs. | Abhishek Udupa, R. Govindarajan, Matthew J. Thazhuthaveetil |
| 2008 | CGO | Comprehensive path-sensitive data-flow analysis. | Aditya V. Thakur, R. Govindarajan |
| 2008 | Interspeech | Online unsupervised pattern discovery in speech using parallelization. | Mrugesh R. Gajjar, R. Govindarajan, T. V. Sreenivas |
| 2008 | ICS | Focused prefetching: performance oriented prefetching based on commit stalls. | R. Manikantan, R. Govindarajan |
| 2008 | VLSID | Memory Architecture Exploration Framework for Cache Based Embedded SOC. | T. S. Rajesh Kumar, C. P. Ravikumar, R. Govindarajan |
| 2007 | ASPDAC | MODLEX: A Multi Objective Data Layout EXploration Framework for Embedded Systems-on-Chip. | T. S. Rajesh Kumar, C. P. Ravikumar, R. Govindarajan |
| 2007 | CC | Register Allocation and Optimal Spill Code Scheduling in Software Pipelined Loops Using 0-1 Integer Linear Programming Formulation. | Santosh Nagarakatte, R. Govindarajan |
| 2007 | CC | An Array Allocation Scheme for Energy Reduction in Partitioned Memory Architectures. | K. Shyam, R. Govindarajan |
| 2007 | HiPC | Compiler-Directed Dynamic Voltage Scaling Using Program Phases. | K. Shyam, R. Govindarajan |
| 2007 | VLSID | MAX: A Multi Objective Memory Architecture eXploration Framework for Embedded Systems-on-Chip. | T. S. Rajesh Kumar, C. P. Ravikumar, R. Govindarajan |
| 2006 | ICS | A scalable low power issue queue for large instruction window processors. | Rajesh Vivekanandham, Bharadwaj S. Amrutur, R. Govindarajan |
| 2005 | HiPC | Offloading Bloom Filter Operations to Network Processor for Parallel Query Processing in Cluster of Workstations. | V. Santhosh Kumar, Matthew J. Thazhuthaveetil, R. Govindarajan |
| 2003 | HiPC | An Efficient Web Cache Replacement Policy. | A. Radhika Sarma, R. Govindarajan |
| 2003 | VLSID | Optimal Code and Data Layout in Embedded Systems. | T. S. Rajesh Kumar, R. Govindarajan, C. P. Ravikumar |
| 2003 | SCOPES | Unified Instruction Reordering and Algebraic Transformations for Minimum Cost Offset Assignment. | V. V. N. S. Sarvani, R. Govindarajan |
| 2002 | HiPC | Dynamic Path Profile Aided Recompilation in a JAVA Just-In-Time Compiler. | R. Vinodh Kumar, B. Lakshmi Narayanan, R. Govindarajan |
| 2001 | HiPC | Hidden Costs in Avoiding False Sharing in Software DSMs. | K. V. Manjunath, R. Govindarajan |
| 2001 | ICCAD | Area and Power Reduction of Embedded DSP Systems using Instruction Compression and Re-Configurable Encoding. | Subash Chandar G., Mahesh Mehendale, R. Govindarajan |
| 1998 | HiPC | Modulo-variable expansion sensitive scheduling. | Madhavi Gopal Valluri, R. Govindarajan |
| 1998 | SMC | Performance bounds for distributed memory multithreaded architectures. | Wlodzimierz M. Zuberek, R. Govindarajan |
| 1997 | ICPADS | Distributed Shared Memory on IBM SP2. | S. Ramesh, R. Lakshmi, R. Govindarajan |
| 1992 | COMPSAC | Software fault-tolerance in functional programming. | R. Govindarajan |
| 1992 | ICASSP | Well-behaved dataflow programs for DSP computation. | Guang R. Gao, R. Govindarajan, Prakash Panangaden |
| 1991 | COMPSAC | ParC project: practical constructs for parallel programming languages. | R. Govindarajan, Lifu Guo, Sheng Yu, P. Wang |