| 2024 | ICCD | PCCL: Energy-Efficient LLM Training with Power-Aware Collective Communication. | Ziyang Jia, Laxmi N. Bhuyan, Daniel Wong |
| 2023 | PPoPP | Improving Energy Saving of One-Sided Matrix Decompositions on CPU-GPU Heterogeneous Systems. | Jieyang Chen, Xin Liang, Kai Zhao, Hadi Zamani Sabzi, Laxmi N. Bhuyan, Zizhong Chen |
| 2022 | HPCA | Cottage: Coordinated Time Budget Assignment for Latency, Quality and Power Optimization in Web Search. | Liang Zhou, Laxmi N. Bhuyan, K. K. Ramakrishnan |
| 2022 | ICNP | Synergy: A SmartNIC Accelerated 5G Dataplane and Monitor for Mobility Prediction. | Sourav Panda, K. K. Ramakrishnan, Laxmi N. Bhuyan |
| 2021 | CoNEXT | SmartWatch: accurate traffic analysis and flow-state tracking for intrusion prevention using SmartNICs. | Sourav Panda, Yixiao Feng, Sameer G. Kulkarni, K. K. Ramakrishnan, Nick Duffield, Laxmi N. Bhuyan |
| 2021 | ICNP | pMACH: Power and Migration Aware Container scHeduling. | Sourav Panda, K. K. Ramakrishnan, Laxmi N. Bhuyan |
| 2020 | ISLPED | Slumber: static-power management for GPGPU register files. | Devashree Tripathy, Hadi Zamani, Debiprasanna Sahoo, Laxmi N. Bhuyan, Manoranjan Satpathy |
| 2020 | ISLPED | SAOU: safe adaptive overclocking and undervolting for energy-efficient GPU computing. | Hadi Zamani, Devashree Tripathy, Laxmi N. Bhuyan, Zizhong Chen |
| 2020 | ISLPED | Swan: a two-step power management for distributed search engines. | Liang Zhou, Laxmi N. Bhuyan, K. K. Ramakrishnan |
| 2020 | MICRO | Gemini: Learning to Manage CPU Power for Latency-Critical Search Engines. | Liang Zhou, Laxmi N. Bhuyan, K. K. Ramakrishnan |
| 2019 | HPCA | μDPM: Dynamic Power Management for the Microsecond Era. | Chih-Hsun Chou, Laxmi N. Bhuyan, Daniel Wong |
| 2019 | ICDCS | Goldilocks: Adaptive Resource Provisioning in Containerized Data Centers. | Liang Zhou, Laxmi N. Bhuyan, K. K. Ramakrishnan |
| 2019 | ICS | GreenMM: energy efficient GPU matrix multiplication through undervolting. | Hadi Zamani, Yuanlai Liu, Devashree Tripathy, Laxmi N. Bhuyan, Zizhong Chen |
| 2018 | PPoPP | Juggler: a dependence-aware task-based execution framework for GPUs. | Mehmet E. Belviranli, Seyong Lee, Jeffrey S. Vetter, Laxmi N. Bhuyan |
| 2017 | ICDCS | TailCut: Power Reduction under Quality and Latency Constraints in Distributed Search Systems. | Chih-Hsun Chou, Laxmi N. Bhuyan, Shaolei Ren |
| 2017 | SC | Enabling Work-Efficiency for High Performance Vertex-Centric Graph Analytics on GPUs. | Farzad Khorasani, Keval Vora, Rajiv Gupta, Laxmi N. Bhuyan |
| 2016 | ICS | CuMAS: Data Transfer Aware Multi-Application Scheduling for Shared GPUs. | Mehmet E. Belviranli, Farzad Khorasani, Laxmi N. Bhuyan, Rajiv Gupta |
| 2016 | ISLPED | DynSleep: Fine-grained Power Management for a Latency-Critical Data Center Application. | Chih-Hsun Chou, Daniel Wong, Laxmi N. Bhuyan |
| 2016 | SC | GreenLA: green linear algebra software for GPU-accelerated heterogeneous computing. | Jieyang Chen, Li Tan, Panruo Wu, Dingwen Tao, Hongbo Li, Xin Liang, Sihuan Li, Rong Ge, Laxmi N. Bhuyan, Zizhong Chen |
| 2015 | ICCD | A multicore vacation scheme for thermal-aware packet processing. | Chih-Hsun Chou, Laxmi N. Bhuyan |
| 2015 | ICS | PeerWave: Exploiting Wavefront Parallelism on GPUs with Peer-SM Synchronization. | Mehmet E. Belviranli, Peng Deng, Laxmi N. Bhuyan, Rajiv Gupta, Qi Zhu |
| 2015 | MICRO | Efficient warp execution in presence of divergence with collaborative context collection. | Farzad Khorasani, Rajiv Gupta, Laxmi N. Bhuyan |
| 2014 | HPDC | A paradigm shift in GP-GPU computing: task based execution of applications with dynamic data dependencies. | Mehmet E. Belviranli, Chih-Hsun Chou, Laxmi N. Bhuyan, Rajiv Gupta |
| 2014 | HPDC | CuSha: vertex-centric graph processing on GPUs. | Farzad Khorasani, Keval Vora, Rajiv Gupta, Laxmi N. Bhuyan |
| 2014 | ICPADS | fAHRW | Qin Liu, Laxmi N. Bhuyan |
| 2013 | FPL | Shared memory heterogeneous computation on PCIe-supported platforms. | Sambit Kumar Shukla, Yang Yang, Laxmi N. Bhuyan, Philip Brisk |
| 2013 | HiPC | A hybrid shared memory heterogeneous execution platform for PCIe-based GPGPUs. | Sambit Kumar Shukla, Laxmi N. Bhuyan |
| 2012 | DAC | Traffic-aware power optimization for network applications on multicore servers. | Jilong Kuang, Laxmi N. Bhuyan, Raymond Klefstad |
| 2012 | IPCCC | An efficient dynamic multiple-candidate motion vector approach for GPU-based hierarchical motion estimation. | Dung Vu, Yang Yang, Laxmi N. Bhuyan |
| 2012 | IWQoS | Improving the throughput and delay performance of network processors by applying push model. | Bin Liu, Bo Yuan, Huichen Dai, Hongbo Zhao, Jia Yu, Laxmi N. Bhuyan |
| 2012 | PPoPP | Speculative parallelization on GPGPUs. | Min Feng, Rajiv Gupta, Laxmi N. Bhuyan |
| 2011 | HPCA | A new server I/O architecture for high speed networks. | Guangdeng Liao, Xia Zhu, Laxmi N. Bhuyan |
| 2011 | INFOCOM | A QoS aware multicore hash scheduler for network applications. | Danhua Guo, Laxmi N. Bhuyan |
| 2010 | DAC | LATA: a latency and throughput-aware packet processing system. | Jilong Kuang, Laxmi N. Bhuyan |
| 2010 | DAC | A new IP lookup cache for high performance IP routers. | Guangdeng Liao, Heeyeol Yu, Laxmi N. Bhuyan |
| 2010 | GLOBECOM | Experience on Applying Push Model to Packet Processors in High Performance Routers. | Bo Yuan, Hongbo Zhao, Chengchen Hu, Bin Liu, Jia Yu, Laxmi N. Bhuyan |
| 2010 | HOTI | Understanding Power Efficiency of TCP/IP Packet Processing over 10GbE. | Guangdeng Liao, Xia Zhu, Steen Larsen, Laxmi N. Bhuyan, Ram Huggahalli |
| 2010 | INFOCOM | A Balanced Consistency Maintenance Protocol for Structured P2P Systems. | Yi Hu, Min Feng, Laxmi N. Bhuyan |
| 2010 | INFOCOM | Optimizing Throughput and Latency under Given Power Budget for Network Packet Processing. | Jilong Kuang, Laxmi N. Bhuyan |
| 2009 | HOTI | Performance Measurement of an Integrated NIC Architecture with 10GbE. | Guangdeng Liao, Laxmi N. Bhuyan |
| 2009 | ICNP | A Hash-based Scalable IP lookup using Bloom and Fingerprint Filters. | Heeyeol Yu, Rabi N. Mahapatra, Laxmi N. Bhuyan |
| 2009 | INFOCOM | Budget-Based Self-Optimized Incentive Search in Unstructured P2P Networks. | Yi Hu, Min Feng, Laxmi N. Bhuyan, Vana Kalogeraki |
| 2008 | DSD | Revisiting the Cache Effect on Multicore Multithreaded Network Processors. | Zhen Liu, Jia Yu, Xiaojun Wang, Bin Liu, Laxmi N. Bhuyan |
| 2008 | ICCCN | A Novel Service-Aware Message Scheduler for Cisco Application Oriented Networking Systems. | Jingnan Yao, Jianxun Jason Ding, Laxmi N. Bhuyan |
| 2008 | ICDCS | Quantum-Adaptive Scheduling for Multi-Core Network Processors. | Yue Zhang, Bin Liu, Lei Shi, Jingnan Yao, Laxmi N. Bhuyan |
| 2008 | INFOCOM | Cyber-Fraud is One Typo Away. | Anirban Banerjee, Dhiman Barman, Michalis Faloutsos, Laxmi N. Bhuyan |
| 2007 | DAC | Program Mapping onto Network Processors by Recursive Bipartitioning and Refining. | Jia Yu, Jingnan Yao, Laxmi N. Bhuyan, Jun Yang |
| 2007 | GLOBECOM | Clustered K-Center: Effective Replica Placement in Peer-to-Peer Systems. | Jian Zhou, Xin Zhang, Laxmi N. Bhuyan, Bin Liu |
| 2007 | INFOCOM | Lexicographic Fairness in WDM Optical Cross-Connects. | Satya Ranjan Mohanty, Laxmi N. Bhuyan |
| 2007 | INFOCOM | Adaptive Max-Min Fair Scheduling in Buffered Crossbar Switches Without Speedup. | Xiao Zhang, Satya Ranjan Mohanty, Laxmi N. Bhuyan |
| 2007 | IPCCC | Scalable and Decentralized Content-Aware Dispatching in Web Clusters. | Zhiyong Xu, Jizhong Han, Laxmi N. Bhuyan |
| 2007 | Networking | The P2P War: Someone Is Monitoring Your Activities! | Anirban Banerjee, Michalis Faloutsos, Laxmi N. Bhuyan |
| 2006 | CCGRID | Effective Load Balancing in P2P Systems. | Zhiyong Xu, Laxmi N. Bhuyan |
| 2006 | IPCCC | Efficient server cooperation mechanism in content delivery network. | Zhiyong Xu, Yiming Hu, Laxmi N. Bhuyan |
| 2006 | LCN | Fair Scheduling over multiple servers with flow-dependent server rate. | Satya Ranjan Mohanty, Laxmi N. Bhuyan |
| 2006 | LCN | Computing Real Time Jobs in P2P Networks. | Jingnan Yao, Jian Zhou, Laxmi N. Bhuyan |
| 2005 | DAC | Low power network processor design using clock gating. | Yan Luo, Jia Yu, Jun Yang, Laxmi N. Bhuyan |
| 2005 | GLOBECOM | Guaranteed smooth switch scheduling with low complexity. | Satya Ranjan Mohanty, Laxmi N. Bhuyan |
| 2005 | GLOBECOM | QoS-aware object replica placement in CDNs. | Zhiyong Xu, Laxmi N. Bhuyan |
| 2005 | GLOBECOM | Distributed packet processing in P2P networks. | Jingnan Yao, Laxmi N. Bhuyan |
| 2005 | GLOBECOM | Optimal network processor topologies for efficient packet processing. | Jingnan Yao, Yan Luo, Laxmi N. Bhuyan, Ravishankar R. Iyer |
| 2005 | GLOBECOM | Achieving fairness and throughput for best-effort traffic in input-queued crossbar switches. | Xiao Zhang, Laxmi N. Bhuyan |
| 2005 | HOTI | Performance Characterization of a 10-Gigabit Ethernet TOE. | Wu-chun Feng, Pavan Balaji, Christopher Baron, Laxmi N. Bhuyan, Dhabaleswar K. Panda |
| 2005 | HOTI | Design and Implementation of a Content-Aware Switch Using a Network Processor. | Li Zhao, Yan Luo, Laxmi N. Bhuyan, Ravishankar R. Iyer |
| 2005 | ICCCN | On fair scheduling in heterogeneous link aggregated services. | Satya Ranjan Mohanty, Laxmi N. Bhuyan |
| 2005 | ICCD | Hardware Support for Bulk Data Movement in Server Platforms. | Li Zhao, Ravi R. Iyer, Srihari Makineni, Laxmi N. Bhuyan, Donald Newell |
| 2005 | INFOCOM | An efficient packet scheduling algorithm in network processors. | Jiani Guo, Jingnan Yao, Laxmi N. Bhuyan |
| 2005 | IPCCC | Efficient file sharing strategy in DHT based P2P systems. | Zhinyong Xu, Xubin He, Laxmi N. Bhuyan |
| 2005 | ISPASS | Anatomy and Performance of SSL Processing. | Li Zhao, Ravi R. Iyer, Srihari Makineni, Laxmi N. Bhuyan |
| 2004 | DATE | Utilizing Formal Assertions for System Design of Network Processors. | Xi Chen, Yan Luo, Harry Hsieh, Laxmi N. Bhuyan, Felice Balarin |
| 2004 | GLOBECOM | Scheduling real-time multimedia tasks in network processors. | Jingnan Yao, Jiani Guo, Laxmi N. Bhuyan, Zhiyong Xu |
| 2004 | GLOBECOM | An efficient scheduling algorithm for combined input-crosspoint-queued (CICQ) switches. | Xiao Zhang, Laxmi N. Bhuyan |
| 2004 | ICPADS | Load Balancing of DNS-Based Distributed Web Server Systems with Page Caching. | Zhong Xu, Rong Huang, Laxmi N. Bhuyan |
| 2003 | CASES | Power efficient encoding techniques for off-chip data buses. | Dinesh C. Suresh, Banit Agrawal, Jun Yang, Walid A. Najjar, Laxmi N. Bhuyan |
| 2002 | INFOCOM | Fair Scheduling and Buffer Management in Internet Routers. | Nan Ni, Laxmi N. Bhuyan |
| 2001 | NCA | Execution-Driven Simulation of IP Router Architectures. | Laxmi N. Bhuyan, Hu-Jun Wang |
| 2000 | ICCD | Hierarchical Simulation of a Multiprocessor Architecture. | Marius Pirvu, Laxmi N. Bhuyan, Rabi N. Mahapatra |
| 2000 | ICS | Hardware spatial forwarding for widely shared data. | Marius Pirvu, Laxmi N. Bhuyan |
| 1999 | HPCA | Switch Cache: A Framework for Improving the Remote Memory Access Latency of CC-NUMA Multiprocessors. | Ravi R. Iyer, Laxmi N. Bhuyan |
| 1999 | HPCA | The Impact of Link Arbitration on Switch Performance. | Marius Pirvu, Laxmi N. Bhuyan, Nan Ni |
| 1999 | ICS | Comparing the memory system performance of the HP V-class and SGI Origin 2000 multiprocessors using microbenchmarks and scientific applications. | Ravi R. Iyer, Nancy M. Amato, Lawrence Rauchwerger, Laxmi N. Bhuyan |
| 1998 | ICCD | Circular buffered switch design with wormhole routing and virtual channels. | Nan Ni, Marius Pirvu, Laxmi N. Bhuyan |
| 1996 | ICPP | An Efficient Hybrid Cache Coherence Protocol for Shared Memory Multiprocessors. | Yeimkuan Chang, Laxmi N. Bhuyan |
| 1996 | ICS | Evaluating Virtual Channels for Cache-Coherent Shared-Memory Multiprocessors. | Akhilesh Kumar, Laxmi N. Bhuyan |
| 1995 | ICCCN | valuation of multi-queue buffered multistage interconnection networks under uniform and nonuniform traffic patterns. | Jianxun Jason Ding, Laxmi N. Bhuyan |
| 1995 | ICCD | A dynamic cache sub-block design to reduce false sharing. | Murali Kadiyala, Laxmi N. Bhuyan |
| 1995 | ICPP | Partitioning an Arbitrary Multicomputer Architecture. | Laxmi N. Bhuyan, Sumon Shahed, Yeimkuan Chang |
| 1995 | ICPP | A Submesh Allocation Scheme for Mesh-Connected Multiprocessor Systems. | Tong Liu, Wei-Kang Huang, Fabrizio Lombardi, Laxmi N. Bhuyan |
| 1994 | ICPP | Performance and Reliability of the Multistage Bus Network. | Laxmi N. Bhuyan, Ashwini K. Nanda, Tahsin Askar |
| 1994 | ICPP | A Distributed Cache Coherence Protocol for Hypercube Multiprocessors. | Yeimkuan Chang, Laxmi N. Bhuyan, Akhilesh Kumar |
| 1994 | SC | Efficient and scalable cache coherence schemes for shared memory hypercube multiprocessors. | Akhilesh Kumar, Phanindra K. Mannava, Laxmi N. Bhuyan |
| 1993 | ICPP | Fault Tolerant Subcube Allocation in Hypercubes. | Yeimkuan Chang, Laxmi N. Bhuyan |
| 1993 | ICPP | An Adaptive Submesh Allocation Strategy For Two-Dimensional Mesh Connected Systems. | Jianxun Ding, Laxmi N. Bhuyan |
| 1993 | ICPP | An Adaptive System-Level Diagnosis Approach for Mesh Connected Multiprocessors. | Chao Feng, Laxmi N. Bhuyan, Fabrizio Lombardi |
| 1993 | ICPP | Parallel FFT Algorithms for Cache Based Shared Memory Multiprocessors. | Akhilesh Kumar, Laxmi N. Bhuyan |
| 1992 | ICPP | Extending Multistage Interconnection Networks for Multitasking. | Yeimkuan Chang, Laxmi N. Bhuyan |
| 1992 | ICPP | A Formal Specification and Verification Technique for Cache Coherence Protocols. | Ashwini K. Nanda, Laxmi N. Bhuyan |
| 1992 | SC | Mapping Applications onto a Cache Coherent Multiprocessor. | Ashwini K. Nanda, Laxmi N. Bhuyan |
| 1991 | ICDCS | Load balancing with network cooperation. | Margaret A. Schaar, Kemal Efe, Lois M. L. Delcambre, Laxmi N. Bhuyan |
| 1991 | ICPP | Performance Evaluation of Multistage Interconnection Networks with Finite Buffers. | Jianxun Ding, Laxmi N. Bhuyan |
| 1991 | ICPP | Performance Analysis of Layered Task Graphs. | Hong Jiang, Laxmi N. Bhuyan |
| 1990 | ICPP | Approximate Analysis of Multiprocessing Task Graphs. | Hong Jiang, Laxmi N. Bhuyan, Dipak Ghosal |
| 1989 | ICCD | A systolic approach to multistage interconnection network design. | Chung-Han Chen, Laxmi N. Bhuyan |
| 1989 | ICPP | From Interconnection Network To Task Level Analysis. | Laxmi N. Bhuyan, Hong Jiang, Dipak Ghosal |
| 1989 | ICPP | Analysis of MIN Based Multiprocessors with Private Cache Memories. | Laxmi N. Bhuyan, Bao-Chyn Liu, Irshad Ahmed |
| 1989 | ISCA | Analysis of Computation-Communication Issues in Dynamic Dataflow Architectures. | Dipak Ghosal, Satish K. Tripathi, Laxmi N. Bhuyan, Hong Jiang |
| 1988 | ICPP | A Queueing Network Model for a Cache Coherence Protocol on Multiple-bus Multiprocessors. | Qing Yang, Laxmi N. Bhuyan |
| 1988 | INFOCOM | Design and analysis of multiple token ring networks. | C. H. Chen, Laxmi N. Bhuyan |
| 1988 | SIGMETRICS | Approximate Analysis of Task Graphs for Parallel Processing Systems. | Dipak Ghosal, Laxmi N. Bhuyan, Uday Choudhury |
| 1987 | ICPP | Performance Analysis of the MIT Tagged Token Dataflow Architecture. | Dipak Ghosal, Laxmi N. Bhuyan |
| 1987 | ICPP | Design and Analysis of a Decentralized Multiple-Bus Multiprocessor. | Qing Yang, Laxmi N. Bhuyan |
| 1987 | ISCA | Analytical Modeling and Architectural Modifications of a Dataflow Computer. | Dipak Ghosal, Laxmi N. Bhuyan |
| 1987 | RTSS | Performance Analysis of Packet-Switched Multiple-Bus Multiprocessor Systems. | Qing Yang, Laxmi N. Bhuyan, R. Pavaskar |
| 1986 | ICPP | Effect of Arbitration Policies on the Performance of Interconnection Networks. | Laxmi N. Bhuyan |
| 1986 | ICPP | Dependability Evaluation of Multicomputer Networks. | Laxmi N. Bhuyan, Chita R. Das |
| 1985 | ICPP | Reliability Simulation of Multiprocessor Systems. | Chita R. Das, Laxmi N. Bhuyan |
| 1985 | ICPP | Computation Availability of Multiple-Bus Multiprocessors. | Chita R. Das, Laxmi N. Bhuyan |
| 1984 | ISCA | On the Performance of Loosely Coupled Multiprocessors. | Laxmi N. Bhuyan |
| 1983 | ICPP | An Interference Analysis of Interconnection Networks. | Laxmi N. Bhuyan, C. W. Lee |
| 1982 | ICDCS | VLSI Performance of Multistage Interconnection Network Using 4*4 Switches. | Laxmi N. Bhuyan, Dharma P. Agrawal |
| 1982 | ICPP | Design and performance of a general class of interconnection networks. | Laxmi N. Bhuyan, Dharma P. Agrawal |
| 1982 | ISCA | A general class of processor interconnection strategies. | Laxmi N. Bhuyan, Dharma P. Agrawal |