| 2026 | ISCA | TTP: A Hardware-Efficient Design for Precise Prefetching in Ray Tracing. | Yavuz Selim Tozlu, Anshul Naithani, Huiyang Zhou |
| 2025 | HPCA | SpecMPK: Efficient In-Process Isolation with Speculative and Secure Permission Update Instruction. | Debpratim Adak, Huiyang Zhou, Eric Rotenberg, Amro Awad |
| 2025 | ISCA | Genesis: A Compiler for Hamiltonian Simulation on Hybrid CV-DV Quantum Computers. | Zihan Chen, Jiakang Li, Minghao Guo, Henry Chen, Zirui Li, Joel Bierman, Yipeng Huang, Huiyang Zhou, Yuan Liu, Eddy Z. Zhang |
| 2025 | ISCA | CoopRT: Accelerating BVH Traversal for Ray Tracing via Cooperative Threads. | Yavuz Selim Tozlu, Huiyang Zhou |
| 2025 | MICRO | CryptoBTB: A Secure Hierarchical BTB for Diverse Instruction Footprint Workloads. | Debpratim Adak, Eric Rotenberg, Amro Awad, Huiyang Zhou |
| 2025 | QCE | Benchmarking Fidelity Metrics of Quantum Computers. | Shubdeep Mohapatra, Hrushikesh Pramod Patil, Ji Liu, Huiyang Zhou |
| 2025 | QCE | Q-Cluster: Quantum Error Mitigation Through Noise-Aware Unsupervised Learning. | Hrushikesh Pramod Patil, Dror Baron, Huiyang Zhou |
| 2024 | HPCA | Salus: Efficient Security Support for CXL-Expanded GPU Memory. | Rahaf Abdullah, Hyokeun Lee, Huiyang Zhou, Amro Awad |
| 2024 | ISCA | Tetris: A Compilation Framework for VQA Applications in Quantum Computing. | Yuwei Jin, Zirui Li, Fei Hua, Tianyi Hao, Huiyang Zhou, Yipeng Huang, Eddy Z. Zhang |
| 2024 | ISCA | QuTracer: Mitigating Quantum Gate and Measurement Errors by Tracing Subsets of Qubits. | Peiyi Li, Ji Liu, Alvin Gonzales, Zain H. Saleem, Huiyang Zhou, Paul D. Hovland |
| 2024 | ISPASS | SEFsim: A Statistically-Guided Fast DRAM Simulator. | Debpratim Adak, Hyokeun Lee, Ben Feinberg, Gwendolyn Voskuilen, Clayton Hughes, Huiyang Zhou, Amro Awad |
| 2024 | QCE | Qubit-Wise Majority Vote: Maximum Likelihood Quantum Error Mitigation for Algorithms with a Single Correct Output. | Dror Baron, Hrushikesh Pramod Patil, Huiyang Zhou |
| 2024 | QCE | Posters Program: 2024 IEEE International Conference on Quantum Computing and Engineering. | Fan Chen, Huiyang Zhou |
| 2024 | QCE | Dynamic Runtime Assertions in Quantum Ternary Systems. | Ehsan Faghih, Huiyang Zhou |
| 2024 | QCE | Error Mitigation of Hamiltonian Simulations from an Analog-Based Compiler (SimuQ). | Amey Meher, Yuan Liu, Huiyang Zhou |
| 2024 | QCE | Understanding Error Sensitivity of Quantum Circuits. | Shubdeep Mohapatra, Huiyang Zhou |
| 2023 | HPCA | Plutus: Bandwidth-Efficient Memory Security for GPUs. | Rahaf Abdullah, Huiyang Zhou, Amro Awad |
| 2023 | HPCA | SecPB: Architectures for Secure Non-Volatile Memory with Battery-Backed Persist Buffers. | Alexander Freij, Huiyang Zhou, Yan Solihin |
| 2023 | ICCD | Enhancing Virtual Distillation with Circuit Cutting for Quantum Error Mitigation. | Peiyi Li, Ji Liu, Hrushikesh Pramod Patil, Paul D. Hovland, Huiyang Zhou |
| 2023 | QCE | Folding-Free ZNE: A Comprehensive Quantum Zero-Noise Extrapolation Approach for Mitigating Depolarizing and Decoherence Noise. | Hrushikesh Pramod Patil, Peiyi Li, Ji Liu, Huiyang Zhou |
| 2022 | HPCA | Not All SWAPs Have the Same Cost: A Case for Optimization-Aware Qubit Routing. | Ji Liu, Peiyi Li, Huiyang Zhou |
| 2022 | HPCA | Adaptive Security Support for Heterogeneous Memory on GPUs. | Shougang Yuan, Amro Awad, Ardhi Wiratama Baskara Yudha, Yan Solihin, Huiyang Zhou |
| 2022 | ICCD | Exploiting Quantum Assertions for Error Mitigation and Quantum Program Debugging. | Peiyi Li, Ji Liu, Yangjia Li, Huiyang Zhou |
| 2022 | ICS | LITE: a low-cost practical inter-operable GPU TEE. | Ardhi Wiratama Baskara Yudha, Jake Meyer, Shougang Yuan, Huiyang Zhou, Yan Solihin |
| 2021 | CGO | Relaxed Peephole Optimization: A Novel Compiler Optimization for Quantum Circuits. | Ji Liu, Luciano Bello, Huiyang Zhou |
| 2021 | HiPC | PILOT: a Runtime System to Manage Multi-tenant GPU Unified Memory Footprint. | John Ravi, Tri Nguyen, Huiyang Zhou, Michela Becchi |
| 2021 | HPCA | Systematic Approaches for Precise and Approximate Quantum State Runtime Assertion. | Ji Liu, Huiyang Zhou |
| 2021 | ICS | PSSM: achieving secure memory for GPUs with partitioned and sectored security metadata. | Shougang Yuan, Yan Solihin, Huiyang Zhou |
| 2021 | ISPASS | Analyzing Secure Memory Architecture for GPUs. | Shougang Yuan, Ardhi Wiratama Baskara Yudha, Yan Solihin, Huiyang Zhou |
| 2021 | MICRO | Bonsai Merkle Forests: Efficiently Achieving Crash Consistency in Secure Persistent Memory. | Alexander Freij, Huiyang Zhou, Yan Solihin |
| 2020 | ASPLOS | Quantum Circuits for Dynamic Runtime Assertions in Quantum Computation. | Ji Liu, Gregory T. Byrd, Huiyang Zhou |
| 2020 | ICS | MKPipe: a compiler framework for optimizing multi-kernel workloads in OpenCL for FPGA. | Ji Liu, Abdullah-Al Kafi, Xipeng Shen, Huiyang Zhou |
| 2020 | MICRO | Persist Level Parallelism: Streamlining Integrity Tree Updates for Secure Persistent Memory. | Alexander Freij, Shougang Yuan, Huiyang Zhou, Yan Solihin |
| 2019 | ASPLOS | Scatter-and-Gather Revisited: High-Performance Side-Channel-Resistant AES on GPUs. | Zhen Lin, Utkarsh Mathur, Huiyang Zhou |
| 2018 | HPCA | Accelerate GPU Concurrent Kernel Execution by Mitigating Memory Pipeline Stalls. | Hongwen Dai, Zhen Lin, Chao Li, Chen Zhao, Fei Wang, Nanning Zheng, Huiyang Zhou |
| 2017 | DAC | Developing Dynamic Profiling and Debugging Support in OpenCL for FPGAs. | Anshuman Verma, Huiyang Zhou, Skip Booth, Robbie King, James Coole, Andy Keep, John Marshall, Wu-chun Feng |
| 2017 | PPoPP | EffiSha: A Software Framework for Enabling Effficient Preemptive Scheduling of GPU. | Guoyang Chen, Yue Zhao, Xipeng Shen, Huiyang Zhou |
| 2016 | DAC | A model-driven approach to warp/thread-block level GPU cache bypassing. | Hongwen Dai, Chao Li, Huiyang Zhou, Saurabh Gupta, Christos Kartsaklis, Mike Mantor |
| 2016 | ICCD | Tuning Stencil codes in OpenCL for FPGAs. | Qi Jia, Huiyang Zhou |
| 2016 | ICPADS | Selectively GPU Cache Bypassing for Un-Coalesced Loads. | Chen Zhao, Fei Wang, Zhen Lin, Huiyang Zhou, Nanning Zheng |
| 2016 | SC | Enabling efficient preemption for SIMT architectures with lightweight context switching. | Zhen Lin, Lars Nyland, Huiyang Zhou |
| 2016 | SC | Optimizing memory efficiency for deep convolutional neural networks on GPUs. | Chao Li, Yi Yang, Min Feng, Srimat T. Chakradhar, Huiyang Zhou |
| 2015 | CCGRID | Revisiting ILP Designs for Throughput-Oriented GPGPU Architecture. | Ping Xiang, Yi Yang, Mike Mantor, Norm Rubin, Huiyang Zhou |
| 2015 | CGO | Automatic data placement into GPU on-chip memory resources. | Chao Li, Yi Yang, Zhen Lin, Huiyang Zhou |
| 2015 | ICPP | Spatial Locality-Aware Cache Partitioning for Effective Cache Sharing. | Saurabh Gupta, Huiyang Zhou |
| 2015 | ICS | Locality-Driven Dynamic GPU Cache Bypassing. | Chao Li, Shuaiwen Leon Song, Hongwen Dai, Albert Sidelnik, Siva Kumar Sastry Hari, Huiyang Zhou |
| 2015 | ISPASS | Analyzing graphics processor unit (GPU) instruction set architectures. | Kothiya Mayank, Hongwen Dai, Jizeng Wei, Huiyang Zhou |
| 2014 | HPCA | Warp-level divergence in GPUs: Characterization, impact, and mitigation. | Ping Xiang, Yi Yang, Huiyang Zhou |
| 2014 | ISPASS | Understanding the tradeoffs between software-managed vs. hardware-managed caches in GPUs. | Chao Li, Yi Yang, Hongwen Dai, Shengen Yan, Frank Mueller, Huiyang Zhou |
| 2014 | PPoPP | CUDA-NP: realizing nested thread-level parallelism in GPGPU applications. | Yi Yang, Huiyang Zhou |
| 2014 | PPoPP | yaSpMV: yet another SpMV framework on GPUs. | Shengen Yan, Chao Li, Yunquan Zhang, Huiyang Zhou |
| 2014 | SBAC-PAD | RACB: Resource Aware Cache Bypass on GPUs. | Hongwen Dai, Christos Kartsaklis, Chao Li, Tomislav Janjusic, Huiyang Zhou |
| 2013 | ICS | Exploiting uniform vector instructions for GPGPU performance, energy efficiency, and opportunistic reliability enhancement. | Ping Xiang, Yi Yang, Mike Mantor, Norm Rubin, Lisa R. Hsu, Huiyang Zhou |
| 2013 | PLDI | Analyzing locality of memory references in GPU architectures. | Saurabh Gupta, Ping Xiang, Huiyang Zhou |
| 2012 | HPCA | CPU-assisted GPGPU on fused CPU-GPU architectures. | Yi Yang, Ping Xiang, Mike Mantor, Huiyang Zhou |
| 2012 | ICPP | Fixing Performance Bugs: An Empirical Study of Open-Source GPGPU Programs. | Yi Yang, Ping Xiang, Mike Mantor, Huiyang Zhou |
| 2010 | ASPLOS | Accelerating MATLAB Image Processing Toolbox functions on GPUs. | Jingfei Kong, Martin Dimitrov, Yi Yang, Janaka Liyanage, Lin Cao, Jacob Staples, Mike Mantor, Huiyang Zhou |
| 2010 | DSN | Improving privacy and lifetime of PCM-based main memory. | Jingfei Kong, Huiyang Zhou |
| 2010 | PLDI | A GPGPU compiler for memory optimization and parallelism management. | Yi Yang, Ping Xiang, Jingfei Kong, Huiyang Zhou |
| 2010 | PPoPP | An optimizing compiler for GPGPU programs with input-data sharing. | Yi Yang, Ping Xiang, Jingfei Kong, Huiyang Zhou |
| 2009 | ASPLOS | Understanding software approaches for GPGPU reliability. | Martin Dimitrov, Mike Mantor, Huiyang Zhou |
| 2009 | ASPLOS | Anomaly-based bug prediction, isolation, and validation: an automated approach for software debugging. | Martin Dimitrov, Huiyang Zhou |
| 2009 | HPCA | Hardware-software integrated approaches to defend against software cache-based side channel attacks. | Jingfei Kong, Onur Aciimez, Jean-Pierre Seifert, Huiyang Zhou |
| 2008 | CCS | Deconstructing new cache designs for thwarting software cache-based side channel attacks. | Jingfei Kong, Onur Aciimez, Jean-Pierre Seifert, Huiyang Zhou |
| 2008 | HPCA | Address-branch correlation: A novel locality for long-latency hard-to-predict branches. | Hongliang Gao, Yi Ma, Martin Dimitrov, Huiyang Zhou |
| 2006 | ASPLOS | Improving software security via runtime instruction-level taint checking. | Jingfei Kong, Cliff Changchun Zou, Huiyang Zhou |
| 2006 | ICCD | Efficient Transient-Fault Tolerance for Multithreaded Processors using Dual-Thread Execution. | Yi Ma, Huiyang Zhou |
| 2003 | ISCA | Detecting Global Stride Locality in Value Streams. | Huiyang Zhou, Jill Flanagan, Thomas M. Conte |
| 2003 | ICS | Enhancing memory level parallelism via recovery-free value prediction. | Huiyang Zhou, Thomas M. Conte |