David W. Nellans
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
28
Venues
9
Active years
2010–2023
Best venue rank
A*
Where they publish
Papers
28 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | CGO | Parsimony: Enabling SIMD/Vector Programming in Standard Compiler Flows. | Vijay Kandiah, Daniel Lustig, Oreste Villa, David W. Nellans, Nikos Hardavellas |
| 2023 | HPCA | FinePack: Transparently Improving the Efficiency of Fine-Grained Transfers in Multi-GPU Systems. | Harini Muthukrishnan, Daniel Lustig, Oreste Villa, Thomas F. Wenisch, David W. Nellans |
| 2023 | MICRO | Architectural Support for Optimizing Huge Page Selection Within the OS. | Aninda Manocha, Zi Yan, Esin Tureci, Juan L. Aragn, David W. Nellans, Margaret Martonosi |
| 2021 | HPCA | Need for Speed: Experiences Building a Trustworthy System-Level GPU Simulator. | Oreste Villa, Daniel Lustig, Zi Yan, Evgeny Bolotin, Yaosheng Fu, Niladrish Chatterjee, Nan Jiang, David W. Nellans |
| 2021 | ISCA | Efficient Multi-GPU Shared Memory via Automatic Optimization of Fine-Grained Transfers. | Harini Muthukrishnan, David W. Nellans, Daniel Lustig, Jeffrey A. Fessler, Thomas F. Wenisch |
| 2021 | MICRO | GPS: A Global Publish-Subscribe Model for Multi-GPU Memory Management. | Harini Muthukrishnan, Daniel Lustig, David W. Nellans, Thomas F. Wenisch |
| 2020 | HPCA | HMG: Extending Cache Coherence Protocols Across Modern Hierarchical Multi-GPU Systems. | Xiaowei Ren, Daniel Lustig, Evgeny Bolotin, Aamer Jaleel, Oreste Villa, David W. Nellans |
| 2020 | ISCA | Buddy Compression: Enabling Larger Memory for Deep Learning and HPC Workloads on GPUs. | Esha Choukse, Michael B. Sullivan, Mike O'Connor, Mattan Erez, Jeff Pool, David W. Nellans, Stephen W. Keckler |
| 2020 | MICRO | Locality-Centric Data and Threadblock Management for Massive GPUs. | Mahmoud Khairy, Vadim Nikiforov, David W. Nellans, Timothy G. Rogers |
| 2019 | ASPLOS | Nimble Page Management for Tiered Memory Systems. | Zi Yan, Daniel Lustig, David W. Nellans, Abhishek Bhattacharjee |
| 2019 | HPCA | Understanding the Future of Energy Efficiency in Multi-Module GPUs. | Akhil Arunkumar, Evgeny Bolotin, David W. Nellans, Carole-Jean Wu |
| 2019 | ISCA | Translation ranger: operating system support for contiguity-aware TLBs. | Zi Yan, Daniel Lustig, David W. Nellans, Abhishek Bhattacharjee |
| 2019 | MICRO | NVBit: A Dynamic Binary Instrumentation Framework for NVIDIA GPUs. | Oreste Villa, Mark Stephenson, David W. Nellans, Stephen W. Keckler |
| 2018 | MICRO | Combining HW/SW Mechanisms to Improve NUMA Performance of Multi-GPU Systems. | Vinson Young, Aamer Jaleel, Evgeny Bolotin, Eiman Ebrahimi, David W. Nellans, Oreste Villa |
| 2017 | ISCA | MCM-GPU: Multi-Chip-Module GPUs for Continued Performance Scalability. | Akhil Arunkumar, Evgeny Bolotin, Benjamin Y. Cho, Ugljesa Milic, Eiman Ebrahimi, Oreste Villa, Aamer Jaleel, Carole-Jean Wu, David W. Nellans |
| 2017 | MICRO | Beyond the socket: NUMA-aware GPUs. | Ugljesa Milic, Oreste Villa, Evgeny Bolotin, Akhil Arunkumar, Eiman Ebrahimi, Aamer Jaleel, Alex Ramrez, David W. Nellans |
| 2016 | HPCA | Selective GPU caches to eliminate CPU-GPU HW cache coherence. | Neha Agarwal, David W. Nellans, Eiman Ebrahimi, Thomas F. Wenisch, John Danskin, Stephen W. Keckler |
| 2016 | HPCA | Towards high performance paged memory for GPUs. | Tianhao Zheng, David W. Nellans, Arslan Zulfiqar, Mark Stephenson, Stephen W. Keckler |
| 2015 | ASPLOS | Page Placement Strategies for GPUs within Heterogeneous Memory Systems. | Neha Agarwal, David W. Nellans, Mark Stephenson, Mike O'Connor, Stephen W. Keckler |
| 2015 | HPCA | Unlocking bandwidth for GPUs in CC-NUMA systems. | Neha Agarwal, David W. Nellans, Mike O'Connor, Stephen W. Keckler, Thomas F. Wenisch |
| 2015 | ISCA | Flexible software profiling of GPU architectures. | Mark Stephenson, Siva Kumar Sastry Hari, Yunsup Lee, Eiman Ebrahimi, Daniel R. Johnson, David W. Nellans, Mike O'Connor, Stephen W. Keckler |
| 2014 | SC | Scaling the Power Wall: A Path to Exascale. | Oreste Villa, Daniel R. Johnson, Mike O'Connor, Evgeny Bolotin, David W. Nellans, Justin Luitjens, Nikolai Sakharnykh, Peng Wang, Paulius Micikevicius, Anthony Scudiero, Stephen W. Keckler, William J. Dally |
| 2013 | SOSP | Better flash access via shape-shifting virtual memory pages. | Anirudh Badam, Vivek S. Pai, David W. Nellans |
| 2013 | SYSTOR | Linux block IO: introducing multi-queue SSD access on multi-core systems. | Matias Bjrling, Jens Axboe, David W. Nellans, Philippe Bonnet |
| 2011 | HPCA | Beyond block I/O: Rethinking traditional storage primitives. | Xiangyong Ouyang, David W. Nellans, Robert Wipfel, David Flynn, Dhabaleswar K. Panda |
| 2010 | ASPLOS | Micro-pages: increasing DRAM efficiency with locality-aware data placement. | Kshitij Sudan, Niladrish Chatterjee, David W. Nellans, Manu Awasthi, Rajeev Balasubramonian, Al Davis |
| 2010 | ISCA | Improving Server Performance on Multi-cores via Selective Off-Loading of OS Functionality. | David W. Nellans, Kshitij Sudan, Erik Brunvand, Rajeev Balasubramonian |
| 2010 | ISPASS | Hardware prediction of OS run-length for fine-grained resource customization. | David W. Nellans, Kshitij Sudan, Rajeev Balasubramonian, Erik Brunvand |