| 2025 | BIBE | A Graph-Based Approach for Early and Explainable Health Risk Assessment. | Sonal Jha, Wu-chun Feng |
| 2025 | HPDC | Top-Down SBP: Turning Graph Clustering Upside Down. | Frank Wanye, Vitaliy Gleyzer, Edward K. Kao, Wu-chun Feng |
| 2025 | Networking | STAGS: A Graph-Sampling Approach for GNN-based Network Anomaly Detection. | Saikat Dey, Mark K. Gardner, Jeffry Lang, Wu-chun Feng |
| 2024 | CLUSTER | Welcome Message from the IEEE Cluster 2024 Program Chairs. | Yutong Lu, Wu-chun Feng, Mohamed Wahib |
| 2022 | ICPP | On the Parallelization of MCMC for Community Detection. | Frank Wanye, Vitaliy Gleyzer, Edward K. Kao, Wu-chun Feng |
| 2021 | SC | Mitigating Catastrophic Forgetting in Deep Learning in a Streaming Setting Using Historical Summary. | Sajal Dash, Junqi Yin, Mallikarjun Shankar, Feiyi Wang, Wu-chun Feng |
| 2020 | CCGRID | Alleviating Load Imbalance in Data Processing for Large-Scale Deep Learning. | Sarunya Pumma, Daniele Buono, Fabio Checconi, Xinyu Que, Wu-chun Feng |
| 2020 | CCGRID | SparkLeBLAST: Scalable Parallelization of BLAST Sequence Alignment Using Spark. | Karim Youssef, Wu-chun Feng |
| 2018 | HPDC | CommAnalyzer: automated estimation of communication cost and scalability on HPC clusters from sequential code. | Ahmed E. Helal, Changhee Jung, Wu-chun Feng, Yasser Y. Hanafy |
| 2018 | ICPP | A Framework for Auto-Parallelization and Code Generation: An Integrative Case Study with Legacy FORTRAN Codes. | Konstantinos Krommydas, Paul Sathre, Ruchira Sasanka, Wu-chun Feng |
| 2017 | DAC | Developing Dynamic Profiling and Debugging Support in OpenCL for FPGAs. | Anshuman Verma, Huiyang Zhou, Skip Booth, Robbie King, James Coole, Andy Keep, John Marshall, Wu-chun Feng |
| 2017 | HPCC | Portable Parallel Design of Weighted Multi-Dimensional Scaling for Real-Time Data Analysis. | Sajal Dash, Anshuman Verma, Chris North, Wu-chun Feng |
| 2017 | HPCC | Towards Scalable Deep Learning via I/O Analysis and Optimization. | Sarunya Pumma, Min Si, Wu-chun Feng, Pavan Balaji |
| 2017 | ICPADS | Parallel I/O Optimizations for Scalable Deep Learning. | Sarunya Pumma, Min Si, Wu-chun Feng, Pavan Balaji |
| 2017 | ICS | Fast segmented sort on GPUs. | Kaixi Hou, Weifeng Liu, Hao Wang, Wu-chun Feng |
| 2017 | ICS | Demystifying automata processing: GPUs, FPGAs or Micron's AP? | Marziyeh Nourian, Xiang Wang, Xiaodong Yu, Wu-chun Feng, Michela Becchi |
| 2016 | CCGRID | Online Power Estimation of Graphics Processing Units. | Vignesh Adhinarayanan, Balaji Subramaniam, Wu-chun Feng |
| 2016 | CCGRID | cuART: Fine-Grained Algebraic Reconstruction Technique for Computed Tomography Images on GPUs. | Xiaodong Yu, Hao Wang, Wu-chun Feng, Hao Gong, Guohua Cao |
| 2016 | FCCM | Bridging the Performance-Programmability Gap for FPGAs via OpenCL: A Case Study with OpenDwarfs. | Konstantinos Krommydas, Ahmed E. Helal, Anshuman Verma, Wu-chun Feng |
| 2016 | ICS | Parallel Transposition of Sparse Data Structures. | Hao Wang, Weifeng Liu, Kaixi Hou, Wu-chun Feng |
| 2016 | ISPASS | An automated framework for characterizing and subsetting GPGPU workloads. | Vignesh Adhinarayanan, Wu-chun Feng |
| 2016 | SC | MetaMorph: a library framework for interoperable kernels on multi- and many-core clusters. | Ahmed E. Helal, Paul Sathre, Wu-chun Feng |
| 2015 | CLUSTER | Automatic Command Queue Scheduling for Task-Parallel Workloads in OpenCL. | Ashwin Mandayam Aji, Antonio J. Pea, Pavan Balaji, Wu-chun Feng |
| 2015 | ICPP | GLAF: A Visual Programming and Auto-tuning Framework for Parallel Computing. | Konstantinos Krommydas, Ruchira Sasanka, Wu-chun Feng |
| 2015 | INFOCOM | Rapid and parallel content screening for detecting transformed data exposure. | Xiaokui Shu, Jing Zhang, Danfeng Yao, Wu-chun Feng |
| 2015 | ICS | ASPaS: A Framework for Automatic SIMDization of Parallel Sorting on x86-based Many-core Processors. | Kaixi Hou, Hao Wang, Wu-chun Feng |
| 2014 | CCGRID | Runtime Adaptation for Autonomic Heterogeneous Computing. | Thomas R. W. Scogland, Wu-chun Feng |
| 2014 | CCGRID | Enabling Efficient Power Provisioning for Enterprise Applications. | Balaji Subramaniam, Wu-chun Feng |
| 2014 | HPDC | SLAM: scalable locality-aware middleware for I/O in scientific analysis and visualization. | Jiangling Yin, Jun Wang, Wu-chun Feng, Xuhong Zhang, Junyao Zhang |
| 2014 | SC | On the Energy Proportionality of Distributed NoSQL Data Stores. | Balaji Subramaniam, Wu-chun Feng |
| 2014 | TrustCom | Aeromancer: A Workflow Manager for Large-Scale MapReduce-Based Scientific Workflows. | Mohamed Nabeel, Nabanita Maji, Jing Zhang, Nataliya Timoshevskaya, Wu-chun Feng |
| 2013 | CCGRID | Optimizing Burrows-Wheeler Transform-Based Sequence Alignment on Multicore Architectures. | Jing Zhang, Heshan Lin, Pavan Balaji, Wu-chun Feng |
| 2013 | GLOBECOM | Cascaded TCP: Applying pipelining to TCP for efficient communication over wide-area networks. | Umar Kalim, Mark K. Gardner, Eric J. Brown, Wu-chun Feng |
| 2013 | HPCC | Consolidating Applications for Energy Efficiency in Heterogeneous Computing Systems. | Jing Zhang, Hao Wang, Heshan Lin, Wu-chun Feng |
| 2013 | HPDC | On the efficacy of GPU-integrated MPI for scientific applications. | Ashwin M. Aji, Lokendra S. Panwar, Feng Ji, Milind Chabbi, Karthik Murthy, Pavan Balaji, Keith R. Bisset, James Dinan, Wu-chun Feng, John M. Mellor-Crummey, Xiaosong Ma, Rajeev Thakur |
| 2013 | ICCCN | Seamless Migration of Virtual Machines across Networks. | Umar Kalim, Mark K. Gardner, Eric J. Brown, Wu-chun Feng |
| 2013 | ICDCS | pVOCL: Power-Aware Dynamic Placement and Migration in Virtualized GPU Environments. | Palden Lama, Yan Li, Ashwin M. Aji, Pavan Balaji, James Dinan, Shucai Xiao, Yunquan Zhang, Wu-chun Feng, Rajeev Thakur, Xiaobo Zhou |
| 2013 | ICPADS | Wideband Channelization for Software-Defined Radio via Mobile Graphics Processors. | Vignesh Adhinarayanan, Wu-chun Feng |
| 2013 | ICPADS | On the Portability of the OpenCL Dwarfs on Fixed and Reconfigurable Parallel Platforms. | Konstantinos Krommydas, Muhsen Owaida, Christos D. Antonopoulos, Nikolaos Bellas, Wu-chun Feng |
| 2013 | ICPADS | On the Programmability and Performance of Heterogeneous Platforms. | Konstantinos Krommydas, Thomas R. W. Scogland, Wu-chun Feng |
| 2013 | ICPADS | Online Performance Projection for Clusters with Heterogeneous GPUs. | Lokendra S. Panwar, Ashwin M. Aji, Jiayuan Meng, Pavan Balaji, Wu-chun Feng |
| 2013 | SC | SDAFT: a novel scalable data access framework for parallel BLAST. | Jiangling Yin, Junyao Zhang, Jun Wang, Wu-chun Feng |
| 2012 | CCGRID | Transparent Accelerator Migration in a Virtualized GPU Environment. | Shucai Xiao, Pavan Balaji, James Dinan, Qian Zhu, Rajeev Thakur, Susan Coghlan, Heshan Lin, Gaojin Wen, Jue Hong, Wu-chun Feng |
| 2012 | HPCC | MPI-ACC: An Integrated and Extensible Approach to Data Movement in Accelerator-based Systems. | Ashwin M. Aji, James Dinan, Darius Buntinas, Pavan Balaji, Wu-chun Feng, Keith R. Bisset, Rajeev Thakur |
| 2012 | HPCC | DMA-Assisted, Intranode Communication in GPU Accelerated Systems. | Feng Ji, Ashwin M. Aji, James Dinan, Darius Buntinas, Pavan Balaji, Rajeev Thakur, Wu-chun Feng, Xiaosong Ma |
| 2012 | SC | Abstract: Cascaded TCP: BIG Throughput for BIG DATA Applications in Distributed HPC. | Umar Kalim, Mark K. Gardner, Eric J. Brown, Wu-chun Feng |
| 2012 | SC | Poster: Cascaded TCP: BIG Throughput for BIG DATA Applications in Distributed HPC. | Umar Kalim, Mark K. Gardner, Eric J. Brown, Wu-chun Feng |
| 2011 | CLUSTER | Performance Characterization and Optimization of Atomic Operations on AMD GPUs. | Marwa K. Elteir, Heshan Lin, Wu-chun Feng |
| 2011 | HPDC | Energy-efficient E-puting everywhere. | Wu-chun Feng |
| 2011 | ICCCN | Restoring End-to-End Resilience in the Presence of Middleboxes. | Eric J. Brown, Mark K. Gardner, Umar Kalim, Wu-chun Feng |
| 2011 | ICPADS | Architecture-Aware Mapping and Optimization on a 1600-Core GPU. | Mayank Daga, Thomas Scogland, Wu-chun Feng |
| 2011 | ICPADS | StreamMR: An Optimized MapReduce Framework for AMD GPUs. | Marwa K. Elteir, Heshan Lin, Wu-chun Feng, Tom Scogland |
| 2011 | ICPADS | CU2CL: A CUDA-to-OpenCL Translator for Multi- and Many-Core Architectures. | Gabriel Martinez, Mark K. Gardner, Wu-chun Feng |
| 2011 | ICPADS | Optimizing Dynamic Programming on Graphics Processing Units via Adaptive Thread-Level Parallelism. | Chao-Chin Wu, Jenn-Yang Ke, Heshan Lin, Wu-chun Feng |
| 2011 | SC | Poster: characterizing the impact of memory-access techniques on AMD fusion. | Kenneth S. Lee, Heshan Lin, Wu-chun Feng |
| 2010 | HPDC | MOON: MapReduce On Opportunistic eNvironments. | Heshan Lin, Xiaosong Ma, Jeremy S. Archuleta, Wu-chun Feng, Mark K. Gardner, Zhe Zhang |
| 2010 | ICCCN | On the Goodput of TCP NewReno in Mobile Networks. | Sushant Sharma, Donald W. Gillies, Wu-chun Feng |
| 2010 | ICPADS | Enhancing MapReduce via Asynchronous Data Processing. | Marwa K. Elteir, Heshan Lin, Wu-chun Feng |
| 2010 | ISCAS | To GPU synchronize or not GPU synchronize? | Wu-chun Feng, Shucai Xiao |
| 2010 | ITiCSE | Broadening accessibility to computer science for K-12 education. | Mark K. Gardner, Wu-chun Feng |
| 2009 | ICPADS | On the Robust Mapping of Dynamic Programming onto a Graphics Processing Unit. | Shucai Xiao, Ashwin M. Aji, Wu-chun Feng |
| 2009 | ICPP | GePSeA: A General-Purpose Software Acceleration Framework for Lightweight Task Offloading. | Ajeet Singh, Pavan Balaji, Wu-chun Feng |
| 2008 | BIBE | Optimizing performance, cost, and sensitivity in pairwise sequence search on a cluster of PlayStations. | Ashwin M. Aji, Wu-chun Feng |
| 2008 | HiPC | Making a Case for Proactive Flow Control in Optical Circuit-Switched Networks. | Mithilesh Kumar, Vineeta Chaube, Pavan Balaji, Wu-chun Feng, Hyun-Wook Jin |
| 2008 | HPDC | Semantic-based distributed i/o with the paramedic framework. | Pavan Balaji, Wu-chun Feng, Heshan Lin |
| 2008 | ICCCN | Impact of Network Sharing in Multi-Core Architectures. | Ganesh Narayanaswamy, Pavan Balaji, Wu-chun Feng |
| 2008 | PPoPP | Semantics-based distributed I/O for mpiBLAST. | Pavan Balaji, Wu-chun Feng, Jeremy S. Archuleta, Heshan Lin, Rajkumar Kettimuthu, Rajeev Thakur, Xiaosong Ma |
| 2008 | SC | Massively parallel genomic sequence search on the Blue Gene/P architecture. | Heshan Lin, Pavan Balaji, Ruth Poole, Carlos P. Sosa, Xiaosong Ma, Wu-chun Feng |
| 2008 | SC | Asymmetric interactions in symmetric multi-core systems: analysis, enhancements and evaluation. | Thomas Scogland, Pavan Balaji, Wu-chun Feng, Ganesh Narayanaswamy |
| 2007 | HOTI | An Analysis of 10-Gigabit Ethernet Protocol Stacks in Multicore Environments. | Ganesh Narayanaswamy, Pavan Balaji, Wu-chun Feng |
| 2007 | ICPP | CPU MISER: A Performance-Directed, Run-Time System for Power-Aware Clusters. | Rong Ge, Xizhou Feng, Wu-chun Feng, Kirk W. Cameron |
| 2007 | SC | Analyzing the impact of supporting out-of-order communication on in-order performance with iWARP. | Pavan Balaji, Wu-chun Feng, Sitha Bhagvat, Dhabaleswar K. Panda, Rajeev Thakur, William Gropp |
| 2006 | CCGRID | A Feedback Mechanism for Network Scheduling in LambdaGrids. | Pallab Datta, Sushant Sharma, Wu-chun Feng |
| 2006 | HPDC | Exploring I/O Strategies for Parallel Sequence-Search Tools with S3aSim. | Avery Ching, Wu-chun Feng, Heshan Lin, Xiaosong Ma, Alok N. Choudhary |
| 2006 | ICCCN | When Optical Networking Meets Grid Computing? | Wu-chun Feng, Mark K. Gardner, Gigi Karmous-Edwards, Jerry Sobieski, Malathi Veeraraghavan |
| 2006 | SC | Grid networks and portals - End-system aware, rate-adaptive protocol for network transport in LambdaGrid environments. | Pallab Datta, Wu-chun Feng, Sushant Sharma |
| 2006 | SC | Grid applications - Parallel genomic sequence-searching on an ad-hoc grid: experiences, lessons learned, and implications. | Mark K. Gardner, Wu-chun Feng, Jeremy S. Archuleta, Heshan Lin, Xiaosong Ma |
| 2005 | CLUSTER | Head-to-TOE Evaluation of High-Performance Sockets over Protocol Offload Engines. | Pavan Balaji, Wu-chun Feng, Qi Gao, Ranjit Noronha, Weikuan Yu, Dhabaleswar K. Panda |
| 2005 | CLUSTER | A Feasibility Analysis of Power Awareness in Commodity-Based High-Performance Clusters. | Chung-Hsing Hsu, Wu-chun Feng |
| 2005 | HOTI | Performance Characterization of a 10-Gigabit Ethernet TOE. | Wu-chun Feng, Pavan Balaji, Christopher Baron, Laxmi N. Bhuyan, Dhabaleswar K. Panda |
| 2005 | INFOCOM | Q-Composer and CpR: a probabilistic synthesizer and regulator of traffic (a probabilistic control of buffer occupancy). | Sami Ayyorgun, Sarut Vanichpun, Wu-chun Feng |
| 2005 | SC | A Power-Aware Run-Time System for High-Performance Computing. | Chung-Hsing Hsu, Wu-chun Feng |
| 2004 | CBMS | A Multimodal Interface for the Immediate Transcription of Radiology Dictation. | Wu-chun Feng |
| 2004 | ICCCN | A Systematic Approach for Providing End-to-End Probabilistic QoS Guarantees. | Sami Ayyorgun, Wu-chun Feng |
| 2003 | CCGRID | MAGNET: A Tool for Debugging, Analyzing and Adapting Computing Systems. | Mark K. Gardner, Wu-chun Feng, Michael Broxton, Adam Engelhart, Justin Gus Hurwitz |
| 2003 | HOTI | Initial end-to-end performance evaluation of 10-Gigabit Ethernet. | Justin Gus Hurwitz, Wu-chun Feng |
| 2003 | HPDC | Optimizing GridFTP through Dynamic Right-Sizing. | Sunil Thulasidasan, Wu-chun Feng, Mark K. Gardner |
| 2003 | SC | Optimizing 10-Gigabit Ethernet for Networks of Workstations, Clusters, and Grids: A Case Study. | Wu-chun Feng, Justin Gus Hurwitz, Harvey B. Newman, Sylvain Ravot, Roger Les Cottrell, Olivier Martin, Fabrizio Coccetti, Cheng Jin, David X. Wei, Steven H. Low |
| 2002 | CLUSTER | The Bladed Beowulf: A Cost-Effective Alternative to Traditional Beowulf. | Wu-chun Feng, Michael S. Warren, Eric Weigle |
| 2002 | GLOBECOM | GREEN: proactive queue management over a best-effort network. | Wu-chun Feng, Apu Kapadia, Sunil Thulasidasan |
| 2002 | HPDC | Dynamic Right-Sizing in FTP (drsFTP): Enhancing Grid Performance in User-Space. | Mark K. Gardner, Wu-chun Feng, Mike Fisk |
| 2002 | HPDC | A Comparison of TCP Automatic Tuning Techniques for Distributed Computing. | Eric Weigle, Wu-chun Feng |
| 2002 | ICCCN | On the transient behavior of TCP Vegas. | Sarut Vanichpun, Wu-chun Feng |
| 2002 | ICPP | Honey, I Shrunk the Beowulf! | Wu-chun Feng, Michael S. Warren, Eric Weigle |
| 2002 | SC | High-density computing: a 240-processor Beowulf in one cubic meter. | Michael S. Warren, Eric Weigle, Wu-chun Feng |
| 2001 | HOTI | The Quadrics network (QsNet): high-performance clustering technology. | Fabrizio Petrini, Wu-chun Feng, Adolfy Hoisie, Salvador Coll, Eitan Frachtenberg |
| 2001 | HPDC | A Case for TCP Vegas in High-Performance Computational Grids. | Eric Weigle, Wu-chun Feng |
| 2001 | ICCCN | MAGNeT: monitor for application-generated network traffic. | Wu-chun Feng, Jeffrey R. Hay, Mark K. Gardner |
| 2001 | ICCCN | Dynamic right-sizing: a simulation study. | Eric Weigle, Wu-chun Feng |
| 2001 | ICDCS | The Effects of Inter-packet Spacing on the Delivery of Multimedia Content. | Apu Kapadia, Annette C. Feng, Wu-chun Feng |
| 2000 | ICDCS | Scheduling with Global Information in Distributed Systems. | Fabrizio Petrini, Wu-chun Feng |
| 2000 | ICDCS | On the Burstiness of the TCP Congestion-Control Mechanism in a Distributed Computing System. | Peerapol Tinnakornsrisuphap, Wu-chun Feng, Ian R. Philp |
| 2000 | ICPP | The Adverse Impact of the TCP Congestion-Control Mechanism in Heterogeneous Computing Systems. | Wu-chun Feng, Peerapol Tinnakornsrisuphap |
| 2000 | JSSPP | Time-Sharing Parallel Jobs in the Presence of Multiple Resource Requirements. | Fabrizio Petrini, Wu-chun Feng |
| 2000 | SC | The Failure of TCP in High-Performance Computational Grids. | Wu-chun Feng, Peerapol Tinnakornsrisuphap |
| 1999 | COMPSAC | Dynamic Client-Side Scheduling in a Real-Time CORBA System. | Wu-chun Feng |