Skip to content

Zhiling Lan

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

64

Venues

15

Active years

2000–2026

Best venue rank

A*

Where they publish

Papers

64 indexed papers, newest first.

YearVenueTitleAuthors
2026AAAISMART: A Surrogate Model for Predicting Application Runtime in Dragonfly Systems.Xin Wang, Pietro Lodi Rizzini, Sourav Medya, Zhiling Lan
2026ICSSmartCap: Coordinated CPU-GPU Power Capping for Performance-Assurance Energy Efficiency.Zhong Zheng, Xingfu Wu, Valerie E. Taylor, Michael E. Papka, Zhiling Lan
2025CCGRIDEvaluating Energy Efficiency of Ai Accelerators Using Two Mlperf Benchmarks.Farah Ferdaus, Xingfu Wu, Valerie Taylor, Zhiling Lan, Sanjif Shanmugavelu, Venkatram Vishwanath, Michael E. Papka
2025JSSPPMore for Less: Integrating Capability-Predominant and Capacity-Predominant Computing.Zhong Zheng, Michael E. Papka, Zhiling Lan
2025PADSDirecting PDES and Surrogate Models in Loosely Coupled Hybrid Simulations.Kevin A. Brown, Elkin Cruz-Camacho, Kazutomo Yoshii, Xin Wang, Zhiling Lan, Christopher D. Carothers, Robert B. Ross
2025PADSCQSim+: Symbiotic Simulation for Multi-Resource Scheduling in High-Performance Computing.Yash Kurkure, Shambhawi Sharma, Xin Wang, Michael E. Papka, Zhiling Lan
2025PADSSynopsis: MFNetSim: A Multi-Fidelity Network Simulation Framework for Multi-Traffic Modeling of Dragonfly Systems.Xin Wang, Kevin A. Brown, Robert B. Ross, Christopher D. Carothers, Zhiling Lan
2025PDPOn the Effectiveness of Unified Memory in Multi-GPU Collective Communication.Riccardo Strina, Ian Di Dio Lavore, Marco D. Santambrogio, Michael E. Papka, Zhiling Lan
2025SCMinimizing Power Waste in Heterogenous Computing via Adaptive Uncore Scaling.Zhong Zheng, Seyfal Sultanov, Michael E. Papka, Zhiling Lan
2024CLUSTERLASSI: An LLM-Based Automated Self-Correcting Pipeline for Translating Parallel Scientific Codes.Matthew T. Dearing, Yiheng Tao, Xingfu Wu, Zhiling Lan, Valerie Taylor
2024PADSSurrogate Modeling for HPC Application Iteration Times Forecasting with Network Features.Xiongxiao Xu, Kevin A. Brown, Tanwi Mallick, Xin Wang, Elkin Cruz-Camacho, Robert B. Ross, Christopher D. Carothers, Zhiling Lan, Kai Shu
2023MASCOTSInterpretable Modeling of Deep Reinforcement Learning Driven Scheduling.Boyang Li, Zhiling Lan, Michael E. Papka
2023PADSHybrid PDES Simulation of HPC Networks Using Zombie Packets.Elkin Cruz-Camacho, Kevin A. Brown, Xin Wang, Xiongxiao Xu, Kai Shu, Zhiling Lan, Robert B. Ross, Christopher D. Carothers
2023PADSWorkload Interference Prevention with Intelligent Routing and Flexible Job Placement on Dragonfly.Yao Kang, Xin Wang, Zhiling Lan
2023PADSMachine Learning for Interconnect Network Traffic Forecasting: Investigation and Exploitation.Xiongxiao Xu, Xin Wang, Elkin Cruz-Camacho, Christopher D. Carothers, Kevin A. Brown, Robert B. Ross, Zhiling Lan, Kai Shu
2022CLUSTERMRSch: Multi-Resource Scheduling for HPC.Boyang Li, Yuping Fan, Matthew T. Dearing, Zhiling Lan, Paul Rich, William E. Allcock, Michael E. Papka
2022JSSPPEncoding for Reinforcement Learning Driven Scheduling.Boyang Li, Yuping Fan, Michael E. Papka, Zhiling Lan
2022SCStudy of Workload Interference with Intelligent Routing on Dragonfly.Yao Kang, Xin Wang, Zhiling Lan
2021CLUSTERA Dynamic Power Capping Library for HPC Applications.Sahil Sharma, Zhiling Lan, Xingfu Wu, Valerie Taylor
2021HPDCQ-adaptive: A Multi-Agent Reinforcement Learning Based Routing on Dragonfly Network.Yao Kang, Xin Wang, Zhiling Lan
2019HPDCScheduling Beyond CPUs for HPC.Yuping Fan, Zhiling Lan, Paul Rich, William E. Allcock, Michael E. Papka, Brian Austin, David Paul
2019PADSModeling and Analysis of Application Interference on Dragonfly+.Yao Kang, Xin Wang, Neil McGlohon, Misbah Mubarak, Sudheer Chunduri, Zhiling Lan
2017CLUSTERTrade-Off Between Prediction Accuracy and Underestimation Rate in Job Runtime Estimates.Yuping Fan, Paul Rich, William E. Allcock, Michael E. Papka, Zhiling Lan
2017CLUSTERPreliminary Interference Study About Job Placement and Routing Algorithms in the Fat-Tree Topology for HPC Applications.Peixin Qiao, Xin Wang, Xu Yang, Yuping Fan, Zhiling Lan
2017CLUSTERA Preliminary Study of Intra-Application Interference on Dragonfly Network.Xin Wang, Xu Yang, Misbah Mubarak, Robert B. Ross, Zhiling Lan
2017JSSPPExperience and Practice of Batch Scheduling on Leadership Supercomputers at Argonne.William E. Allcock, Paul Rich, Yuping Fan, Zhiling Lan
2016CLUSTERExploring Plan-Based Scheduling for Large-Scale Computing Systems.Xingwu Zhang, Zhou Zhou, Xu Yang, Zhiling Lan, Jia Wang
2016EuroParExploring Partial Replication to Improve Lightweight Silent Data Corruption Detection for HPC Applications.Eduardo Berrocal, Leonardo Bautista-Gomez, Sheng Di, Zhiling Lan, Franck Cappello
2016ICPADSStudy of Intra- and Interjob Interference on Torus Networks.Xu Yang, John Jenkins, Misbah Mubarak, Xin Wang, Robert B. Ross, Zhiling Lan
2016SCA data driven scheduling approach for power management on HPC systems.Sean Wallace, Xu Yang, Venkatram Vishwanath, William E. Allcock, Susan Coghlan, Michael E. Papka, Zhiling Lan
2016SCWatch out for the bully!: job interference study on dragonfly network.Xu Yang, John Jenkins, Misbah Mubarak, Robert B. Ross, Zhiling Lan
2015CLUSTERComparison of Vendor Supplied Environmental Data Collection Mechanisms.Sean Wallace, Venkatram Vishwanath, Susan Coghlan, Zhiling Lan, Michael E. Papka
2015CLUSTERI/O-Aware Batch Scheduling for Petascale Computing Systems.Zhou Zhou, Xu Yang, Dongfang Zhao, Paul Rich, Wei Tang, Jia Wang, Zhiling Lan
2015HPDCLightweight Silent Data Corruption Detection Based on Runtime Data Analysis for HPC Applications.Eduardo Berrocal, Leonardo Arturo Bautista-Gomez, Sheng Di, Zhiling Lan, Franck Cappello
2014CLUSTERExploring void search for fault detection on extreme scale systems.Eduardo Berrocal, Li Yu, Sean Wallace, Michael E. Papka, Zhiling Lan
2014CLUSTERBalancing job performance with system performance via locality-aware scheduling on torus-connected systems.Xu Yang, Zhou Zhou, Wei Tang, Xingwu Zheng, Jia Wang, Zhiling Lan
2013CLUSTERApplication power profiling on IBM Blue Gene/Q.Sean Wallace, Venkatram Vishwanath, Susan Coghlan, John R. Tramm, Zhiling Lan, Michael E. Papka
2013JSSPPReducing Energy Costs for IBM Blue Gene/P via Power-Aware Job Scheduling.Zhou Zhou, Zhiling Lan, Wei Tang, Narayan Desai
2013SCIntegrating dynamic pricing of electricity into energy aware scheduling for HPC systems.Xu Yang, Zhou Zhou, Sean Wallace, Zhiling Lan, Wei Tang, Susan Coghlan, Michael E. Papka
2012DSNFiltering log data: Finding the needles in the Haystack.Li Yu, Ziming Zheng, Zhiling Lan, Terry R. Jones, Jim M. Brandt, Ann C. Gentile
2012SCHierarchical task mapping of cell-based AMR cosmology simulations.Jingjin Wu, Zhiling Lan, Xuanxing Xiong, Nickolay Y. Gnedin, Andrey V. Kravtsov
2011CLUSTERPerformance Emulation of Cell-Based AMR Cosmology Simulations.Jingjin Wu, Roberto E. Gonzlez, Zhiling Lan, Nickolay Y. Gnedin, Andrey V. Kravtsov, Douglas H. Rudd, Yongen Yu
2011CLUSTEREvaluating Performance Impacts of Delayed Failure Repairing on Large-Scale Systems.Zhou Zhou, Wei Tang, Ziming Zheng, Zhiling Lan, Narayan Desai
2011DSNPractical online failure prediction for Blue Gene/P: Period-based vs event-driven.Li Yu, Ziming Zheng, Zhiling Lan, Susan Coghlan
2010DSNA practical failure prediction with location and lead time for Blue Gene/P.Ziming Zheng, Zhiling Lan, Rinku Gupta, Susan Coghlan, Peter H. Beckman
2010SCAutomatic and coordinated job recovery for high performance computing.Wei Tang, Zhiling Lan, Narayan Desai, Daniel Buettner
2009CCGRIDPerformance under Failures of DAG-based Parallel Computing.Hui Jin, Xian-He Sun, Ziming Zheng, Zhiling Lan, Bing Xie
2009CLUSTERFault-aware, utility-based job scheduling on Blue, Gene/P systems.Wei Tang, Zhiling Lan, Narayan Desai, Daniel Buettner
2009CLUSTERReliability-aware scalability models for high performance computing.Ziming Zheng, Zhiling Lan
2009DSNSystem log pre-processing to improve failure prediction.Ziming Zheng, Zhiling Lan, Byung-Hoon Park, Al Geist
2008DSNA fast restart mechanism for checkpoint/recovery protocols in networked environments.Yawei Li, Zhiling Lan
2008ICPPDynamic Meta-Learning for Failure Prediction in Large-Scale Systems: A Case Study.Jiexing Gu, Ziming Zheng, Zhiling Lan, John White, Eva Hocks, Byung-Hoon Park
2007CLUSTERAnomaly localization in large-scale clusters.Ziming Zheng, Yawei Li, Zhiling Lan
2007ICPPA Meta-Learning Failure Predictor for Blue Gene/L Systems.Prashasta Gujrati, Yawei Li, Zhiling Lan, Rajeev Thakur, John White
2007ICPPFault-Driven Re-Scheduling For Improving System-level Fault Resilience.Yawei Li, Prashasta Gujrati, Zhiling Lan, Xian-He Sun
2006CCGRIDEvaluating Performance and Scalability of Advanced Accelerator Simulations.Jungmin Lee, Zhiling Lan, James F. Amundson, Panagiotis Spentzouris
2006CCGRIDExploit Failure Prediction for Adaptive Fault-Tolerance in Cluster Computing.Yawei Li, Zhiling Lan
2006SCPoster reception - Improving fault resilience of high performance applications.Yawei Li, Zhiling Lan
2005CCGRIDA novel workload migration scheme for heterogeneous distributed computing.Yawei Li, Zhiling Lan
2004CISA Survey of Load Balancing in Grid Computing.Yawei Li, Zhiling Lan
2003CLUSTERPerformance Analysis of a Large-Scale Cosmology Application on Three Cluster Systems.Zhiling Lan, Prathibha Deshikachar
2001ICPPDynamic Load Balancing for Structured Adaptive Mesh Refinement Applications.Zhiling Lan, Valerie E. Taylor, Greg Bryan
2001SCDynamic load balancing of SAMR applications on distributed systems.Zhiling Lan, Valerie E. Taylor, Greg Bryan
2000HPDCProphesy: An Infrastructure for Analyzing and Modeling the Performance of Parallel and Distributed Applications.Xingfu Wu, Valerie E. Taylor, Jonathan Geisler, Xin Li, Zhiling Lan, Rick L. Stevens, Mark Hereld, Ivan R. Judson