Skip to content

Dongbin Zhao

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

92

Venues

16

Active years

2003–2026

Best venue rank

A*

Where they publish

Papers

92 indexed papers, newest first.

YearVenueTitleAuthors
2026ACLSpec-o3: A Tool-Augmented Vision-Language Agent for Rare Celestial Object Candidate Vetting via Automated Spectral Inspection.Minghui Jia, Qichao Zhang, Ali Luo, Linjing Li, Shuo Ye, Hailing Lu, Wen Hou, Dongbin Zhao
2026ACLBeyond Query Memorization: Large Language Model Routing with Query Decomposition and Historical Matching.Bo Lv, Jingbo Sun, Jianwei Lv, Chen Tang, Shaojie Zhang, Nayu Liu, Guoxin Yu, Zihao Li, Qichao Zhang, Dongbin Zhao, Ping Luo, Yue Yu
2026ACLAutoSearch: Adaptive Search Depth for Efficient Agentic RAG via Reinforcement Learning.Jingbo Sun, Wenyue Chong, Songjun Tu, Qichao Zhang, Yaocheng Zhang, Jiajun Chai, Xiaohan Wang, Wei Lin, Guojun Yin, Dongbin Zhao
2026ACLCriticSearch: Fine-Grained Credit Assignment for Search Agents via a Retrospective Critic.Yaocheng Zhang, Haohuan Huang, Zijun Song, Zijie Zhao, Qichao Zhang, Yuanheng Zhu, Dongbin Zhao
2025AAAIIn-Dataset Trajectory Return Regularization for Offline Preference-based Reinforcement Learning.Songjun Tu, Jingbo Sun, Qichao Zhang, Yaocheng Zhang, Jia Liu, Ke Chen, Dongbin Zhao
2025EMNLPRLAE: Reinforcement Learning-Assisted Ensemble for LLMs.Yuqian Fu, Yuanheng Zhu, Jiajun Chai, Guojun Yin, Wei Lin, Qichao Zhang, Dongbin Zhao
2025ICCVWorld4Drive: End-to-End Autonomous Driving via Intention-Aware Physical Latent World Model.Yupeng Zheng, Pengxuan Yang, Zebin Xing, Qichao Zhang, Yuhang Zheng, Yinfeng Gao, Pengfei Li, Teng Zhang, Zhongpu Xia, Peng Jia, Xianpeng Lang, Dongbin Zhao
2025ICLREmpowering LLM Agents with Zero-Shot Optimal Decision-Making through Q-learning.Jiajun Chai, Sicheng Li, Yuqian Fu, Dongbin Zhao, Yuanheng Zhu
2025ICLRINS: Interaction-aware Synthesis to Enhance Offline Multi-agent Reinforcement Learning.Yuqian Fu, Yuanheng Zhu, Jian Zhao, Jiajun Chai, Dongbin Zhao
2025ICLRDivergence-Regularized Discounted Aggregation: Equilibrium Finding in Multiplayer Partially Observable Stochastic Games.Runyu Lu, Yuanheng Zhu, Dongbin Zhao
2025ICLRUnsupervised Zero-Shot Reinforcement Learning via Dual-Value Forward-Backward Representation.Jingbo Sun, Songjun Tu, Qichao Zhang, Haoran Li, Xin Liu, Yaran Chen, Ke Chen, Dongbin Zhao
2025ICMLConstrained Exploitability Descent: An Offline Reinforcement Learning Method for Finding Mixed-Strategy Nash Equilibrium.Runyu Lu, Yuanheng Zhu, Dongbin Zhao
2025ICMLDipLLM: Fine-Tuning LLM for Strategic Decision-making in Diplomacy.Kaixuan Xu, Jiajun Chai, Sicheng Li, Yuqian Fu, Yuanheng Zhu, Dongbin Zhao
2025IROSAdvancing Object-Goal Navigation through LLM-enhanced Object Affinities Transfer.Mengying Lin, Shugao Liu, Dingxi Zhang, Yaran Chen, Zhaoran Wang, Haoran Li, Dongbin Zhao
2025ICRAUncAD: Towards Safe End-to-end Autonomous Driving via Online Map Uncertainty.Pengxuan Yang, Yupeng Zheng, Qichao Zhang, Kefei Zhu, Zebin Xing, Qiao Lin, Yun-Fu Liu, Zhiguo Su, Dongbin Zhao
2025SMCFusionNav: Enhancing Zero-Shot Object-Goal Navigation via 3D Semantic Fusion and Farsight Value Reasoning.Shugao Liu, Qichao Zhang, Haoran Li, Dongbin Zhao
2024IJCNNHigh-quality Synthetic Data is Efficient for Model-based Offline Reinforcement Learning.Qichao Zhang, Xing Fang, Kaixuan Xu, Weixin Zhao, Haoran Li, Dongbin Zhao
2024WWWUser Response Modeling in Reinforcement Learning for Ads Allocation.Zhiyuan Zhang, Qichao Zhang, Xiaoxu Wu, Xiaowen Shi, Guogang Liao, Yongkang Wang, Xingxing Wang, Dongbin Zhao
2023IJCNNNeuronsMAE: A Novel Multi-Agent Reinforcement Learning Environment for Cooperative and Competitive Multi-Robot Tasks.Guangzheng Hu, Haoran Li, Shasha Liu, Yuanheng Zhu, Dongbin Zhao
2023IJCNNDense Attention: A Densely Connected Attention Mechanism for Vision Transformer.Nannan Li, Yaran Chen, Dongbin Zhao
2023IJCNNAdvantage Constrained Proximal Policy Optimization in Multi-Agent Reinforcement Learning.Weifan Li, Yuanheng Zhu, Dongbin Zhao
2023ICRASTEPS: Joint Self-supervised Nighttime Image Enhancement and Depth Estimation.Yupeng Zheng, Chengliang Zhong, Pengfei Li, Huan-ang Gao, Yuhang Zheng, Bu Jin, Ling Wang, Hao Zhao, Guyue Zhou, Qichao Zhang, Dongbin Zhao
2022IJCNNNeurons Perception Dataset for RoboMaster AI Challenge.Haoran Li, Zicheng Duan, Jiaqi Li, Mingjun Ma, Yaran Chen, Dongbin Zhao
2020IJCNNShift-Invariant Convolutional Network Search.Nannan Li, Yaran Chen, Zixiang Ding, Dongbin Zhao
2020IJCNNAn Improved Minimax-Q Algorithm Based on Generalized Policy Iteration to Solve a Chaser-Invader Game.Minsong Liu, Yuanheng Zhu, Dongbin Zhao
2020IJCNNRailNet: An Information Aggregation Network for Rail Track Segmentation.Haoran Li, Qichao Zhang, Dongbin Zhao, Yaran Chen
2020IJCNNCooperative Multi-Agent Deep Reinforcement Learning with Counterfactual Reward.Kun Shao, Yuanheng Zhu, Zhentao Tang, Dongbin Zhao
2020ISNNContourRend: A Segmentation Method for Improving Contours by Rendering.Junwen Chen, Yi Lu, Yaran Chen, Dongbin Zhao, Zhonghua Pang
2020ISNNDynamically Weighted Model Predictive Control of Affine Nonlinear Systems Based on Two-Timescale Neurodynamic Optimization.Jiasen Wang, Jun Wang, Dongbin Zhao
2019IJCNNLane Change Decision-making through Deep Reinforcement Learning with Rule-based Constraints.Junjie Wang, Qichao Zhang, Dongbin Zhao, Yaran Chen
2019IJCNNModel-Free Reinforcement Learning based Lateral Control for Lane Keeping.Qichao Zhang, Rui Luo, Dongbin Zhao, Chaomin Luo, Dianwei Qian
2019IJCNNOptimal Pedestrian Evacuation in Building with Consecutive Differential Dynamic Programming.Yuanheng Zhu, Haibo He, Dongbin Zhao, Zhongsheng Hou
2019ISNNGraph-FCN for Image Semantic Segmentation.Yi Lu, Yaran Chen, Dongbin Zhao, Jianxin Chen
2019SMCDeep Kalman Filter with Optical Flow for Multiple Object Tracking.Yaran Chen, Dongbin Zhao, Haoran Li
2018ICONIPValue Iteration Algorithm for Optimal Consensus Control of Multi-agent Systems.Qichao Zhang, Dongbin Zhao
2018ICONIPDriving Control with Deep and Reinforcement Learning in The Open Racing Car Simulator.Yuanheng Zhu, Dongbin Zhao
2018IJCNNA temporal-based deep learning method for multiple objects detection in autonomous driving.Yaran Chen, Dongbin Zhao, Haoran Li, Dong Li, Ping Guo
2018IJCNNDeepSign: Deep Learning based Traffic Sign Recognition.Dong Li, Dongbin Zhao, Yaran Chen, Qichao Zhang
2018IJCNNVisual Navigation with Actor-Critic Deep Reinforcement Learning.Kun Shao, Dongbin Zhao, Yuanheng Zhu, Qichao Zhang
2018IJCNNModel-Free Reinforcement Learning for Fully Cooperative Multi-Agent Graphical Games.Qichao Zhang, Dongbin Zhao, Frank L. Lewis
2017ICONIPFMR-GA - A Cooperative Multi-agent Reinforcement Learning Algorithm Based on Gradient Ascent.Zhen Zhang, Dongqing Wang, Dongbin Zhao, Tingting Song
2017ICONIPOff-Policy Reinforcement Learning for Partially Unknown Nonzero-Sum Games.Qichao Zhang, Dongbin Zhao, Sibo Zhang
2017IJCNNPolicy gradient methods with Gaussian process modelling acceleration.Dong Li, Dongbin Zhao, Qichao Zhang, Chaomin Luo
2017ISNNMulti-task Learning with Cartesian Product-Based Multi-objective Combination for Dangerous Object Detection.Yaran Chen, Dongbin Zhao
2016IJCNNEnsemble LSDD-based change detection tests.Li Bu, Cesare Alippi, Dongbin Zhao
2016IJCNNA general adaptive dynamic programming approach with experience replay.Bin Wang, Dongbin Zhao, Jin Cheng, Yuan Xu, Yueyang Li
2016IJCNNModel-free reinforcement learning for nonlinear zero-sum games with simultaneous explorations.Qichao Zhang, Dongbin Zhao, Yuanheng Zhu, Xi Chen
2016IJCNNConvolutional fitted Q iteration for vision-based control problems.Dongbin Zhao, Yuanheng Zhu, Le Lv, Yaran Chen, Qichao Zhang
2015IJCNNThermal comfort control based on MEC algorithm for HVAC systems.Dong Li, Dongbin Zhao, Yuanheng Zhu, Zhongpu Xia
2015IJCNNOnline reinforcement learning by Bayesian inference.Zhongpu Xia, Dongbin Zhao
2015ISNNEvent-Triggered H ∞ Control for Continuous-Time Nonlinear System.Dongbin Zhao, Qichao Zhang, Xiangjun Li, Lingda Kong
2014IJCNNA hierarchical classification algorithm for evaluating energy consumption behaviors.Li Bu, Dongbin Zhao, Yu Liu, Qiang Guan
2014IJCNNA Kaiman filter-based actor-critic learning approach.Bin Wang, Dongbin Zhao
2014IJCNNEvent-triggered reinforcement learning approach for unknown nonlinear continuous-time system.Xiangnan Zhong, Zhen Ni, Haibo He, Xin Xu, Dongbin Zhao
2013ICONIPOnline Model-Free RLSPI Algorithm for Nonlinear Discrete-Time Non-affine Systems.Yuanheng Zhu, Dongbin Zhao
2013IJCNNA prior-free encode-decode change detection test to inspect datastreams for concept drift.Cesare Alippi, Li Bu, Dongbin Zhao
2012ICONIPSVM-Based Just-in-Time Adaptive Classifiers.Cesare Alippi, Li Bu, Dongbin Zhao
2012ICONIPThe Optimal Control of Discrete-Time Delay Nonlinear System with Dual Heuristic Dynamic Programming.Bin Wang, Dongbin Zhao
2012IJCNNReinforcement learning control based on multi-goal representation using hierarchical heuristic dynamic programming.Zhen Ni, Haibo He, Dongbin Zhao, Danil V. Prokhorov
2012IJCNNNeural and fuzzy dynamic programming for under-actuated systems.Dongbin Zhao, Yuanheng Zhu, Haibo He
2012ISNNA Hierarchical Neural Network Architecture for Classification.Jing Wang, Haibo He, Yuan Cao, Jin Xu, Dongbin Zhao
2011IJCNNNeural-network-based optimal control for a class of nonlinear cdiscrete-time systems with control constraints using the citerative GDHP algorithm.Derong Liu, Ding Wang, Dongbin Zhao
2010IJCNNA comparative study of urban traffic signal control with reinforcement learning and Adaptive Dynamic Programming.Yujie Dai, Dongbin Zhao, Jianqiang Yi
2009IJCNNCoordinated multiple ramps metering based on neuro-fuzzy adaptive dynamic programming.Xuerui Bai, Dongbin Zhao, Jianqiang Yi
2009IROSFuzzy logic based adjustment control of a cable-driven auto-leveling parallel robot.Yi Yu, Jianqiang Yi, Chengdong Li, Dongbin Zhao, Jianhong Zhang
2008IJCNNRamp metering based on on-line ADHDP (lambda) controller.Xuerui Bai, Dongbin Zhao, Jianqiang Yi, Jing Xu
2008ICRAControl of a class of under-actuated systems with saturation using hierarchical sliding mode.Dianwei Qian, Jianqiang Yi, Dongbin Zhao
2008IJCNNAdaptive dynamic neuro-fuzzy system for traffic signal control.Tao Li, Dongbin Zhao, Jianqiang Yi
2008ICRATrajectory tracking control of omnidirecitonal wheeled mobile manipulators: Robust neural network based sliding mode approach.Dong Xu, Dongbin Zhao, Jianqiang Yi, Xiang-min Tan, Zonghai Chen
2007ACCStabilization of an Underactuated Surface Vessel via Discontinuous Control.Jin Cheng, Jianqiang Yi, Dongbin Zhao
2007ACCHierarchical Sliding Mode Control to Swing up a Pendubot.Dianwei Qian, Jianqiang Yi, Dongbin Zhao
2007ICRARobust Control Using Sliding Mode for a Class of Under-Actuated Systems With Mismatched Uncertainties.Dianwei Qian, Jianqiang Yi, Dongbin Zhao
2007IROSRobust adaptive tracking control of omnidirecitonal wheeled mobile manipulators.Dong Xu, Dongbin Zhao, Jianqiang Yi, Xiang-min Tan
2007IROSMotion regulation of redundantly actuated omni-directional Wheeled Mobile Robots with internal force control.Dongbin Zhao, Jianqiang Yi, Xuyue Deng
2007ISNNApproximate Dynamic Programming for Ship Course Control.Xuerui Bai, Jianqiang Yi, Dongbin Zhao
2007ISNNA Comparison of Four Data Mining Models: Bayes, Neural Network, SVM and Decision Trees in Identifying Syndromes in Coronary Heart Disease.Jianxin Chen, Yanwei Xing, Guangcheng Xi, Jing Chen, Jianqiang Yi, Dongbin Zhao, Jie Wang
2007ISNNApplication of ADP to Intersection Signal Control.Tao Li, Dongbin Zhao, Jianqiang Yi
2007ISNNMultiple Approximate Dynamic Programming Controllers for Congestion Control.Yanping Xiang, Jianqiang Yi, Dongbin Zhao
2006ICNCRobot Planning with Artificial Potential Field Guided Ant Colony Optimization Algorithm.Dongbin Zhao, Jianqiang Yi
2006IROSHierarchical Sliding Mode Control for Series Double Inverted Pendulums System.Dianwei Qian, Jianqiang Yi, Dongbin Zhao, Yinxing Hao
2006ISNNA Particle Swarm Optimized Fuzzy Neural Network Control for Acrobot.Dongbin Zhao, Jianqiang Yi
2005IROSTracking control of mobile manipulator with dynamical uncertainties.Zuoshi Song, Dongbin Zhao, Jianqiang Yi, Xinchun Li
2005IROSDouble layer sliding mode control for second-order underactuated mechanical systems.Wei Wang, Jianqiang Yi, Dongbin Zhao, Xiaojing Liu
2005IROSCascade sliding-mode controller for large-scale underactuated systems.Jianqiang Yi, Wei Wang, Dongbin Zhao, Xiaojing Liu
2005ICRAPose Estimation and Structure Recovery from Point Pairs.Zhiguang Zhong, Jianqiang Yi, Dongbin Zhao
2005ISNNAdaptive Inverse Control System Based on Least Squares Support Vector Machines.Xiaojing Liu, Jianqiang Yi, Dongbin Zhao
2005ISNNA Reinforcement Learning Based Radial-Bassis Function Network Control System.Jianing Li, Jianqiang Yi, Dongbin Zhao, Guangcheng Xi
2004ICARCVA robust doorplate recognition system.Yiping Hong, Jianqiang Yi, Dongbin Zhao, Xinzheng Li
2004ICARCVSmooth time-varying regulation of nonholonomic chained systems.Zuoshi Song, Jianqiang Yi, Dongbin Zhao, Xinchun Li
2004ICARCVMotion vision for mobile robot localization.Zhiguang Zhong, Jianqiang Yi, Dongbin Zhao, Yiping Hong, Xinzheng Li
2004ICRAPassive Adaptive Grasp Multi-fingered Humanoid Robot Hand with High Under-actuated Function.Wenzeng Zhang, Qiang Chen, Zhenguo Sun, Dongbin Zhao
2003ICRAUnder-actuated passive adaptive grasp humanoid robot hand with control of grasping force.Wenzeng Zhang, Qiang Chen, Zhenguo Sun, Dongbin Zhao