Skip to content

Hengshuai Yao

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

15

Venues

8

Active years

2008–2023

Best venue rank

A*

Where they publish

Papers

15 indexed papers, newest first.

YearVenueTitleAuthors
2023AAAIThe Sufficiency of Off-Policyness and Soft Clipping: PPO Is Still Insufficient according to an Off-Policy Measure.Xing Chen, Dongcui Diao, Hechang Chen, Hengshuai Yao, Haiyin Piao, Zhixiao Sun, Zhiwei Yang, Randy Goebel, Bei Jiang, Yi Chang
2022UAIUnderstanding and mitigating the limitations of prioritized experience replay.Yangchen Pan, Jincheng Mei, Amir-massoud Farahmand, Martha White, Hengshuai Yao, Mohsen Rohani, Jun Luo
2021ICMLBreaking the Deadly Triad with a Target Network.Shangtong Zhang, Hengshuai Yao, Shimon Whiteson
2020ICMLProvably Convergent Two-Timescale Off-Policy Actor-Critic with Function Approximation.Shangtong Zhang, Bo Liu, Hengshuai Yao, Shimon Whiteson
2020IJCAIWeakly Supervised Few-shot Object Segmentation using Co-Attention with Visual and Semantic Embeddings.Mennatullah Siam, Naren Doraiswamy, Boris N. Oreshkin, Hengshuai Yao, Martin Jgersand
2020ICRAMapless Navigation among Dynamics with Social-safety-awareness: a reinforcement learning approach from 2D laser scans.Jun Jin, Nhat M. Nguyen, Nazmus Sakib, Daniel Graves, Hengshuai Yao, Martin Jgersand
2019AAAIACE: An Actor Ensemble Algorithm for Continuous Control with Tree Search.Shangtong Zhang, Hengshuai Yao
2019AAAIQUOTA: The Quantile Option Architecture for Reinforcement Learning.Shangtong Zhang, Hengshuai Yao
2019ICDMM-estimation in Low-Rank Matrix Factorization: A General Framework.Wei Tu, Peng Liu, Jingyu Zhao, Yi Liu, Linglong Kong, Guodong Li, Bei Jiang, Guangjian Tian, Hengshuai Yao
2019ICMLDistributional Reinforcement Learning for Efficient Exploration.Borislav Mavrin, Hengshuai Yao, Linglong Kong, Kaiwen Wu, Yaoliang Yu
2019IJCAIHill Climbing on Value Estimates for Search-control in Dyna.Yangchen Pan, Hengshuai Yao, Amir-massoud Farahmand, Martha White
2014WWWLearning to predict trending queries: classification - based.Chi-Hoon Lee, Hengshuai Yao, Xu He, Su Han Chan, JieYang Chang, Farzin Maghoul
2012AAAIApproximate Policy Iteration with Linear Action Models.Hengshuai Yao, Csaba Szepesvri
2008ICMLPreconditioned temporal difference learning.Hengshuai Yao, Zhi-Qiang Liu
2008ISAIMMinimal Residual Approaches for Policy Evaluation in Large Sparse Markov Chains.Hengshuai Yao, Zhi-Qiang Liu