Shixiang Gu
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
17
Venues
5
Active years
2016–2023
Best venue rank
A*
Where they publish
Papers
17 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2023 | EMNLP | Large Language Models Can Self-Improve. | Jiaxin Huang, Shixiang Gu, Le Hou, Yuexin Wu, Xuezhi Wang, Hongkun Yu, Jiawei Han |
| 2021 | ICLR | Deployment-Efficient Reinforcement Learning via Model-Based Offline Optimization. | Tatsuya Matsushima, Hiroki Furuta, Yutaka Matsuo, Ofir Nachum, Shixiang Gu |
| 2020 | EMNLP | Human-centric dialog training via offline reinforcement learning. | Natasha Jaques, Judy Hanwen Shen, Asma Ghandeharioun, Craig Ferguson, gata Lapedriza, Noah Jones, Shixiang Gu, Rosalind W. Picard |
| 2020 | ICLR | Dynamics-Aware Unsupervised Discovery of Skills. | Archit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar, Karol Hausman |
| 2019 | CoRL | A Divergence Minimization Perspective on Imitation Learning Methods. | Seyed Kamyar Seyed Ghasemipour, Richard S. Zemel, Shixiang Gu |
| 2019 | ICLR | Near-Optimal Representation Learning for Hierarchical Reinforcement Learning. | Ofir Nachum, Shixiang Gu, Honglak Lee, Sergey Levine |
| 2019 | ICLR | Doubly Reparameterized Gradient Estimators for Monte Carlo Objectives. | George Tucker, Dieterich Lawson, Shixiang Gu, Chris J. Maddison |
| 2018 | ICLR | Leave no Trace: Learning to Reset for Safe and Autonomous Reinforcement Learning. | Benjamin Eysenbach, Shixiang Gu, Julian Ibarz, Sergey Levine |
| 2018 | ICLR | Temporal Difference Models: Model-Free Deep RL for Model-Based Control. | Vitchyr Pong, Shixiang Gu, Murtaza Dalal, Sergey Levine |
| 2018 | ICLR | The Mirage of Action-Dependent Baselines in Reinforcement Learning. | George Tucker, Surya Bhupatiraju, Shixiang Gu, Richard E. Turner, Zoubin Ghahramani, Sergey Levine |
| 2018 | ICML | The Mirage of Action-Dependent Baselines in Reinforcement Learning. | George Tucker, Surya Bhupatiraju, Shixiang Gu, Richard E. Turner, Zoubin Ghahramani, Sergey Levine |
| 2017 | ICLR | Q-Prop: Sample-Efficient Policy Gradient with An Off-Policy Critic. | Shixiang Gu, Timothy P. Lillicrap, Zoubin Ghahramani, Richard E. Turner, Sergey Levine |
| 2017 | ICLR | Categorical Reparameterization with Gumbel-Softmax. | Eric Jang, Shixiang Gu, Ben Poole |
| 2017 | ICLR | Tuning Recurrent Neural Networks with Reinforcement Learning. | Natasha Jaques, Shixiang Gu, Richard E. Turner, Douglas Eck |
| 2017 | ICML | Sequence Tutor: Conservative Fine-Tuning of Sequence Generation Models with KL-control. | Natasha Jaques, Shixiang Gu, Dzmitry Bahdanau, Jos Miguel Hernndez-Lobato, Richard E. Turner, Douglas Eck |
| 2017 | ICRA | Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates. | Shixiang Gu, Ethan Holly, Timothy P. Lillicrap, Sergey Levine |
| 2016 | ICML | Continuous Deep Q-Learning with Model-based Acceleration. | Shixiang Gu, Timothy P. Lillicrap, Ilya Sutskever, Sergey Levine |