Shixiang Shane Gu
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
10
Venues
3
Active years
2019–2024
Best venue rank
A*
Where they publish
Papers
10 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2024 | ICLR | Multimodal Web Navigation with Instruction-Finetuned Foundation Models. | Hiroki Furuta, Kuang-Huei Lee, Ofir Nachum, Yutaka Matsuo, Aleksandra Faust, Shixiang Shane Gu, Izzeddin Gur |
| 2023 | ICLR | A System for Morphology-Task Generalization via Unified Representation and Behavior Distillation. | Hiroki Furuta, Yusuke Iwasawa, Yutaka Matsuo, Shixiang Shane Gu |
| 2023 | ICLR | Mind's Eye: Grounded Language Model Reasoning through Simulation. | Ruibo Liu, Jason Wei, Shixiang Shane Gu, Te-Yen Wu, Soroush Vosoughi, Claire Cui, Denny Zhou, Andrew M. Dai |
| 2022 | ICLR | Generalized Decision Transformer for Offline Hindsight Information Matching. | Hiroki Furuta, Yutaka Matsuo, Shixiang Shane Gu |
| 2022 | ICML | Why Should I Trust You, Bellman? The Bellman Error is a Poor Replacement for Value Error. | Scott Fujimoto, David Meger, Doina Precup, Ofir Nachum, Shixiang Shane Gu |
| 2022 | ICML | Blocks Assemble! Learning to Assemble with Large-Scale Structured Reinforcement Learning. | Seyed Kamyar Seyed Ghasemipour, Satoshi Kataoka, Byron David, Daniel Freeman, Shixiang Shane Gu, Igor Mordatch |
| 2021 | ICML | Variational Empowerment as Representation Learning for Goal-Conditioned Reinforcement Learning. | Jongwook Choi, Archit Sharma, Honglak Lee, Sergey Levine, Shixiang Shane Gu |
| 2021 | ICML | Policy Information Capacity: Information-Theoretic Measure for Task Complexity in Deep Reinforcement Learning. | Hiroki Furuta, Tatsuya Matsushima, Tadashi Kozuno, Yutaka Matsuo, Sergey Levine, Ofir Nachum, Shixiang Shane Gu |
| 2021 | ICML | EMaQ: Expected-Max Q-Learning Operator for Simple Yet Effective Offline and Online RL. | Seyed Kamyar Seyed Ghasemipour, Dale Schuurmans, Shixiang Shane Gu |
| 2019 | CoRL | Multi-Agent Manipulation via Locomotion using Hierarchical Sim2Real. | Ofir Nachum, Michael Ahn, Hugo Ponte, Shixiang Shane Gu, Vikash Kumar |