Ofir Nachum
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
37
Venues
9
Active years
2017–2024
Best venue rank
A*
Where they publish
Papers
37 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2024 | ICLR | Multimodal Web Navigation with Instruction-Finetuned Foundation Models. | Hiroki Furuta, Kuang-Huei Lee, Ofir Nachum, Yutaka Matsuo, Aleksandra Faust, Shixiang Shane Gu, Izzeddin Gur |
| 2024 | IROS | The Design of the Barkour Benchmark for Robot Agility. | Wenhao Yu, Ken Caluwaerts, Atil Iscen, J. Chase Kew, Tingnan Zhang, Daniel Freeman, Lisa Lee, Stefano Saliceti, Vincent Zhuang, Nathan Batchelor, Steven Bohez, Federico Casarini, Jos Enrique Chen, Erwin Coumans, Adil Dostmohamed, Gabriel Dulac-Arnold, Alejandro Escontrela, Erik Frey, Roland Hafner, Deepali Jain, Bauyrjan Jyenis, Yuheng Kuang, Tsang-Wei Edward Lee, Ofir Nachum, Ken Oslund, Francesco Romano, Fereshteh Sadeghi, Baruch Tabanpour, Daniel Zheng, Michael Neunert, Raia Hadsell, Nicolas Heess, Francesco Nori, Jeff Seto, Carolina Parada, Vikas Sindhwani, Vincent Vanhoucke, Jie Tan, Kuang-Huei Lee |
| 2023 | CoRL | Q-Transformer: Scalable Offline Reinforcement Learning via Autoregressive Q-Functions. | Yevgen Chebotar, Quan Vuong, Karol Hausman, Fei Xia, Yao Lu, Alex Irpan, Aviral Kumar, Tianhe Yu, Alexander Herzog, Karl Pertsch, Keerthana Gopalakrishnan, Julian Ibarz, Ofir Nachum, Sumedh Anand Sontakke, Grecia Salazar, Huong T. Tran, Jodilyn Peralta, Clayton Tan, Deeksha Manjunath, Jaspiar Singh, Brianna Zitkovich, Tomas Jackson, Kanishka Rao, Chelsea Finn, Sergey Levine |
| 2023 | CoRL | Contrastive Value Learning: Implicit Models for Simple Offline RL. | Bogdan Mazoure, Benjamin Eysenbach, Ofir Nachum, Jonathan Tompson |
| 2023 | EMNLP | Understanding HTML with Large Language Models. | Izzeddin Gur, Ofir Nachum, Yingjie Miao, Mustafa Safdari, Austin V. Huang, Aakanksha Chowdhery, Sharan Narang, Noah Fiedel, Aleksandra Faust |
| 2023 | ICLR | A Mixture-of-Expert Approach to RL-based Dialogue Management. | Yinlam Chow, Aza Tulepbergenov, Ofir Nachum, Dhawal Gupta, Moonkyung Ryu, Mohammad Ghavamzadeh, Craig Boutilier |
| 2023 | ICLR | Dichotomy of Control: Separating What You Can Control from What You Cannot. | Sherry Yang, Dale Schuurmans, Pieter Abbeel, Ofir Nachum |
| 2023 | ICML | Multi-Environment Pretraining Enables Transfer to Action Limited Datasets. | David Venuto, Sherry Yang, Pieter Abbeel, Doina Precup, Igor Mordatch, Ofir Nachum |
| 2022 | AISTATS | Offline Policy Selection under Uncertainty. | Mengjiao Yang, Bo Dai, Ofir Nachum, George Tucker, Dale Schuurmans |
| 2022 | ICLR | Policy Gradients Incorporating the Future. | David Venuto, Elaine Lau, Doina Precup, Ofir Nachum |
| 2022 | ICLR | TRAIL: Near-Optimal Imitation Learning with Suboptimal Data. | Mengjiao Yang, Sergey Levine, Ofir Nachum |
| 2022 | ICML | Model Selection in Batch Policy Optimization. | Jonathan Lee, George Tucker, Ofir Nachum, Bo Dai |
| 2022 | ICML | Why Should I Trust You, Bellman? The Bellman Error is a Poor Replacement for Value Error. | Scott Fujimoto, David Meger, Doina Precup, Ofir Nachum, Shixiang Shane Gu |
| 2022 | IROS | PI-ARS: Accelerating Evolution-Learned Visual-Locomotion with Predictive Information Representations. | Kuang-Huei Lee, Ofir Nachum, Tingnan Zhang, Sergio Guadarrama, Jie Tan, Wenhao Yu |
| 2021 | ICLR | OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning. | Anurag Ajay, Aviral Kumar, Pulkit Agrawal, Sergey Levine, Ofir Nachum |
| 2021 | ICLR | Benchmarks for Deep Off-Policy Evaluation. | Justin Fu, Mohammad Norouzi, Ofir Nachum, George Tucker, Ziyu Wang, Alexander Novikov, Mengjiao Yang, Michael R. Zhang, Yutian Chen, Aviral Kumar, Cosmin Paduraru, Sergey Levine, Tom Le Paine |
| 2021 | ICLR | Deployment-Efficient Reinforcement Learning via Model-Based Offline Optimization. | Tatsuya Matsushima, Hiroki Furuta, Yutaka Matsuo, Ofir Nachum, Shixiang Gu |
| 2021 | ICLR | Autoregressive Dynamics Models for Offline Policy Evaluation and Optimization. | Michael R. Zhang, Thomas Paine, Ofir Nachum, Cosmin Paduraru, George Tucker, Ziyu Wang, Mohammad Norouzi |
| 2021 | ICML | Policy Information Capacity: Information-Theoretic Measure for Task Complexity in Deep Reinforcement Learning. | Hiroki Furuta, Tatsuya Matsushima, Tadashi Kozuno, Yutaka Matsuo, Sergey Levine, Ofir Nachum, Shixiang Shane Gu |
| 2021 | ICML | Offline Reinforcement Learning with Fisher Divergence Critic Regularization. | Ilya Kostrikov, Rob Fergus, Jonathan Tompson, Ofir Nachum |
| 2021 | ICML | Representation Matters: Offline Pretraining for Sequential Decision Making. | Mengjiao Yang, Ofir Nachum |
| 2020 | AISTATS | Identifying and Correcting Label Bias in Machine Learning. | Heinrich Jiang, Ofir Nachum |
| 2020 | CoRL | Safe Policy Learning for Continuous Control. | Yinlam Chow, Ofir Nachum, Aleksandra Faust, Edgar A. Duez-Guzmn, Mohammad Ghavamzadeh |
| 2020 | ICLR | Imitation Learning via Off-Policy Distribution Matching. | Ilya Kostrikov, Ofir Nachum, Jonathan Tompson |
| 2020 | IJCAI | BRPO: Batch Residual Policy Optimization. | Sungryull Sohn, Yinlam Chow, Jayden Ooi, Ofir Nachum, Honglak Lee, Ed H. Chi, Craig Boutilier |
| 2019 | AISTATS | Robustness Guarantees for Density Clustering. | Heinrich Jiang, Jennifer Jang, Ofir Nachum |
| 2019 | CoRL | Multi-Agent Manipulation via Locomotion using Hierarchical Sim2Real. | Ofir Nachum, Michael Ahn, Hugo Ponte, Shixiang Shane Gu, Vikash Kumar |
| 2019 | ICLR | Near-Optimal Representation Learning for Hierarchical Reinforcement Learning. | Ofir Nachum, Shixiang Gu, Honglak Lee, Sergey Levine |
| 2019 | ICLR | The Laplacian in RL: Learning Representations with Efficient Approximations. | Yifan Wu, George Tucker, Ofir Nachum |
| 2019 | ICML | DeepMDP: Learning Continuous Latent Space Models for Representation Learning. | Carles Gelada, Saurabh Kumar, Jacob Buckman, Ofir Nachum, Marc G. Bellemare |
| 2018 | CVPR | MorphNet: Fast & Simple Resource-Constrained Structure Learning of Deep Networks. | Ariel Gordon, Elad Eban, Ofir Nachum, Bo Chen, Hao Wu, Tien-Ju Yang, Edward Choi |
| 2018 | ICLR | Trust-PCL: An Off-Policy Trust Region Method for Continuous Control. | Ofir Nachum, Mohammad Norouzi, Kelvin Xu, Dale Schuurmans |
| 2018 | ICML | Path Consistency Learning in Tsallis Entropy Regularized MDPs. | Yinlam Chow, Ofir Nachum, Mohammad Ghavamzadeh |
| 2018 | ICML | Smoothed Action Value Functions for Learning Gaussian Policies. | Ofir Nachum, Mohammad Norouzi, George Tucker, Dale Schuurmans |
| 2018 | ICRA | Deep Reinforcement Learning for Vision-Based Robotic Grasping: A Simulated Comparative Evaluation of Off-Policy Methods. | Deirdre Quillen, Eric Jang, Ofir Nachum, Chelsea Finn, Julian Ibarz, Sergey Levine |
| 2017 | ICLR | Learning to Remember Rare Events. | Lukasz Kaiser, Ofir Nachum, Aurko Roy, Samy Bengio |
| 2017 | ICLR | Improving Policy Gradient by Exploring Under-appreciated Rewards. | Ofir Nachum, Mohammad Norouzi, Dale Schuurmans |