Yonathan Efroni
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
24
Venues
6
Active years
2018–2026
Best venue rank
A*
Where they publish
Papers
24 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2026 | EACL | Imbalanced Gradients in RL Post-Training of Multi-Task LLMs. | Runzhe Wu, Ankur Samanta, Ayush Jain, Scott Fujimoto, Jeongyeol Kwon, Ben Kretzu, Youliang Yu, Kaveh Hassani, Boris Vidolov, Yonathan Efroni |
| 2025 | ICLR | Time After Time: Deep-Q Effect Estimation for Interventions on When and What to do. | Yoav Wald, Mark Goldstein, Yonathan Efroni, Wouter A. C. van Amsterdam, Rajesh Ranganath |
| 2025 | ICLR | Exploiting Structure in Offline Multi-Agent RL: The Benefits of Low Interaction Rank. | Wenhao Zhan, Scott Fujimoto, Zheqing Zhu, Jason D. Lee, Daniel Jiang, Yonathan Efroni |
| 2025 | ICML | Aligned Multi Objective Optimization. | Yonathan Efroni, Ben Kretzu, Daniel Jiang, Jalaj Bhandari, Zheqing Zhu, Karen Ullrich |
| 2024 | ICML | PcLast: Discovering Plannable Continuous Latent States. | Anurag Koul, Shivakanth Sujit, Shaoru Chen, Ben Evans, Lili Wu, Byron Xu, Rajan Chari, Riashat Islam, Raihan Seraj, Yonathan Efroni, Lekan P. Molu, Miroslav Dudk, John Langford, Alex Lamb |
| 2024 | ICML | Prospective Side Information for Latent MDPs. | Jeongyeol Kwon, Yonathan Efroni, Shie Mannor, Constantine Caramanis |
| 2023 | ICML | Principled Offline RL in the Presence of Rich Exogenous Information. | Riashat Islam, Manan Tomar, Alex Lamb, Yonathan Efroni, Hongyu Zang, Aniket Rajiv Didolkar, Dipendra Misra, Xin Li, Harm van Seijen, Remi Tachet des Combes, John Langford |
| 2023 | ICML | Reward-Mixing MDPs with Few Latent Contexts are Learnable. | Jeongyeol Kwon, Yonathan Efroni, Constantine Caramanis, Shie Mannor |
| 2022 | COLT | Sample-Efficient Reinforcement Learning in the Presence of Exogenous Information. | Yonathan Efroni, Dylan J. Foster, Dipendra Misra, Akshay Krishnamurthy, John Langford |
| 2022 | ICLR | Provably Filtering Exogenous Distractors using Multistep Inverse Dynamics. | Yonathan Efroni, Dipendra Misra, Akshay Krishnamurthy, Alekh Agarwal, John Langford |
| 2022 | ICLR | Mirror Descent Policy Optimization. | Manan Tomar, Lior Shani, Yonathan Efroni, Mohammad Ghavamzadeh |
| 2022 | ICML | Provable Reinforcement Learning with a Short-Term Memory. | Yonathan Efroni, Chi Jin, Akshay Krishnamurthy, Sobhan Miryoosefi |
| 2022 | ICML | Sparsity in Partially Controllable Linear Systems. | Yonathan Efroni, Sham M. Kakade, Akshay Krishnamurthy, Cyril Zhang |
| 2022 | ICML | Coordinated Attacks against Contextual Bandits: Fundamental Limits and Defense Mechanisms. | Jeongyeol Kwon, Yonathan Efroni, Constantine Caramanis, Shie Mannor |
| 2021 | AAAI | Reinforcement Learning with Trajectory Feedback. | Yonathan Efroni, Nadav Merlis, Shie Mannor |
| 2021 | ICML | Confidence-Budget Matching for Sequential Budgeted Learning. | Yonathan Efroni, Nadav Merlis, Aadirupa Saha, Shie Mannor |
| 2021 | UAI | Bandits with partially observable confounded data. | Guy Tennenholtz, Uri Shalit, Shie Mannor, Yonathan Efroni |
| 2020 | AAAI | Adaptive Trust Region Policy Optimization: Global Convergence and Faster Rates for Regularized MDPs. | Lior Shani, Yonathan Efroni, Shie Mannor |
| 2020 | ICML | Optimistic Policy Optimization with Bandit Feedback. | Lior Shani, Yonathan Efroni, Aviv Rosenberg, Shie Mannor |
| 2020 | ICML | Multi-step Greedy Reinforcement Learning Algorithms. | Manan Tomar, Yonathan Efroni, Mohammad Ghavamzadeh |
| 2019 | AAAI | How to Combine Tree-Search Methods in Reinforcement Learning. | Yonathan Efroni, Gal Dalal, Bruno Scherrer, Shie Mannor |
| 2019 | ICML | Exploration Conscious Reinforcement Learning Revisited. | Lior Shani, Yonathan Efroni, Shie Mannor |
| 2019 | ICML | Action Robust Reinforcement Learning and Applications in Continuous Control. | Chen Tessler, Yonathan Efroni, Shie Mannor |
| 2018 | ICML | Beyond the One-Step Greedy Approach in Reinforcement Learning. | Yonathan Efroni, Gal Dalal, Bruno Scherrer, Shie Mannor |