Kiant Brantley
Publication record assembled from the DBLP archive of ranked conferences.
Papers indexed
13
Venues
6
Active years
2019–2025
Best venue rank
A*
Where they publish
Papers
13 indexed papers, newest first.
| Year | Venue | Title | Authors |
|---|---|---|---|
| 2025 | ICLR | Regressing the Relative Future: Efficient Policy Optimization for Multi-turn RLHF. | Zhaolin Gao, Wenhao Zhan, Jonathan Daniel Chang, Gokul Swamy, Kiant Brantley, Jason D. Lee, Wen Sun |
| 2025 | ICLR | Diffusing States and Matching Scores: A New Framework for Imitation Learning. | Runzhe Wu, Yiding Chen, Gokul Swamy, Kiant Brantley, Wen Sun |
| 2024 | ICLR | Adversarial Imitation Learning via Boosting. | Jonathan D. Chang, Dhruv Sreenivas, Yingbing Huang, Kiant Brantley, Wen Sun |
| 2024 | ICML | When is Transfer Learning Possible? | My Phan, Kiant Brantley, Stephanie Milani, Soroush Mehri, Gokul Swamy, Geoffrey J. Gordon |
| 2024 | ICML | Coactive Learning for Large Language Models using Implicit User Feedback. | Aaron David Tucker, Kiant Brantley, Adam Cahall, Thorsten Joachims |
| 2024 | WSDM | Ranking with Long-Term Constraints. | Kiant Brantley, Zhichong Fang, Sarah Dean, Thorsten Joachims |
| 2023 | ACL | lilGym: Natural Language Visual Reasoning with Reinforcement Learning. | Anne Wu, Kiant Brantley, Noriyuki Kojima, Yoav Artzi |
| 2023 | EMNLP | Interactive Text Generation. | Felix Faltings, Michel Galley, Kiant Brantley, Baolin Peng, Weixin Cai, Yizhe Zhang, Jianfeng Gao, Bill Dolan |
| 2023 | ICLR | Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy Optimization. | Rajkumar Ramamurthy, Prithviraj Ammanabrolu, Kiant Brantley, Jack Hessel, Rafet Sifa, Christian Bauckhage, Hannaneh Hajishirzi, Yejin Choi |
| 2021 | AAAI | Successor Feature Sets: Generalizing Successor Representations Across Policies. | Kiant Brantley, Soroush Mehri, Geoffrey J. Gordon |
| 2020 | ACL | Active Imitation Learning with Noisy Guidance. | Kiant Brantley, Hal Daum III, Amr Sharaf |
| 2020 | ICLR | Disagreement-Regularized Imitation Learning. | Kiant Brantley, Wen Sun, Mikael Henaff |
| 2019 | ICML | Non-Monotonic Sequential Text Generation. | Sean Welleck, Kiant Brantley, Hal Daum III, Kyunghyun Cho |