Skip to content

The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data.

Thomas Pouplin, Kasia Kobalczyk, Hao Sun, Mihaela van der Schaar

VenueA*ICML
Year2025
ProceedingsICML

Browse the full ICML paper archive.