The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data.
Thomas Pouplin, Kasia Kobalczyk, Hao Sun, Mihaela van der Schaar
Browse the full ICML paper archive.
Thomas Pouplin, Kasia Kobalczyk, Hao Sun, Mihaela van der Schaar
Browse the full ICML paper archive.