Grounding Large Language Models in Interactive Environments with Online Reinforcement Learning.
Thomas Carta, Clment Romac, Thomas Wolf, Sylvain Lamprier, Olivier Sigaud, Pierre-Yves Oudeyer
Browse the full ICML paper archive.
Thomas Carta, Clment Romac, Thomas Wolf, Sylvain Lamprier, Olivier Sigaud, Pierre-Yves Oudeyer
Browse the full ICML paper archive.