Offline-to-Online: Case-Based Knowledge Distillation with Large Language Models for Reinforcement Learning.
Hongzhe Liu, Quan Liu, Lan Wu, Meilong Shi, Zhiming Cui
Browse the full ICCBR paper archive.
Hongzhe Liu, Quan Liu, Lan Wu, Meilong Shi, Zhiming Cui
Browse the full ICCBR paper archive.