Reward-on-the-Line: A Novel Offline Reinforcement Learning Method for Building Legal Conversational Agents.
Xubo Lin, Mingze Wang, Grace Hui Yang, Daniel Chen
Browse the full AIES paper archive.
Xubo Lin, Mingze Wang, Grace Hui Yang, Daniel Chen
Browse the full AIES paper archive.