Policy Learning from Large Vision-Language Model Feedback Without Reward Modeling.
Tung Minh Luu, Donghoon Lee, Younghwan Lee, Chang D. Yoo
Browse the full IROS paper archive.
Tung Minh Luu, Donghoon Lee, Younghwan Lee, Chang D. Yoo
Browse the full IROS paper archive.