PrefMMT: Modeling Human Preferences in Preference-based Reinforcement Learning with Multimodal Transformers.
Dezhong Zhao, Ruiqi Wang, Dayoon Suh, Taehyeon Kim, Ziqin Yuan, Byung-Cheol Min, Guohua Chen
Browse the full IROS paper archive.
Dezhong Zhao, Ruiqi Wang, Dayoon Suh, Taehyeon Kim, Ziqin Yuan, Byung-Cheol Min, Guohua Chen
Browse the full IROS paper archive.