Aligning Dialogue Agents with Global Feedback via Large Language Model Multimodal Reward Decomposition.
Don (Dong Won) Lee, Hae Won Park, Cynthia Breazeal, Louis-Philippe Morency
Browse the full EMNLP paper archive.
Don (Dong Won) Lee, Hae Won Park, Cynthia Breazeal, Louis-Philippe Morency
Browse the full EMNLP paper archive.