Skip to content

CAPO: A Unified Policy Gradient Approach for Reward and Cost Optimization in Safe Reinforcement Learning (Student Abstract).

Xiaotao Liu, Mohit Prashant, Arvind Easwaran

VenueA*AAAI
Year2026
ProceedingsAAAI

Browse the full AAAI paper archive.