Skip to content

Sample-Efficient Iterative Lower Bound Optimization of Deep Reactive Policies for Planning in Continuous MDPs.

Siow Meng Low, Akshat Kumar, Scott Sanner

VenueA*AAAI
Year2022
ProceedingsAAAI

Browse the full AAAI paper archive.