Skip to content

Howard's Policy Iteration is Subexponential for Deterministic Markov Decision Problems with Rewards of Fixed Bit-size and Arbitrary Discount Factor.

Dibyangshu Mukherjee, Shivaram Kalyanakrishnan

VenueA*ICAPS
Year2025
ProceedingsICAPS

Browse the full ICAPS paper archive.