Howard's Policy Iteration is Subexponential for Deterministic Markov Decision Problems with Rewards of Fixed Bit-size and Arbitrary Discount Factor.
Dibyangshu Mukherjee, Shivaram Kalyanakrishnan
Browse the full ICAPS paper archive.
Dibyangshu Mukherjee, Shivaram Kalyanakrishnan
Browse the full ICAPS paper archive.