Skip to content

On learning history-based policies for controlling Markov decision processes.

Gandharv Patil, Aditya Mahajan, Doina Precup

Year2024
ProceedingsAISTATS

Browse the full AISTATS paper archive.