A Reduction from Reinforcement Learning to No-Regret Online Learning.
Ching-An Cheng, Remi Tachet des Combes, Byron Boots, Geoffrey J. Gordon
Browse the full AISTATS paper archive.
Ching-An Cheng, Remi Tachet des Combes, Byron Boots, Geoffrey J. Gordon
Browse the full AISTATS paper archive.