Distributional reinforcement learning with linear function approximation.
Marc G. Bellemare, Nicolas Le Roux, Pablo Samuel Castro, Subhodeep Moitra
Browse the full AISTATS paper archive.
Marc G. Bellemare, Nicolas Le Roux, Pablo Samuel Castro, Subhodeep Moitra
Browse the full AISTATS paper archive.