Skip to content

Reward Learning as Doubly Nonparametric Bandits: Optimal Design and Scaling Laws.

Kush Bhatia, Wenshuo Guo, Jacob Steinhardt

Year2023
ProceedingsAISTATS

Browse the full AISTATS paper archive.