Skip to content

BRAIn: Bayesian Reward-conditioned Amortized Inference for natural language generation from feedback.

Gaurav Pandey, Yatin Nandwani, Tahira Naseem, Mayank Mishra, Guangxuan Xu, Dinesh Raghu, Sachindra Joshi, Asim Munawar, Ramn Fernandez Astudillo

VenueA*ICML
Year2024
ProceedingsICML

Browse the full ICML paper archive.