Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment.
Yuang Cai, Yuyu Yuan, Jinsheng Shi, Qinhong Lin
Browse the full AAAI paper archive.
Yuang Cai, Yuyu Yuan, Jinsheng Shi, Qinhong Lin
Browse the full AAAI paper archive.