Skip to content

Cache-Efficient Posterior Sampling for Reinforcement Learning with LLM-Derived Priors Across Discrete and Continuous Domains.

Ibne Farabi Shihab, Sanjeda Akter, Anuj Sharma

VenueA*EMNLP
Year2025
ProceedingsEMNLP

Browse the full EMNLP paper archive.