Skip to content

FAST-Q: Fast-track Exploration with Adversarially Balanced State Representations for Counterfactual Action Estimation in Offline Reinforcement Learning.

Pulkit Agrawal, Rukma Talwadker, Aditya Pareek, Tridib Mukherjee

VenueA*WWW
Year2025
ProceedingsWWW (Companion Volume)

Browse the full WWW paper archive.