Skip to content

Can Small LLMs Learn a Robust Theory of Mind via RLVR? Investigating Generalization through the False-Belief Task.

Sneheel Sarangi, Hanan Salam

VenueA*ACL
Year2026
ProceedingsACL (Findings)

Browse the full ACL paper archive.