Skip to content

Learning When Not to Answer: a Ternary Reward Structure for Reinforcement Learning Based Question Answering.

Frderic Godin, Anjishnu Kumar, Arpit Mittal

VenueANAACL
Year2019
ProceedingsNAACL-HLT (2)

Browse the full NAACL paper archive.