| 2025 | EMNLP | Playpen: An Environment for Exploring Learning From Dialogue Game Feedback. | Nicola Horst, Davide Mazzaccara, Antonia Schmidt, Michael Sullivan, Filippo Moment, Luca Franceschetti, Philipp Sadler, Sherzod Hakimov, Alberto Testoni, Raffaella Bernardi, Raquel Fernndez, Alexander Koller, Oliver Lemon, David Schlangen, Mario Giulianelli, Alessandro Suglia |
| 2024 | COLING | Sharing the Cost of Success: A Game for Evaluating and Learning Collaborative Multi-Agent Instruction Giving and Following Policies. | Philipp Sadler, Sherzod Hakimov, David Schlangen |
| 2024 | INLG | The Unreasonable Ineffectiveness of Nucleus Sampling on Mitigating Text Memorization. | Luka Borec, Philipp Sadler, David Schlangen |
| 2023 | ACL | Yes, this Way! Learning to Ground Referring Expressions into Actions with Intra-episodic Feedback from Supportive Teachers. | Philipp Sadler, Sherzod Hakimov, David Schlangen |
| 2023 | EACL | Pento-DIARef: A Diagnostic Dataset for Learning the Incremental Algorithm for Referring Expression Generation from Examples. | Philipp Sadler, David Schlangen |
| 2023 | EMNLP | clembench: Using Game Play to Evaluate Chat-Optimized Language Models as Conversational Agents. | Kranti Chalamalasetti, Jana Gtze, Sherzod Hakimov, Brielen Madureira, Philipp Sadler, David Schlangen |
| 2020 | INLG | From "Before" to "After": Generating Natural Language Instructions from Image Pairs in a Simple Visual Domain. | Robin Rojowiec, Jana Gtze, Philipp Sadler, Henrik Voigt, Sina Zarrie, David Schlangen |
| 2019 | INLG | Can Neural Image Captioning be Controlled via Forced Attention? | Philipp Sadler, Tatjana Scheffler, David Schlangen |