| 2021 | Contextualizing Variation in Text Style Transfer Datasets. | Stephanie Schoch, Wanyu Du, Yangfeng Ji |
| 2021 | TUDA-Reproducibility @ ReproGen: Replicability of Human Evaluation of Text-to-Text and Concept-to-Text Generation. | Christian Richter, Yanran Chen, Steffen Eger |
| 2021 | Automatic Verification of Data Summaries. | Rayhane Rezgui, Mohammed Saeed, Paolo Papotti |
| 2021 | A Reproduction Study of an Annotation-based Human Evaluation of MT Outputs. | Maja Popovic, Anya Belz |
| 2021 | Grounding NBA Matchup Summaries. | Tadashi Nomoto |
| 2021 | Predicting Antonyms in Context using BERT. | Ayana Niwa, Keisuke Nishiguchi, Naoaki Okazaki |
| 2021 | Shared Task on Feedback Comment Generation for Language Learners. | Ryo Nagata, Masato Hagiwara, Kazuaki Hanawa, Masato Mita, Artem Chernodub, Olena Nahorna |
| 2021 | Underreporting of errors in NLG output, and what to do about it. | Emiel van Miltenburg, Miruna-Adriana Clinciu, Ondrej Dusek, Dimitra Gkatzia, Stephanie Inglis, Leo Leppnen, Saad Mahamood, Emma Manning, Stephanie Schoch, Craig Thomson, Luou Wen |
| 2021 | Another PASS: A Reproduction Study of the Human Evaluation of a Football Report Generation System. | Simon Mille, Thiago Castro Ferreira, Anya Belz, Brian Davis |
| 2021 | Neural Methodius Revisited: Do Discourse Relations Help with Pre-Trained Models Too? | Aleksandre Maskharashvili, Symon Jory Stevens-Guille, Xintong Li, Michael White |
| 2021 | Explaining Decision-Tree Predictions by Addressing Potential Conflicts between Predictions and Plausible Expectations. | Sameen Maruf, Ingrid Zukerman, Ehud Reiter, Gholamreza Haffari |
| 2021 | Exploring Structural Encoding for Data-to-Text Generation. | Joy Mahapatra, Utpal Garain |
| 2021 | Reproducing a Comparison of Hedged and Non-hedged NLG Texts. | Saad Mahamood |
| 2021 | Goal-Oriented Script Construction. | Qing Lyu, Li Zhang, Chris Callison-Burch |
| 2021 | Self-Training for Compositional Neural NLG in Task-Oriented Dialogue. | Xintong Li, Symon Jory Stevens-Guille, Aleksandre Maskharashvili, Michael White |
| 2021 | WeaSuL: Weakly Supervised Dialogue Policy Learning: Reward Estimation for Multi-turn Dialogue. | Anant Khandelwal |
| 2021 | Formulating Neural Sentence Ordering as the Asymmetric Traveling Salesman Problem. | Vishal Keswani, Harsh Jhamtani |
| 2021 | Text-in-Context: Token-Level Error Detection for Table-to-Text Generation. | Zdenek Kasner, Simon Mille, Ondrej Dusek |
| 2021 | Exploring Input Representation Granularity for Generating Questions Satisfying Question-Answer Congruence. | Madeeswaran Kannan, Haemanth Santhi Ponnusamy, Kordula De Kuthy, Lukas Stein, Detmar Meurers |
| 2021 | BERT-based distractor generation for Swedish reading comprehension questions using a small-scale dataset. | Dmytro Kalpakchi, Johan Boye |
| 2021 | Attention Is Indeed All You Need: Semantically Attention-Guided Decoding for Data-to-Text NLG. | Juraj Juraska, Marilyn A. Walker |
| 2021 | Using BERT for choosing classifiers in Mandarin. | Jani Jrnfors, Guanyi Chen, Kees van Deemter, Rint Sybesma |
| 2021 | Generating Racing Game Commentary from Vision, Language, and Structured Data. | Tatsuya Ishigaki, Goran Topic, Yumi Hamazono, Hiroshi Noji, Ichiro Kobayashi, Yusuke Miyao, Hiroya Takamura |
| 2021 | Generating Diverse Descriptions from Semantic Graphs. | Jiuzhou Han, Daniel Beck, Trevor Cohn |
| 2021 | Shared Task in Evaluating Accuracy: Leveraging Pre-Annotations in the Validation Process. | Nicolas Garneau, Luc Lamontagne |