| 2021 | Clustering Word Embeddings with Self-Organizing Maps. Application on LaRoSeDa - A Large Romanian Sentiment Data Set. | Anca Maria Tache, Mihaela Gaman, Radu Tudor Ionescu |
| 2021 | Story Centaur: Large Language Model Few Shot Learning as a Creative Writing Tool. | Ben Swanson, Kory W. Mathewson, Ben Pietrzak, Sherol Chen, Monica Dinalescu |
| 2021 | A New View of Multi-modal Language Analysis: Audio and Video Features as Text "Styles". | Zhongkai Sun, Prathusha Kameswara Sarma, Yingyu Liang, William A. Sethares |
| 2021 | Cross-Cultural Similarity Features for Cross-Lingual Transfer Learning of Pragmatically Motivated Tasks. | Jimin Sun, Hwijeen Ahn, Chan Young Park, Yulia Tsvetkov, David R. Mortensen |
| 2021 | An Empirical Study on the Generalization Power of Neural Representations Learned via Visual Guessing Games. | Alessandro Suglia, Yonatan Bisk, Ioannis Konstas, Antonio Vergari, Emanuele Bastianelli, Andrea Vanzo, Oliver Lemon |
| 2021 | Benchmarking Machine Reading Comprehension: A Psychological Perspective. | Saku Sugawara, Pontus Stenetorp, Akiko Aizawa |
| 2021 | Non-Autoregressive Text Generation with Pre-trained Language Models. | Yixuan Su, Deng Cai, Yan Wang, David Vandyke, Simon Baker, Piji Li, Nigel Collier |
| 2021 | Metric-Type Identification for Multi-Level Header Numerical Tables in Scientific Papers. | Lya Hulliyyatus Suadaa, Hidetaka Kamigaito, Manabu Okumura, Hiroya Takamura |
| 2021 | Relating Relations: Meta-Relation Extraction from Online Health Forum Posts. | Daniel Stickley |
| 2021 | Recipes for Adapting Pre-trained Monolingual and Multilingual Models to Machine Translation. | Asa Cooper Stickland, Xian Li, Marjan Ghazvininejad |
| 2021 | How to Evaluate a Summarizer: Study Design and Statistical Analysis for Manual Linguistic Quality Evaluation. | Julius Steen, Katja Markert |
| 2021 | Why Is MBTI Personality Detection from Texts a Difficult Task? | Sanja Stajner, Seren Yenikent |
| 2021 | Complex Question Answering on knowledge graphs using machine translation and multi-task learning. | Saurabh Srivastava, Mayur Patidar, Sudip Chowdhury, Puneet Agarwal, Indrajit Bhattacharya, Gautam Shroff |
| 2021 | NLQuAD: A Non-Factoid Long Question Answering Data Set. | Amir Soleimani, Christof Monz, Marcel Worring |
| 2021 | LSOIE: A Large-Scale Dataset for Supervised Open Information Extraction. | Jacob Solawetz, Stefan Larson |
| 2021 | We Need To Talk About Random Splits. | Anders Sgaard, Sebastian Ebert, Jasmijn Bastings, Katja Filippova |
| 2021 | Exploring Neural Language Models via Analysis of Local and Global Self-Attention Spaces. | Blaz Skrlj, Shane Sheehan, Nika Erzen, Marko Robnik-Sikonja, Saturnino Luz, Senja Pollak |
| 2021 | Learning From Revisions: Quality Assessment of Claims in Argumentation at Scale. | Gabriella Skitalinskaya, Jonas Klaff, Henning Wachsmuth |
| 2021 | DRAG: Director-Generator Language Modelling Framework for Non-Parallel Author Stylized Rewriting. | Hrituraj Singh, Gaurav Verma, Aparna Garimella, Balaji Vasan Srinivasan |
| 2021 | Contrasting distinct structured views to learn sentence embeddings. | Antoine Simoulin, Benot Crabb |
| 2021 | Semantic Oppositeness Assisted Deep Contextual Modeling for Automatic Rumor Detection in Social Networks. | Nisansa de Silva, Dejing Dou |
| 2021 | Adaptive Mixed Component LDA for Low Resource Topic Modeling. | Suzanna Sia, Kevin Duh |
| 2021 | Development of Conversational AI for Sleep Coaching Programme. | Heereen Shim |
| 2021 | With Measured Words: Simple Sentence Selection for Black-Box Optimization of Sentence Compression Algorithms. | Yotam Shichel, Meir Kalech, Oren Tsur |
| 2021 | Leveraging End-to-End ASR for Endangered Language Documentation: An Empirical Study on Yolxochitl Mixtec. | Jiatong Shi, Jonathan D. Amith, Rey Castillo Garca, Esteban Guadalupe Sierra, Kevin Duh, Shinji Watanabe |