| 2021 | A Data-Driven Semi-Automatic Framenet Development Methodology. | Shafqat Mumtaz Virk, Dana Dannlls, Lars Borin, Markus Forsberg |
| 2021 | Mistake Captioning: A Machine Learning Approach for Detecting Mistakes and Generating Instructive Feedback. | Anton Vinogradov, Andrew Miles Byrd, Brent Harrison |
| 2021 | Comparative Analysis of Fine-tuned Deep Learning Language Models for ICD-10 Classification Task for Bulgarian Language. | Boris Velichkov, Sylvia Vassileva, Simeon Gerginov, Boris Kraychev, Ivaylo Ivanov, Philip Ivanov, Ivan Koychev, Svetla Boytcheva |
| 2021 | Contextual-Lexicon Approach for Abusive Language Detection. | Francielle Alves Vargas, Fabiana Rodrigues de Ges, Isabelle Carvalho, Fabrcio Benevenuto, Thiago A. S. Pardo |
| 2021 | Can Multilingual Transformers Fight the COVID-19 Infodemic? | Lasitha Uyangodage, Tharindu Ranasinghe, Hansi Hettiarachchi |
| 2021 | Opinion Prediction with User Fingerprinting. | Kishore Tumarada, Yifan Zhang, Fan Yang, Eduard C. Dragut, Omprakash Gnawali, Arjun Mukherjee |
| 2021 | TR-SEQ: Named Entity Recognition Dataset for Turkish Search Engine Queries. | Berkay Topcu, Ilknur Durgar El-Kahlout |
| 2021 | Serbian NER&Beyond: The Archaic and the Modern Intertwinned. | Branislava Sandrih Todorovic, Cvetana Krstev, Ranka Stankovic, Milica Ikonic Nesic |
| 2021 | An Empirical Analysis of Topic Models: Uncovering the Relationships between Hyperparameters, Document Length and Performance Measures. | Silvia Terragni, Elisabetta Fersini |
| 2021 | Does BERT Understand Idioms? A Probing-Based Empirical Study of BERT Encodings of Idioms. | Minghuan Tan, Jing Jiang |
| 2021 | Learning and Evaluating Chinese Idiom Embeddings. | Minghuan Tan, Jing Jiang |
| 2021 | Tackling Multilinguality and Internationality in Fake News. | Andrey Tagarev, Krasimira Bozhanova, Ivelina Nikolova-Koleva, Ivan Ivanov |
| 2021 | Watching a Language Model Learning Chess. | Andreas Stckl |
| 2021 | How to Obtain Reliable Labels for MBTI Classification from Texts? | Sanja Stajner, Seren Yenikent |
| 2021 | Exploring Reliability of Gold Labels for Emotion Detection in Twitter. | Sanja Stajner |
| 2021 | Exploring German Multi-Level Text Simplification. | Nicolas Spring, Annette Rios, Sarah Ebling |
| 2021 | OCR Processing of Swedish Historical Newspapers Using Deep Hybrid CNN-LSTM Networks. | Molly Brandt Skelbye, Dana Dannlls |
| 2021 | Czert - Czech BERT-like Model for Language Representation. | Jakub Sido, Ondrej Prazk, Pavel Pribn, Jan Pasek, Michal Sejk, Miloslav Konopk |
| 2021 | Towards a Better Understanding of Noise in Natural Language Processing. | Khetam Al Sharou, Zhenhao Li, Lucia Specia |
| 2021 | A Domain-Independent Holistic Approach to Deception Detection. | Sadat Shahriar, Arjun Mukherjee, Omprakash Gnawali |
| 2021 | A Case Study of Deep Learning-Based Multi-Modal Methods for Labeling the Presence of Questionable Content in Movie Trailers. | Mahsa Shafaei, Christos Smailis, Ioannis A. Kakadiaris, Thamar Solorio |
| 2021 | A Lexicon for Profane and Obscene Text Identification in Bengali. | Salim Sazzed |
| 2021 | A Hybrid Approach of Opinion Mining and Comparative Linguistic Analysis of Restaurant Reviews. | Salim Sazzed |
| 2021 | Graph-based Argument Quality Assessment. | Ekaterina Saveleva, Volha Petukhova, Marius Mosbach, Dietrich Klakow |
| 2021 | A Semi-Supervised Approach to Detect Toxic Comments. | Ghivvago Damas Saraiva, Rafael T. Anchita, Francisco Assis Ricarte Neto, Raimundo S. Moura |