| 2023 | Noisy Self-Training with Data Augmentations for Offensive and Hate Speech Detection Tasks. | Joo Augusto Leite, Carolina Scarton, Diego F. Silva |
| 2023 | Challenges of GPT-3-Based Conversational Agents for Healthcare. | Fabian Lechner, Allison Lahnala, Charles Welch, Lucie Flek |
| 2023 | Simultaneous Interpreting as a Noisy Channel: How Much Information Gets Through. | Maria Kunilovskaya, Heike Przybyl, Ekaterina Lapshinova-Koltunski, Elke Teich |
| 2023 | Taxonomy-Based Automation of Prior Approval Using Clinical Guidelines. | Saranya Krishnamoorthy, Ayush Singh |
| 2023 | Sentence Embedding Models for Ancient Greek Using Multilingual Knowledge Distillation. | Kevin Krahn, Derrick Tate, Andrew C. Lamicela |
| 2023 | Advancing Topical Text Classification: A Novel Distance-Based Method with Contextual Embeddings. | Andriy Kosar, Guy De Pauw, Walter Daelemans |
| 2023 | Evaluating Data Augmentation for Medication Identification in Clinical Notes. | Jordan Koontz, Maite Oronoz, Alicia Prez |
| 2023 | A tailored Handwritten-Text-Recognition System for Medieval Latin. | Philipp Koch, Gilary Vera Nuez, Esteban Garces Arias, Christian Heumann, Matthias Schffel, Alexander Hberlin, Matthias Aenmacher |
| 2023 | Bridging the Gap between Subword and Character Segmentation in Pretrained Language Models. | Shun Kiyono, Sho Takase, Shengzhe Li, Toshinori Sato |
| 2023 | Word Sense Disambiguation for Ancient Greek: Sourcing a training corpus through translation alignment. | Alek Keersmaekers, Wouter Mercelis, Toon Van Hal |
| 2023 | Morphological and Semantic Evaluation of Ancient Chinese Machine Translation. | Kai Jin, Dan Zhao, Wuying Liu |
| 2023 | Categorising Fine-to-Coarse Grained Misinformation: An Empirical Study of the COVID-19 Infodemic. | Ye Jiang, Xingyi Song, Carolina Scarton, Iknoor Singh, Ahmet Aker, Kalina Bontcheva |
| 2023 | Pretraining Language- and Domain-Specific BERT on Automatically Translated Text. | Tatsuya Ishigaki, Yui Uehara, Goran Topic, Hiroya Takamura |
| 2023 | Uncertainty Quantification of Text Classification in a Multi-Label Setting for Risk-Sensitive Systems. | Jinha Hwang, Carol Gudumotu, Benyamin Ahmadnia |
| 2023 | Towards a Consensus Taxonomy for Annotating Errors in Automatically Generated Text. | Rudali Huidrom, Anya Belz |
| 2023 | Coding Design of Oracle Bone Inscriptions Input Method Based on "ZhongHuaZiKu" Database. | Dongxin Hu |
| 2023 | Clinical Text Classification to SNOMED CT Codes Using Transformers Trained on Linked Open Medical Ontologies. | Anton Hristov, Petar Ivanov, Anna Aksenova, Tsvetan Asamov, Pavlin Gyurov, Todor Primov, Svetla Boytcheva |
| 2023 | Reading between the Lines: Information Extraction from Industry Requirements. | Ole Magnus Holter, Basil Ell |
| 2023 | Explainable Event Detection with Event Trigger Identification as Rationale Extraction. | Hansi Hettiarachchi, Tharindu Ranasinghe |
| 2023 | Unimodal Intermediate Training for Multimodal Meme Sentiment Classification. | Muzhaffar Hazman, Susan McKeever, Josephine Griffith |
| 2023 | Enriched Pre-trained Transformers for Joint Slot Filling and Intent Detection. | Momchil Hardalov, Ivan Koychev, Preslav Nakov |
| 2023 | Discourse Analysis of Argumentative Essays of English Learners Based on CEFR Level. | Blaise Hanel, Leila Kosseim |
| 2023 | Introducing an Open Source Library for Sumerian Text Analysis. | Hansel Guzman-Soto, Yudong Liu |
| 2023 | Exploring Unsupervised Semantic Similarity Methods for Claim Verification in Health Care News Articles. | Vishwani Gupta, Astrid Viciano, Holger Wormer, Najmehsadat Mousavinezhad |
| 2023 | Data Augmentation for Fake News Detection by Combining Seq2seq and NLI. | Anna V. Glazkova |