| 2024 | NSina: A News Corpus for Sinhala. | Hansi Hettiarachchi, Damith Premasiri, Lasitha Randunu Chandrakantha Uyangodage, Tharindu Ranasinghe |
| 2024 | MedQA-SWE - a Clinical Question & Answer Dataset for Swedish. | Niclas Hertzberg, Anna Lokrantz |
| 2024 | Sparse Logistic Regression with High-order Features for Automatic Grammar Rule Extraction from Treebanks. | Santiago Herrera, Caio Corro, Sylvain Kahane |
| 2024 | ShadowSense: A Multi-annotated Dataset for Evaluating Word Sense Induction. | Ondrej Herman, Milos Jakubcek |
| 2024 | Zero-shot Cross-lingual Automated Essay Scoring. | Junyi He, Xia Li |
| 2024 | The Influence of Automatic Speech Recognition on Linguistic Features and Automatic Alzheimer's Disease Detection from Spontaneous Speech. | Jonathan Heitz, Gerold Schneider, Nicolas Langer |
| 2024 | Automatic Identification of COVID-19-Related Conspiracy Narratives in German Telegram Channels and Chats. | Philipp Heinrich, Andreas Blombach, Bao Minh Doan Dang, Leonardo Zilio, Linda Havenstein, Nathan Dykes, Stephanie Evert, Fabian Schfer |
| 2024 | Demonstration Retrieval-Augmented Generative Event Argument Extraction. | Shiming He, Yu Hong, Shuai Yang, Jianmin Yao, Guodong Zhou |
| 2024 | Mitigating Shortcuts in Language Models with Soft Label Encoding. | Zirui He, Huiqi Deng, Haiyan Zhao, Ninghao Liu, Mengnan Du |
| 2024 | Decoding Probing: Revealing Internal Linguistic Structures in Neural Language Models Using Minimal Pairs. | Linyang He, Peili Chen, Ercong Nie, Yuanning Li, Jonathan R. Brennan |
| 2024 | mForms : Multimodal Form Filling with Question Answering. | Larry Heck, Simon Heck, Anirudh S. Sundar |
| 2024 | Zero-Shot Cross-Lingual Document-Level Event Causality Identification with Heterogeneous Graph Contrastive Transfer Learning. | Zhitao He, Pengfei Cao, Zhuoran Jin, Yubo Chen, Kang Liu, Zhiqiang Zhang, Mengshu Sun, Jun Zhao |
| 2024 | Deriving Entity-Specific Embeddings from Multi-Entity Sequences. | Connor T. Heaton, Prasenjit Mitra |
| 2024 | From Technology to Market. Bilingual Corpus on the Evaluation of Technology Opportunity Discovery. | Amir Hazem, Kazuyuki Motohashi, Chen Zhu |
| 2024 | Reassessing Semantic Knowledge Encoded in Large Language Models through the Word-in-Context Task. | Yoshihiko Hayashi |
| 2024 | ADEA: An Argumentative Dialogue Dataset on Ethical Issues Concerning Future A.I. Applications. | Christian Hauptmann, Adrian Krenzer, Antonia Krause, Frank Puppe |
| 2024 | Tell Me Again! a Large-Scale Dataset of Multiple Summaries for the Same Story. | Hans Ole Hatzel, Chris Biemann |
| 2024 | PromptStream: Self-Supervised News Story Discovery Using Topic-Aware Article Representations. | Arezoo Hatefi, Anton Eklund, Mona Forsman |
| 2024 | Zero- and Few-Shot Prompting with LLMs: A Comparative Study with Fine-tuned Models for Bangla Sentiment Analysis. | Md. Arid Hasan, Shudipta Das, Afiyat Anjum, Firoj Alam, Anika Anjum, Avijit Sarker, Sheak Rashed Haider Noori |
| 2024 | Can GPT-4 Identify Propaganda? Annotation and Detection of Propaganda Spans in News Articles. | Maram Hasanain, Fatema Ahmed, Firoj Alam |
| 2024 | EPOQUE: An English-Persian Quality Estimation Dataset. | Mohammed Hossein Jafari Harandi, Fatemeh Azadi, Mohammad Javad Dousti, Heshaam Faili |
| 2024 | Cognitive Information Bottleneck: Extracting Minimal Sufficient Cognitive Language Processing Signals. | Yuto Harada, Yohei Oseki |
| 2024 | RECIPE4U: Student-ChatGPT Interaction Dataset in EFL Writing Education. | Jieun Han, Haneul Yoo, Junho Myung, Minsun Kim, Tak Yeon Lee, So-Yeon Ahn, Alice Oh |
| 2024 | A Dual-View Approach to Classifying Radiology Reports by Co-Training. | Yutong Han, Yan Yuan, Lili Mou |
| 2024 | Towards Understanding the Relationship between In-context Learning and Compositional Generalization. | Sungjun Han, Sebastian Pad |