| 2024 | Reward Engineering for Generating Semi-structured Explanation. | Jiuzhou Han, Wray L. Buntine, Ehsan Shareghi |
| 2024 | Are Large Language Model-based Evaluators the Solution to Scaling Up Multilingual Evaluation? | Rishav Hada, Varun Gumma, Adrian de Wynter, Harshita Diddee, Mohamed Ahmed, Monojit Choudhury, Kalika Bali, Sunayana Sitaram |
| 2024 | Toward Sentiment Aware Semantic Change Analysis. | Roksana Goworek, Haim Dubossarsky |
| 2024 | A* shortest string decoding for non-idempotent semirings. | Kyle Gorman, Cyril Allauzen |
| 2024 | Where are we Still Split on Tokenization? | Rob Goot |
| 2024 | Multi-Reference Benchmarks for Russian Grammatical Error Correction. | Frank Palma Gomez, Alla Rozovskaya |
| 2024 | Accurate and Well-Calibrated ICD Code Assignment Through Attention Over Diverse Label Embeddings. | Gonalo Gomes, Isabel Coutinho, Bruno Martins |
| 2024 | Large-Scale Label Interpretation Learning for Few-Shot Named Entity Recognition. | Jonas Golde, Felix Hamborg, Alan Akbik |
| 2024 | Anisotropy Is Inherent to Self-Attention in Transformers. | Nathan Godey, ric Villemonte de la Clergerie, Benot Sagot |
| 2024 | Plan-Grounded Large Language Models for Dual Goal Conversational Settings. | Diogo Glria-Silva, Rafael Ferreira, Diogo Tavares, David Semedo, Joo Magalhes |
| 2024 | Should I try multiple optimizers when fine-tuning a pre-trained Transformer for NLP tasks? Should I tune their hyperparameters? | Nefeli Gkouti, Prodromos Malakasiotis, Stavros Toumpis, Ion Androutsopoulos |
| 2024 | SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Models. | Xiang Gao, Jiaxin Zhang, Lalla Mouatadid, Kamalika Das |
| 2024 | Evaluating Unsupervised Argument Aligners via Generation of Conclusions of Structured Scientific Abstracts. | Yingqiang Gao, Nianlong Gu, Jessica Lam, James Henderson, Richard H. R. Hahnloser |
| 2024 | MultiMUC: Multilingual Template Filling on MUC-4. | William Gantt, Shabnam Behzad, Hannah Youngeun An, Yunmo Chen, Aaron Steven White, Benjamin Van Durme, Mahsa Yarmohammadi |
| 2024 | Examining Gender and Racial Bias in Large Vision-Language Models Using a Novel Dataset of Parallel Images. | Kathleen C. Fraser, Svetlana Kiritchenko |
| 2024 | Embible: Reconstruction of Ancient Hebrew and Aramaic Texts Using Transformers. | Niv Fono, Harel Moshayof, Eldar Karol, Itai Assraf, Mark Last |
| 2024 | On the Benefits of Fine-Grained Loss Truncation: A Case Study on Factuality in Summarization. | Lorenzo Jaime Flores, Arman Cohan |
| 2024 | AnnoPlot: Interactive Visualizations of Text Annotations. | Elisabeth Fittschen, Tim Fischer, Daniel Brhl, Julia Spahr, Yuliia Lysa, Phuoc Thang Le |
| 2024 | Establishing degrees of closeness between audio recordings along different dimensions using large-scale cross-lingual models. | Maxime Fily, Guillaume Wisniewski, Severine Guillaume, Gilles Adda, Alexis Michaud |
| 2024 | More Discriminative Sentence Embeddings via Semantic Graph Smoothing. | Chakib Fettal, Lazhar Labiod, Mohamed Nadif |
| 2024 | DepressMind: A Depression Surveillance System for Social Media Analysis. | Roque Fernndez-Iglesias, Marcos Fernndez-Pichel, Mario Ezra Aragn, David E. Losada |
| 2024 | PromptExplainer: Explaining Language Models through Prompt-based Learning. | Zijian Feng, Hanzhang Zhou, Zixiao Zhu, Kezhi Mao |
| 2024 | Who Needs Decoders? Efficient Estimation of Sequence-Level Attributes with Proxies. | Yassir Fathullah, Puria Radmard, Adian Liusie, Mark J. F. Gales |
| 2024 | On-the-fly Denoising for Data Augmentation in Natural Language Understanding. | Tianqing Fang, Wenxuan Zhou, Fangyu Liu, Hongming Zhang, Yangqiu Song, Muhao Chen |
| 2024 | Joint Inference of Retrieval and Generation for Passage Re-ranking. | Wei Fang, Yung-Sung Chuang, James R. Glass |