| 2023 | Can Language Models Be Tricked by Language Illusions? Easier with Syntax, Harder with Semantics. | Yuhan Zhang, Edward Gibson, Forrest Davis |
| 2023 | Frontmatter. | |
| 2023 | Structural Ambiguity and its Disambiguation in Language Model Based Parsers: the Case of Dutch Clause Relativization. | Gijs Wijnholds, Michael Moortgat |
| 2023 | Mind the instructions: a holistic evaluation of consistency and interactions in prompt-based learning. | Lucas Weber, Elia Bruni, Dieuwke Hupkes |
| 2023 | JaSPICE: Automatic Evaluation Metric Using Predicate-Argument Structures for Image Captioning Models. | Yuiga Wada, Kanta Kaneda, Komei Sugiura |
| 2023 | Humans and language models diverge when predicting repeating text. | Aditya R. Vaidya, Javier Turek, Alexander Huth |
| 2023 | The Validity of Evaluation Results: Assessing Concurrence Across Compositionality Benchmarks. | Kaiser Sun, Adina Williams, Dieuwke Hupkes |
| 2023 | Enhancing Code-mixed Text Generation Using Synthetic Data Filtering in Neural Machine Translation. | Dama Sravani, Radhika Mamidi |
| 2023 | Towards Better Evaluation of Instruction-Following: A Case-Study in Summarization. | Ondrej Skopek, Rahul Aralikatte, Sian Gooding, Victor Carbune |
| 2023 | A Minimal Approach for Natural Language Action Space in Text-based Games. | Dongwon Ryu, Meng Fang, Gholamreza Haffari, Shirui Pan, Ehsan Shareghi |
| 2023 | PSST! Prosodic Speech Segmentation with Transformers. | Nathan Roll, Calbert Graham, Simon Todd |
| 2023 | Med-HALT: Medical Domain Hallucination Test for Large Language Models. | Ankit Pal, Logesh Kumar Umapathi, Malaikannan Sankarasubbu |
| 2023 | Future Lens: Anticipating Subsequent Tokens from a Single Hidden State. | Koyena Pal, Jiuding Sun, Andrew Yuan, Byron C. Wallace, David Bau |
| 2023 | Quirk or Palmer: A Comparative Study of Modal Verb Frameworks with Annotated Datasets. | Risako Owan, Maria L. Gini, Dongyeop Kang |
| 2023 | Attribution and Alignment: Effects of Local Context Repetition on Utterance Production and Comprehension in Dialogue. | Aron Molnar, Jaap Jumelet, Mario Giulianelli, Arabella Sinclair |
| 2023 | On the utility of enhancing BERT syntactic bias with Token Reordering Pretraining. | Yassir El Mesbahi, Atif Mahmud, Abbas Ghaddar, Mehdi Rezagholizadeh, Philippe Langlais, Prasanna Parthasarathi |
| 2023 | ToMChallenges: A Principle-Guided Dataset and Diverse Evaluation Tasks for Exploring Theory of Mind. | Xiaomeng Ma, Lingyu Gao, Qihui Xu |
| 2023 | Revising with a Backward Glance: Regressions and Skips during Reading as Cognitive Signals for Revision Policies in Incremental Processing. | Brielen Madureira, Pelin elikkol, David Schlangen |
| 2023 | REFER: An End-to-end Rationale Extraction Framework for Explanation Regularization. | Mohammad Reza Ghasemi Madani, Pasquale Minervini |
| 2023 | Quantifying Information of Tokens for Simple and Flexible Simultaneous Machine Translation. | Donghyun Lee, Minkyung Park, Byung-Jun Lee |
| 2023 | A Comparative Study on Textual Saliency of Styles from Eye Tracking, Annotations, and Language Models. | Karin de Langis, Dongyeop Kang |
| 2023 | Investigating the Nature of Disagreements on Mid-Scale Ratings: A Case Study on the Abstractness-Concreteness Continuum. | Urban Knuples, Diego Frassinelli, Sabine Schulte im Walde |
| 2023 | MuLER: Detailed and Scalable Reference-based Evaluation. | Taelin Karidi, Leshem Choshen, Gal Patel, Omri Abend |
| 2023 | Tree-shape Uncertainty for Analyzing the Inherent Branching Bias of Unsupervised Parsing Models. | Taiga Ishii, Yusuke Miyao |
| 2023 | The Impact of Familiarity on Naming Variation: A Study on Object Naming in Mandarin Chinese. | Yunke He, Xixian Liao, Jialing Liang, Gemma Boleda |