| 2026 | ACL | PiCSAR: Probabilistic Confidence Selection and Ranking for Reasoning Chains. | Joshua Ong Jun Leang, Zheng Zhao, Aryo Pradipta Gema, Sohee Yang, Wai-Chung Kwan, Xuanli He, Wenda Li, Pasquale Minervini, Eleonora Giunchiglia, Shay B. Cohen |
| 2026 | ACL | Can LLMs Compress (and Decompress)? Evaluating Code Understanding and Execution via Invertibility. | Nickil Maveli, Antonio Vergari, Shay B. Cohen |
| 2026 | ACL | Can VLMs Predict Future States? Bootstrapping World Models from Inverse Dynamics. | Yifu Qiu, Yftah Ziser, Anna Korhonen, Shay B. Cohen, Edoardo M. Ponti |
| 2026 | ACL | Thinking in Schemas: Robust Syllogistic Reasoning in LLMs. | Federico Ranaldi, Leonardo Ranaldi, Fabio Massimo Zanzotto, Shay B. Cohen |
| 2025 | ACL | PersonaLens: A Benchmark for Personalization Evaluation in Conversational AI Assistants. | Zheng Zhao, Clara Vania, Subhradeep Kayal, Naila Khan, Shay B. Cohen, Emine Yilmaz |
| 2025 | ACL | Theorem Prover as a Judge for Synthetic Data Generation. | Joshua Ong Jun Leang, Giwon Hong, Wenda Li, Shay B. Cohen |
| 2025 | ACL | Eliciting In-context Retrieval and Reasoning for Long-context Large Language Models. | Yifu Qiu, Varun R. Embar, Yizhe Zhang, Navdeep Jaitly, Shay B. Cohen, Benjamin Han |
| 2025 | CHI | People Attribute Purpose to Autonomous Vehicles When Explaining Their Behavior: Insights from Cognitive Science for Explainable AI. | Balint Gyevnar, Stephanie Droop, Tadeg Quillien, Shay B. Cohen, Neil R. Bramley, Christopher G. Lucas, Stefano V. Albrecht |
| 2025 | EMNLP | CoMAT: Chain of Mathematically Annotated Thought Improves Mathematical Reasoning. | Joshua Ong Jun Leang, Aryo Pradipta Gema, Shay B. Cohen |
| 2025 | EMNLP | Iterative Multilingual Spectral Attribute Erasure. | Shun Shao, Yftah Ziser, Zheng Zhao, Yifu Qiu, Shay B. Cohen, Anna Korhonen |
| 2025 | EMNLP | One More Question is Enough, Expert Question Decomposition (EQD) Model for Domain Quantitative Reasoning. | Mengyu Wang, Sotirios Sabanis, Miguel de Carvalho, Shay B. Cohen, Tiejun Ma |
| 2025 | ICLR | DEPfold: RNA Secondary Structure Prediction as Dependency Parsing. | Ke Wang, Shay B. Cohen |
| 2025 | ICML | PoisonBench: Assessing Language Model Vulnerability to Poisoned Preference Data. | Tingchen Fu, Mrinank Sharma, Philip Torr, Shay B. Cohen, David Krueger, Fazl Barez |
| 2025 | KDD | TSPRank: Bridging Pairwise and Listwise Methods with a Bilinear Travelling Salesman Model. | Weixian Waylon Li, Yftah Ziser, Yifei Xie, Shay B. Cohen, Tiejun Ma |
| 2025 | KDD | Pre-training Time Series Models with Stock Data Customization. | Mengyu Wang, Tiejun Ma, Shay B. Cohen |
| 2025 | NAACL | What can Large Language Models Capture about Code Functional Equivalence? | Nickil Maveli, Antonio Vergari, Shay B. Cohen |
| 2024 | ACL | Can Large Language Models Follow Concept Annotation Guidelines? A Case Study on Scientific and Financial Domains. | Marcio Fonseca, Shay B. Cohen |
| 2024 | ACL | Can Large Language Model Summarizers Adapt to Diverse Scientific Communication Goals? | Marcio Fonseca, Shay B. Cohen |
| 2024 | ACL | Large Language Models Relearn Removed Concepts. | Michelle Lo, Fazl Barez, Shay B. Cohen |
| 2024 | EMNLP | Layer by Layer: Uncovering Where Multi-Task Learning Happens in Instruction-Tuned Large Language Models. | Zheng Zhao, Yftah Ziser, Shay B. Cohen |
| 2024 | EMNLP | Interpreting Context Look-ups in Transformers: Investigating Attention-MLP Interactions. | Clement Neo, Shay B. Cohen, Fazl Barez |
| 2024 | EMNLP | Modeling News Interactions and Influence for Financial Market Prediction. | Mengyu Wang, Shay B. Cohen, Tiejun Ma |
| 2024 | EMNLP | Evaluating Automatic Metrics with Incremental Machine Translation Systems. | Guojun Wu, Shay B. Cohen, Rico Sennrich |
| 2024 | NAACL | LeanReasoner: Boosting Complex Logical Reasoning with Lean. | Dongwei Jiang, Marcio Fonseca, Shay B. Cohen |
| 2024 | NAACL | Are Large Language Model Temporally Grounded? | Yifu Qiu, Zheng Zhao, Yftah Ziser, Anna Korhonen, Edoardo Maria Ponti, Shay B. Cohen |
| 2024 | NAACL | Think While You Write: Hypothesis Verification Promotes Faithful Knowledge-to-Text Generation. | Yifu Qiu, Varun Embar, Shay B. Cohen, Benjamin Han |
| 2024 | SIGIR | CivilSum: A Dataset for Abstractive Summarization of Indian Court Decisions. | Manuj Malik, Zheng Zhao, Marcio Fonseca, Shrisha Rao, Shay B. Cohen |
| 2023 | ACL | The Larger they are, the Harder they Fail: Language Models do not Recognize Identifier Swaps in Python. | Antonio Valerio Miceli Barone, Fazl Barez, Shay B. Cohen, Ioannis Konstas |
| 2023 | ACL | DISCOSQA: A Knowledge Base Question Answering System for Space Debris based on Program Induction. | Paul Darm, Antonio Valerio Miceli Barone, Shay B. Cohen, Annalisa Riccardi |
| 2023 | EACL | BERT Is Not The Count: Learning to Match Mathematical Statements with Proofs. | Weixian Waylon Li, Yftah Ziser, Maximin Coavoux, Shay B. Cohen |
| 2023 | EACL | Gold Doesn't Always Glitter: Spectral Removal of Linear and Nonlinear Guarded Attribute Information. | Shun Shao, Yftah Ziser, Shay B. Cohen |
| 2023 | EMNLP | A Joint Matrix Factorization Analysis of Multilingual Representations. | Zheng Zhao, Yftah Ziser, Bonnie Webber, Shay B. Cohen |
| 2023 | EMNLP | AMR Parsing is Far from Solved: GrAPES, the Granular AMR Parsing Evaluation Suite. | Jonas Groschwitz, Shay B. Cohen, Lucia Donatelli, Meaghan Fowlie |
| 2023 | EMNLP | Detecting and Mitigating Hallucinations in Multilingual Summarisation. | Yifu Qiu, Yftah Ziser, Anna Korhonen, Edoardo Maria Ponti, Shay B. Cohen |
| 2023 | EMNLP | PMIndiaSum: Multilingual and Cross-lingual Headline Summarization for Languages in India. | Ashok Urlana, Pinzhen Chen, Zheng Zhao, Shay B. Cohen, Manish Shrivastava, Barry Haddow |
| 2022 | ACL | Co-training an Unsupervised Constituency Parser with Weak Supervision. | Nickil Maveli, Shay B. Cohen |
| 2022 | EMNLP | Factorizing Content and Budget Decisions in Abstractive Summarization of Long Documents. | Marcio Fonseca, Yftah Ziser, Shay B. Cohen |
| 2022 | EMNLP | Sentence-Incremental Neural Coreference Resolution. | Matt Grenander, Shay B. Cohen, Mark Steedman |
| 2022 | EMNLP | Abstractive Summarization Guided by Latent Hierarchical Document Structure. | Yifu Qiu, Shay B. Cohen |
| 2021 | ACL | A Closer Look into the Robustness of Neural Dependency Parsers Using Better Adversarial Examples. | Yuxuan Wang, Wanxiang Che, Ivan Titov, Shay B. Cohen, Zhilin Lei, Ting Liu |
| 2021 | EMNLP | Open-Domain Contextual Link Prediction and its Complementarity with Entailment Graphs. | Mohammad Javad Hosseini, Shay B. Cohen, Mark Johnson, Mark Steedman |
| 2021 | EMNLP | A Differentiable Relaxation of Graph Segmentation and Alignment for AMR Parsing. | Chunchuan Lyu, Shay B. Cohen, Ivan Titov |
| 2021 | EMNLP | A Root of a Problem: Optimizing Single-Root Dependency Parsing. | Milos Stanojevic, Shay B. Cohen |
| 2021 | NAACL | Text Generation from Discourse Representation Structures. | Jiangming Liu, Shay B. Cohen, Mirella Lapata |
| 2020 | ACL | Learning Dialog Policies from Weak Demonstrations. | Gabriel Gordon-Hall, Philip John Gorinski, Shay B. Cohen |
| 2020 | ACL | Machine Reading of Historical Events. | Or Honovich, Lucas Torroba Hennigen, Omri Abend, Shay B. Cohen |
| 2020 | ACL | Dscorer: A Fast Evaluation Metric for Discourse Representation Structure Parsing. | Jiangming Liu, Shay B. Cohen, Mirella Lapata |
| 2020 | EMNLP | The Role of Reentrancies in Abstract Meaning Representation Parsing. | Marco Damonte, Ida Szubert, Shay B. Cohen, Mark Steedman |
| 2020 | EMNLP | Multi-Step Inference for Reasoning Over Paragraphs. | Jiangming Liu, Matt Gardner, Shay B. Cohen, Mirella Lapata |
| 2020 | EMNLP | Lightweight, Dynamic Graph Convolutional Networks for AMR-to-Text Generation. | Yan Zhang, Zhijiang Guo, Zhiyang Teng, Wei Lu, Shay B. Cohen, Zuozhu Liu, Lidong Bing |
| 2020 | EMNLP | Reducing the Frequency of Hallucinated Quantities in Abstractive Summaries. | Zheng Zhao, Shay B. Cohen, Bonnie Webber |
| 2020 | ICLR | Compositional languages emerge in a neural iterated learning model. | Yi Ren, Shangmin Guo, Matthieu Labeau, Shay B. Cohen, Simon Kirby |
| 2020 | IJCAI | Learning Latent Forests for Medical Relation Extraction. | Zhijiang Guo, Guoshun Nan, Wei Lu, Shay B. Cohen |
| 2020 | IJCNLP | English-to-Chinese Transliteration with Phonetic Auxiliary Task. | Yuan He, Shay B. Cohen |
| 2019 | ACL | Duality of Link Prediction and Entailment Graph Induction. | Mohammad Javad Hosseini, Shay B. Cohen, Mark Johnson, Mark Steedman |
| 2019 | ACL | Discourse Representation Parsing for Sentences and Documents. | Jiangming Liu, Shay B. Cohen, Mirella Lapata |
| 2019 | ACL | Wide-Coverage Neural A* Parsing for Minimalist Grammars. | John Torr, Milos Stanojevic, Mark Steedman, Shay B. Cohen |
| 2019 | EMNLP | Experimenting with Power Divergences for Language Modeling. | Matthieu Labeau, Shay B. Cohen |
| 2019 | EMNLP | Semantic Role Labeling with Iterative Structure Refinement. | Chunchuan Lyu, Shay B. Cohen, Ivan Titov |
| 2019 | EMNLP | Partners in Crime: Multi-view Sequential Inference for Movie Understanding. | Nikos Papasarantopoulos, Lea Frermann, Mirella Lapata, Shay B. Cohen |
| 2019 | NAACL | Discontinuous Constituency Parsing with a Stack-Free Transition System and a Dynamic Oracle. | Maximin Coavoux, Shay B. Cohen |
| 2019 | NAACL | Structural Neural Encoders for AMR-to-text Generation. | Marco Damonte, Shay B. Cohen |
| 2019 | NAACL | Jointly Extracting and Compressing Documents with Summary State Representations. | Afonso Mendes, Shashi Narayan, Sebastio Miranda, Zita Marinho, Andr F. T. Martins, Shay B. Cohen |
| 2018 | AAAI | Canonical Correlation Inference for Mapping Abstract Scenes to Text. | Nikos Papasarantopoulos, Helen Jiang, Shay B. Cohen |
| 2018 | ACL | Stock Movement Prediction from Tweets and Historical Prices. | Yumo Xu, Shay B. Cohen |
| 2018 | ACL | Discourse Representation Structure Parsing. | Jiangming Liu, Shay B. Cohen, Mirella Lapata |
| 2018 | ACL | Document Modeling with External Attention for Sentence Extraction. | Shashi Narayan, Ronald Cardenas, Nikos Papasarantopoulos, Shay B. Cohen, Mirella Lapata, Jiangsheng Yu, Yi Chang |
| 2018 | COLING | Local String Transduction as Sequence Labeling. | Joana Ribeiro, Shashi Narayan, Shay B. Cohen, Xavier Carreras |
| 2018 | EMNLP | Privacy-preserving Neural Representations of Text. | Maximin Coavoux, Shashi Narayan, Shay B. Cohen |
| 2018 | EMNLP | Multilingual Clustering of Streaming News. | Sebastio Miranda, Arturs Znotins, Shay B. Cohen, Guntis Barzdins |
| 2018 | EMNLP | Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization. | Shashi Narayan, Shay B. Cohen, Mirella Lapata |
| 2018 | NAACL | Cross-Lingual Abstract Meaning Representation Parsing. | Marco Damonte, Shay B. Cohen |
| 2018 | NAACL | Abstract Meaning Representation for Paraphrase Detection. | Fuad Issa, Marco Damonte, Shay B. Cohen, Xiaohui Yan, Yi Chang |
| 2018 | NAACL | Ranking Sentences for Extractive Summarization with Reinforcement Learning. | Shashi Narayan, Shay B. Cohen, Mirella Lapata |
| 2017 | EACL | The SUMMA Platform Prototype. | Renars Liepins, Ulrich Germann, Guntis Barzdins, Alexandra Birch, Steve Renals, Susanne Weber, Peggy van der Kreeft, Herv Bourlard, Joo Prieto, Ondrej Klejch, Peter Bell, Alexandros Lazaridis, Afonso Mendes, Sebastian Riedel, Mariana S. C. Almeida, Pedro Balage, Shay B. Cohen, Tomasz Dwojak, Philip N. Garner, Andreas Giefer, Marcin Junczys-Dowmunt, Hina Imran, David Nogueira, Ahmed M. Ali, Sebastio Miranda, Andrei Popescu-Belis, Lesly Miculicich Werlen, Nikos Papasarantopoulos, Abiola Obamuyide, Clive Jones, Fahim Dalvi, Andreas Vlachos, Yang Wang, Sibo Tong, Rico Sennrich, Nikolaos Pappas, Shashi Narayan, Marco Damonte, Nadir Durrani, Sameer Khurana, Ahmed Abdelali, Hassan Sajjad, Stephan Vogel, David Sheppey, Chris Hernon, Jeff Mitchell |
| 2017 | EACL | An Incremental Parser for Abstract Meaning Representation. | Marco Damonte, Shay B. Cohen, Giorgio Satta |
| 2017 | EMNLP | Split and Rephrase. | Shashi Narayan, Claire Gardent, Shay B. Cohen, Anastasia Shimorina |
| 2016 | ACL | Optimizing Spectral Learning for Parsing. | Shashi Narayan, Shay B. Cohen |
| 2016 | AISTATS | Low-Rank Approximation of Weighted Tree Automata. | Guillaume Rabusseau, Borja Balle, Shay B. Cohen |
| 2016 | EMNLP | Semi-Supervised Learning of Sequence Models with Method of Moments. | Zita Marinho, Andr F. T. Martins, Shay B. Cohen, Noah A. Smith |
| 2016 | INLG | Paraphrase Generation from Latent-Variable PCFGs for Semantic Parsing. | Shashi Narayan, Siva Reddy, Shay B. Cohen |
| 2015 | CoNLL | A Coactive Learning View of Online Structured Prediction in Statistical Machine Translation. | Artem Sokolov, Stefan Riezler, Shay B. Cohen |
| 2015 | EMNLP | Conversation Trees: A Grammar Model for Topic Structure in Forums. | Annie Louis, Shay B. Cohen |
| 2015 | EMNLP | Diversity in Spectral Learning for Natural Language Parsing. | Shashi Narayan, Shay B. Cohen |
| 2015 | ICML | Coactive Learning for Interactive Machine Translation. | Artem Sokolov, Stefan Riezler, Shay B. Cohen |
| 2015 | NAACL | Lexical Event Ordering with an Edge-Factored Model. | Omri Abend, Shay B. Cohen, Mark Steedman |
| 2014 | ACL | Lexical Inference over Multi-Word Predicates: A Distributional Approach. | Omri Abend, Shay B. Cohen, Mark Steedman |
| 2014 | ACL | A Provably Correct Learning Algorithm for Latent-Variable PCFGs. | Shay B. Cohen, Michael Collins |
| 2014 | ACL | Spectral Unsupervised Parsing with Additive Tree Metrics. | Ankur P. Parikh, Shay B. Cohen, Eric P. Xing |
| 2014 | EMNLP | Latent-Variable Synchronous CFGs for Hierarchical Translation. | Avneesh Saluja, Chris Dyer, Shay B. Cohen |
| 2013 | ACL | The effect of non-tightness on Bayesian estimation of PCFGs. | Shay B. Cohen, Mark Johnson |
| 2013 | CoNLL | Spectral Learning of Refinement HMMs. | Karl Stratos, Alexander M. Rush, Shay B. Cohen, Michael Collins |
| 2013 | NAACL | Spectral Learning Algorithms for Natural Language Processing. | Shay B. Cohen, Michael Collins, Dean P. Foster, Karl Stratos, Lyle H. Ungar |
| 2013 | NAACL | Approximate PCFG Parsing Using Tensor Decomposition. | Shay B. Cohen, Giorgio Satta, Michael Collins |
| 2013 | NAACL | Experiments with Spectral Learning of Latent-Variable PCFGs. | Shay B. Cohen, Karl Stratos, Michael Collins, Dean P. Foster, Lyle H. Ungar |
| 2012 | ACL | Spectral Learning of Latent-Variable PCFGs. | Shay B. Cohen, Karl Stratos, Michael Collins, Dean P. Foster, Lyle H. Ungar |
| 2011 | EMNLP | Unsupervised Bilingual POS Tagging with Markov Random Fields. | Desai Chen, Chris Dyer, Shay B. Cohen, Noah A. Smith |
| 2011 | EMNLP | Unsupervised Structure Prediction with Non-Parallel Multilingual Guidance. | Shay B. Cohen, Dipanjan Das, Noah A. Smith |
| 2011 | EMNLP | Exact Inference for Generative Probabilistic Non-Projective Dependency Parsing. | Shay B. Cohen, Carlos Gmez-Rodrguez, Giorgio Satta |
| 2010 | ACL | Viterbi Training for PCFGs: Hardness Results and Competitiveness of Uniform Initialization. | Shay B. Cohen, Noah A. Smith |
| 2010 | NAACL | Variational Inference for Adaptor Grammars. | Shay B. Cohen, David M. Blei, Noah A. Smith |
| 2009 | ACL | Variational Inference for Grammar Induction with Prior Knowledge. | Shay B. Cohen, Noah A. Smith |
| 2009 | NAACL | Shared Logistic Normal Distributions for Soft Parameter Tying in Unsupervised Grammar Induction. | Shay B. Cohen, Noah A. Smith |
| 2008 | ICLP | Dynamic Programming Algorithms as Products of Weighted Logic Programs. | Shay B. Cohen, Robert J. Simmons, Noah A. Smith |
| 2007 | EMNLP | Joint Morphological and Syntactic Disambiguation. | Shay B. Cohen, Noah A. Smith |
| 2005 | IJCAI | Feature Selection Based on the Shapley Value. | Shay B. Cohen, Eytan Ruppin, Gideon Dror |