| 2026 | SIGIR | Towards Emotional Intelligence in Conversational AI: How Well Can LLMs Recognise Emotion in Conversations? | Islam A. Hassan, Yvette Graham |
| 2024 | ECAI | REFINE-LM: Mitigating Language Model Stereotypes via Reinforcement Learning. | Rameez Qureshi, Nam Es-Sebbani, Luis Galrraga, Yvette Graham, Miguel Couceiro, Zied Bouraoui |
| 2024 | NAACL | ALoRA: Allocating Low-Rank Adaptation for Fine-tuning Large Language Models. | Zequan Liu, Jiawen Lyn, Wei Zhu, Xing Tian, Yvette Graham |
| 2023 | ACL | Semantic-Aware Dynamic Retrospective-Prospective Reasoning for Event-Level Video Question Answering. | Chenyang Lyu, Tianbo Ji, Yvette Graham, Jennifer Foster |
| 2023 | ACL | Exploiting Rich Textual User-Product Context for Improving Personalized Sentiment Analysis. | Chenyang Lyu, Linyi Yang, Yue Zhang, Yvette Graham, Jennifer Foster |
| 2023 | EMNLP | Do Stochastic Parrots have Feelings Too? Improving Neural Detection of Synthetic Text via Emotion Recognition. | Alan Cowap, Yvette Graham, Jennifer Foster |
| 2022 | ACL | Achieving Reliable Human Assessment of Open-Domain Dialogue Systems. | Tianbo Ji, Yvette Graham, Gareth J. F. Jones, Chenyang Lyu, Qun Liu |
| 2022 | CBMI | An Exploration into the Benefits of the CLIP model for Lifelog Retrieval. | Ly-Duyen Tran, Naushad Alam, Yvette Graham, Linh Khanh Vo, Nghiem Tuong Diep, Binh T. Nguyen, Liting Zhou, Cathal Gurrin |
| 2022 | ICIP | Evaluation of Automatically Generated Video Captions Using Vision and Language Models. | Luis Lebron, Yvette Graham, Noel E. O'Connor, Kevin McGuinness |
| 2022 | LREC | BERTHA: Video Captioning Evaluation Via Transfer-Learned Human Assessment. | Luis Lebron, Yvette Graham, Kevin McGuinness, Konstantinos Kouramas, Noel E. O'Connor |
| 2021 | EMNLP | Improving Unsupervised Question Answering via Summarization-Informed Question Generation. | Chenyang Lyu, Lifeng Shang, Yvette Graham, Jennifer Foster, Xin Jiang, Qun Liu |
| 2020 | CHIIR | Contrasting Human Opinion of Non-factoid Question Answering with Automatic Evaluation. | Tianbo Ji, Yvette Graham, Gareth J. F. Jones |
| 2020 | COLING | Improving Document-Level Sentiment Analysis with User and Product Context. | Chenyang Lyu, Jennifer Foster, Yvette Graham |
| 2020 | ECIR | Incorporating Context and Knowledge for Better Sentiment Analysis of Narrative Text. | Chenyang Lyu, Tianbo Ji, Yvette Graham |
| 2020 | EMNLP | Assessing Human-Parity in Machine Translation on the Segment Level. | Yvette Graham, Christian Federmann, Maria Eskevich, Barry Haddow |
| 2020 | EMNLP | Statistical Power and Translationese in Machine Translation Evaluation. | Yvette Graham, Barry Haddow, Philipp Koehn |
| 2019 | EMNLP | The Second Multilingual Surface Realisation Shared Task (SR'19): Overview and Evaluation Results. | Simon Mille, Anja Belz, Bernd Bohnet, Yvette Graham, Leo Wanner |
| 2019 | MMM | Exploring the Impact of Training Data Bias on Automatic Generation of Video Captions. | Alan F. Smeaton, Yvette Graham, Kevin McGuinness, Noel E. O'Connor, Sen Quinn, Eric Arazo Sanchez |
| 2018 | AAAI | Translating Pro-Drop Languages With Reconstruction Models. | Longyue Wang, Zhaopeng Tu, Shuming Shi, Tong Zhang, Yvette Graham, Qun Liu |
| 2017 | EACL | Improving Evaluation of Document-level Machine Translation Quality Estimation. | Yvette Graham, Qingsong Ma, Timothy Baldwin, Qun Liu, Carla Parra Escartn, Carolina Scarton |
| 2017 | EMNLP | Further Investigation into Reference Bias in Monolingual Evaluation of Machine Translation. | Qingsong Ma, Yvette Graham, Timothy Baldwin, Qun Liu |
| 2016 | COLING | Is all that Glitters in Machine Translation Quality Estimation really Gold? | Yvette Graham, Timothy Baldwin, Meghan Dowling, Maria Eskevich, Teresa Lynn, Lamia Tounsi |
| 2016 | NAACL | Achieving Accurate Conclusions in Evaluation of Automatic Machine Translation Metrics. | Yvette Graham, Qun Liu |
| 2015 | ACL | Improving Evaluation of Machine Translation Quality Estimation. | Yvette Graham |
| 2015 | EMNLP | Re-evaluating Automatic Summarization with BLEU and 192 Shades of ROUGE. | Yvette Graham |
| 2015 | NAACL | Accurate Evaluation of Segment-level Machine Translation Metrics. | Yvette Graham, Timothy Baldwin, Nitika Mathur |
| 2014 | EACL | Is Machine Translation Getting Better over Time? | Yvette Graham, Timothy Baldwin, Alistair Moffat, Justin Zobel |
| 2014 | EMNLP | Testing for Significance of Increased Correlation with Human Judgment. | Yvette Graham, Timothy Baldwin |
| 2008 | EAMT | Packed rules for automatic transfer-rule induction. | Yvette Graham, Josef van Genabith |