| 2026 | ACL | Arg-LLaDA: Argument Summarization via Large Language Diffusion Models and Sufficiency-Aware Refinement. | Hao Li, Yizheng Sun, Viktor Schlegel, Kailai Yang, Riza Batista-Navarro, Goran Nenadic |
| 2026 | ACL | Beyond Static Synthetic Noise: Assessing the Robustness of Large Language Models to Natural Context Variation in the Real World. | Yulong Wu, Viktor Schlegel, Riza Batista-Navarro |
| 2026 | EACL | Evaluation and LLM-Guided Learning of ICD Coding Rationales. | Mingyang Li, Viktor Schlegel, Tingting Mu, Wuraola Oyewusi, Kai Kang, Goran Nenadic |
| 2025 | AAAI | MEDSAGE: Enhancing Robustness of Medical Dialogue Summarization to ASR Errors with LLM-generated Synthetic Dialogues. | Kuluhan Binici, Abhinav Ramesh Kashyap, Viktor Schlegel, Andy T. Liu, Vijay Prakash Dwivedi, Thanh-Tung Nguyen, Xiaoxue Gao, Nancy F. Chen, Stefan Winkler |
| 2025 | ACL | uMedSum: A Unified Framework for Clinical Abstractive Summarization. | Aishik Nagar, Yutong Liu, Andy T. Liu, Viktor Schlegel, Vijay Prakash Dwivedi, Arun-Kumar Kaliya-Perumal, Guna Pratheep Kalanchiam, Yili Tang, Robby T. Tan |
| 2025 | CIKM | Evaluating Differentially Private Generation of Domain-Specific Text. | Yidan Sun, Viktor Schlegel, Srinivasan Kolumam Nandakumar, Iqra Zahid, Yuping Wu, Warren Del-Pinto, Goran Nenadic, Siew-Kei Lam, Jie Zhang, Anil A. Bharath |
| 2025 | ECAI | DocDiscNER: Enhanced Document-Level Discontinuous NER via Coordination Ellipses Resolution and Self-Consistency Decoding. | Areej Alhassan, Viktor Schlegel, Rina Carines Cabral, Riza Batista-Navarro, Soyeon Caren Han, Josiah Poon, Goran Nenadic |
| 2025 | EMNLP | Natural Context Drift Undermines the Natural Language Understanding of Large Language Models. | Yulong Wu, Viktor Schlegel, Riza Batista-Navarro |
| 2025 | ICML | BRIDGE: Bootstrapping Text to Control Time-Series Generation via Multi-Agent Iterative Optimization and Diffusion Modeling. | Hao Li, Yu-Hao Huang, Chang Xu, Viktor Schlegel, Renhe Jiang, Riza Batista-Navarro, Goran Nenadic, Jiang Bian |
| 2024 | ACL | M-QALM: A Benchmark to Assess Clinical Reading Comprehension and Knowledge Recall in Large Language Models via Question Answering. | Anand Subramanian, Viktor Schlegel, Abhinav Ramesh Kashyap, Thanh-Tung Nguyen, Vijay Prakash Dwivedi, Stefan Winkler |
| 2024 | ACL | Which Side Are You On? A Multi-task Dataset for End-to-End Argument Summarisation and Evaluation. | Hao Li, Yuping Wu, Viktor Schlegel, Riza Batista-Navarro, Tharindu Madusanka, Iqra Zahid, Jiayan Zeng, Xiaochi Wang, Xinran He, Yizhi Li, Goran Nenadic |
| 2024 | EACL | A Comprehensive Survey of Sentence Representations: From the BERT Epoch to the CHATGPT Era and Beyond. | Abhinav Ramesh Kashyap, Thanh-Tung Nguyen, Viktor Schlegel, Stefan Winkler, See-Kiong Ng, Soujanya Poria |
| 2024 | EMNLP | Seemingly Plausible Distractors in Multi-Hop Reasoning: Are Large Language Models Attentive Readers? | Neeladri Bhuiya, Viktor Schlegel, Stefan Winkler |
| 2023 | ACL | Do You Hear The People Sing? Key Point Analysis via Iterative Clustering and Abstractive Summarisation. | Hao Li, Viktor Schlegel, Riza Batista-Navarro, Goran Nenadic |
| 2023 | ACL | A Two-Stage Decoder for Efficient ICD Coding. | Thanh-Tung Nguyen, Viktor Schlegel, Abhinav Ramesh Kashyap, Stefan Winkler |
| 2023 | EMNLP | Argument mining as a multi-hop generative machine reading comprehension task. | Boyang Liu, Viktor Schlegel, Riza Batista-Navarro, Sophia Ananiadou |
| 2023 | IJCNLP | Are Machine Reading Comprehension Systems Robust to Context Paraphrasing? | Yulong Wu, Viktor Schlegel, Riza Batista-Navarro |
| 2022 | ACL | WLASL-LEX: a Dataset for Recognising Phonological Properties in American Sign Language. | Federico Tavella, Viktor Schlegel, Marta Romeo, Aphrodite Galata, Angelo Cangelosi |
| 2022 | EMNLP | Can Transformers Reason in Fragments of Natural Language? | Viktor Schlegel, Kamen V. Pavlov, Ian Pratt-Hartmann |
| 2022 | ICWSM | Towards Human-Centred Explainability Benchmarks For Text Classification. | Viktor Schlegel, Erick Mendez Guzman, Riza Batista-Navarro |
| 2022 | LREC | 'Am I the Bad One'? Predicting the Moral Judgement of the Crowd Using Pre-trained Language Models. | Areej Alhassan, Jinkai Zhang, Viktor Schlegel |
| 2022 | LREC | RaFoLa: A Rationale-Annotated Corpus for Detecting Indicators of Forced Labour. | Erick Mendez Guzman, Viktor Schlegel, Riza Batista-Navarro |
| 2022 | LREC | Incorporating Zoning Information into Argument Mining from Biomedical Literature. | Boyang Liu, Viktor Schlegel, Riza Batista-Navarro, Sophia Ananiadou |
| 2021 | AAAI | Semantics Altering Modifications for Evaluating Comprehension in Machine Reading. | Viktor Schlegel, Goran Nenadic, Riza Batista-Navarro |
| 2021 | EACL | Is the Understanding of Explicit Discourse Relations Required in Machine Reading Comprehension? | Yulong Wu, Viktor Schlegel, Riza Batista-Navarro |
| 2020 | LREC | A Framework for Evaluation of Machine Reading Comprehension Gold Standards. | Viktor Schlegel, Marco Valentino, Andr Freitas, Goran Nenadic, Riza Batista-Navarro |
| 2019 | IUI | Vajra: step-by-step programming with natural language. | Viktor Schlegel, Benedikt Lang, Siegfried Handschuh, Andr Freitas |