| 2025 | Can GPT models Follow Human Summarization Guidelines? A Study for Targeted Communication Goals. | Yongxin Zhou, Fabien Ringeval, Franois Portet |
| 2025 | Taming the Titans: A Survey of Efficient LLM Inference Serving. | Ranran Zhen, Juntao Li, Yixin Ji, Zhenlin Yang, Tong Liu, Qingrong Xia, Xinyu Duan, Zhefeng Wang, Baoxing Huai, Min Zhang |
| 2025 | Forecasting Conversation Derailments Through Generation. | Yunfan Zhang, Kathleen McKeown, Smaranda Muresan |
| 2025 | SWI: Speaking with Intent in Large Language Models. | Yuwei Yin, Eunjeong Hwang, Giuseppe Carenini |
| 2025 | PRICoT: Principle Retrieval and Injection from Inference Successes and Failures for CoT Improvement. | Yudai Yamazaki, Naoto Takeda, Yasutaka Nishimura, Kazushi Ikeda |
| 2025 | Benchmarking and Improving LVLMs on Event Extraction from Multimedia Documents. | Fuyu Xing, Zimu Wang, Wei Wang, Haiyang Zhang |
| 2025 | Frontmatter. | |
| 2025 | Analysing Reference Production of Large Language Models. | Chengzhao Wu, Guanyi Chen, Fahime Same, Tingting He |
| 2025 | Truth or Twist? Optimal Model Selection for Reliable Label Flipping Evaluation in LLM-based Counterfactuals. | Qianli Wang, Van Bach Nguyen, Nils Feldhus, Luis Felipe Villa-Arenas, Christin Seifert, Sebastian Mller, Vera Schmitt |
| 2025 | How (un)faithful are explainable LLM-based NLG metrics? | Alex Terentowicz, Mateusz Lango, Ondrej Dusek |
| 2025 | Input Matters: Evaluating Input Structure's Impact on LLM Summaries of Sports Play-by-Play. | Barkavi Sundararajan, Somayajulu Sripada, Ehud Reiter |
| 2025 | Live Football Commentary (LFC): A Large-Scale Dataset for Building Football Commentary Generation Models. | Taiga Someya, Tatsuya Ishigaki, Hiroya Takamura |
| 2025 | Do My Eyes Deceive Me? A Survey of Human Evaluations of Hallucinations in NLG. | Patrcia Schmidtov, Eduardo Cal, Simone Balloccu, Dimitra Gkatzia, Rudali Huidrom, Mateusz Lango, Fahime Same, Vilm Zouhar, Saad Mahamood, Ondrej Dusek |
| 2025 | Automated and Context-Aware Code Documentation Leveraging Advanced LLMs. | Swapnil Sharma Sarker, Tanzina Taher Ifty |
| 2025 | Mining Contextualized Visual Associations from Images for Creativity Understanding. | Ananya Sahu, Amith Ananthram, Kathleen McKeown |
| 2025 | Face the Facts! Evaluating RAG-based Pipelines for Professional Fact-Checking. | Daniel Russo, Stefano Menini, Jacopo Staiano, Marco Guerini |
| 2025 | LogitRouter: a novel Attention variant for reducing Myopic Routing in Mixture of Experts. | Felipe Rodrguez, Marcelo Mendoza |
| 2025 | Enhancing Named Entity Translation from Classical Chinese to Vietnamese in Traditional Vietnamese Medicine Domain: A Hybrid Masking and Dictionary-Augmented Approach. | Nhu Pham, Uyen Nguyen, Long H. B. Nguyen, Dien Dinh |
| 2025 | Are Multi-Agents the new Pipeline Architecture for Data-to-Text Systems? | Chinonso Cynthia Osuji, Brian Timoney, Mark Andrade, Thiago Castro Ferreira, Brian Davis |
| 2025 | Scaling Up Data-to-Text Generation to Longer Sequences: A New Dataset and Benchmark Results for Generation from Large Triple Sets. | Chinonso Cynthia Osuji, Simon Mille, Ornait O'Connell, Thiago Castro Ferreira, Anya Belz, Brian Davis |
| 2025 | Enhancing Coherence and Interestingness in Knowledge-Grounded Dialogue Generation. | Hiroki Onozeki, Michimasa Inaba |
| 2025 | FreshTab: Sourcing Fresh Data for Table-to-Text Generation Evaluation. | Kristna Onderkov, Ondrej Pltek, Zdenek Kasner, Ondrej Dusek |
| 2025 | Fine-Tuning, Prompting and RAG for Knowledge Graph-to-Russian Text Generation. How do these Methods generalise to Out-of-Distribution Data? | Anna Nikiforovskaya, William Soto Martinez, Evan Parker Kelly Chapple, Claire Gardent |
| 2025 | FinStat2SQL: A Text2SQL Pipeline for Financial Statement Analysis. | Quang Hung Nguyen, Phuong-Anh Trinh, Hung Phan Quoc Mai, Phong Tuan Trinh |
| 2025 | KDA: Knowledge Distillation Adapter for Cross-Lingual Transfer. | Ta-Bao Nguyen, Nguyen-Phuong Phan, Tung Le, Huy Tien Nguyen |