| 2025 | SMART: Simulated Students Aligned with Item Response Theory for Question Difficulty Prediction. | Alexander Scarlatos, Nigel Fernandez, Christopher Ormerod, Susan Lottridge, Andrew S. Lan |
| 2025 | Can Large Language Models Outperform Non-Experts in Poetry Evaluation? A Comparative Study Using the Consensual Assessment Technique. | Piotr Sawicki, Marek Grzes, Dan Brown, Fabrcio Ges |
| 2025 | JaCorpTrack: Corporate History Event Extraction for Tracking Organizational Changes. | Yuya Sawada, Hiroki Ouchi, Yuichiro Yasui, Hiroki Teranishi, Yuji Matsumoto, Taro Watanabe, Masayuki Ishii |
| 2025 | Train It and Forget It: Merge Lists are Unnecessary for BPE Inference in Language Models. | Tomohiro Sawada, Kartik Goyal |
| 2025 | Translation in the Hands of Many: Centering Lay Users in Machine Translation Interactions. | Beatrice Savoldi, Alan Ramponi, Matteo Negri, Luisa Bentivogli |
| 2025 | Mind the Inclusivity Gap: Multilingual Gender-Neutral Translation Evaluation with mGeNTE. | Beatrice Savoldi, Giuseppe Attanasio, Eleonora Cupin, Eleni Gkovedarou, Jania Hackenbuchner, Anne Lauscher, Matteo Negri, Andrea Piergentili, Manjinder Thind, Luisa Bentivogli |
| 2025 | polyBART: A Chemical Linguist for Polymer Property Prediction and Generative Design. | Anagha Savit, Harikrishna Sahu, Shivank Shukla, Wei Xiong, Rampi Ramprasad |
| 2025 | Auto prompting without training labels: An LLM cascade for product quality assessment in e-commerce catalogs. | Soham Satyadharma, Fatemeh Sheikholeslami, Swati Kaul, Aziz Umit Batur, Suleiman Ali Khan |
| 2025 | Proactive User Information Acquisition via Chats on User-Favored Topics. | Shiki Sato, Jun Baba, Asahi Hentona, Shinji Iwata, Akifumi Yoshimoto, Koichiro Yoshino |
| 2025 | M-Help: Using Social Media Data to Detect Mental Health Help-Seeking Signals. | MSVPJ Sathvik, Zuhair Hasan Shaik, Vivek Gupta |
| 2025 | Seeing Culture: A Benchmark for Visual Reasoning and Grounding. | Burak Satar, Zhixin Ma, Patrick Amadeus Irawan, Wilfried A. Mulyawan, Jing Jiang, Ee-Peng Lim, Chong-Wah Ngo |
| 2025 | Insights into using temporal coordinated behaviour to explore connections between social media posts and influence. | Elisa Sartori, Serena Tardelli, Maurizio Tesconi, Mauro Conti, Alessandro Galeazzi, Stefano Cresci, Giovanni Da San Martino |
| 2025 | Unsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement. | Gabriele Sarti, Vilm Zouhar, Malvina Nissim, Arianna Bisazza |
| 2025 | Fin-ExBERT: User Intent based Text Extraction in Financial Context using Graph-Augmented BERT and trainable Plugin. | Soumick Sarker, Abhijit Kumar Rai |
| 2025 | Mahānāma: A Unique Testbed for Literary Entity Discovery and Linking. | Sujoy Sarkar, Gourav Sarkar, Manoj Balaji Jagadeeshan, Jivnesh Sandhan, Amrith Krishna, Pawan Goyal |
| 2025 | Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression. | Sreetama Sarkar, Yue Che, Alex Gavin, Peter Anthony Beerel, Souvik Kundu |
| 2025 | Detecting Legal Citations in United Kingdom Court Judgments. | Holli Sargeant, Andreas stling, Mns Magnusson |
| 2025 | What's in a prompt? Language models encode literary style in prompt embeddings. | Raphal Sarfati, Haley Moller, Toni J. B. Liu, Nicolas Boull, Christopher J. Earls |
| 2025 | DIPLomA: Efficient Adaptation of Instructed LLMs to Low-Resource Languages via Post-Training Delta Merging. | Ixak Sarasua, Ander Corral, Xabier Saralegi |
| 2025 | Agentic-ToM: Cognition-Inspired Agentic Processing For Enhancing Theory of Mind Reasoning. | Sneheel Sarangi, Chetan Talele, Hanan Salam |
| 2025 | Mind the Gap: A Closer Look at Tokenization for Multiple-Choice Question Answering with LLMs. | Mario Sanz-Guerrero, Minh Duc Bui, Katharina von der Wense |
| 2025 | Mixing Inference-time Experts for Enhancing LLM Reasoning. | Soumya Sanyal, Tianyi Xiao, Xiang Ren |
| 2025 | Investigating Pedagogical Teacher and Student LLM Agents: Genetic Adaptation Meets Retrieval-Augmented Generation Across Learning Styles. | Debdeep Sanyal, Agniva Maiti, Umakanta Maharana, Dhruv Kumar, Ankur Mali, C. Lee Giles, Murari Mandal |
| 2025 | AutoCVSS: Assessing the Performance of LLMs for Automated Software Vulnerability Scoring. | Davide Sanvito, Giovanni Arriciati, Giuseppe Siracusano, Roberto Bifulco, Michele Carminati |
| 2025 | CAPE: Context-Aware Personality Evaluation Framework for Large Language Models. | Jivnesh Sandhan, Fei Cheng, Tushar Sandhan, Yugo Murawaki |