| 2026 | Instruction-Tuned Urdu LLMs: Efficient Adaptation of Llama Models and Evaluation Resources for Urdu. | Munief Hassan Tahir, Sana Shams, Sarmad Hussain, Miriam Butt |
| 2026 | Automatic Speech Recognition for Documenting Endangered Languages: Case Study of Ikema Miyakoan. | Chihiro Taguchi, Yukinori Takubo, David Chiang |
| 2026 | The Spectrum of Sentiment: Optimistic, Pessimistic, and Neutral Voices in Online Depression Discourse. | Stefana Arina Tabusca, Ana-Maria Bucur, Liviu P. Dinu |
| 2026 | Fill-in-the-Blanks: Automatic Generation and Evaluation of Language Models' Pseudonyms for English and Swedish Texts. | Maria Irena Szawerna, Jacob Lee Suchardt |
| 2026 | AusKidTalk: Developing Transcription Guidelines for Continuous Australian English Child Speech. | Tnde Szalay, Zheng Nan, Renata Huang, Mostafa Shahin, Tharmakulasingam Sirojan, Kirrie J. Ballard, Beena Ahmed |
| 2026 | How Much Data for Stable Formant Values? Pipeline for Convergence Detection Based on Read Speech. | Kayla Sward, Johan Sjons, Axel G. Ekstrm |
| 2026 | Is There Anything More Deceptive than an Obvious Fact? Investigating Implicitness in User-Generated Argumentative Text. | Ekaterina Sviridova, Elena Cabrio, Serena Villata |
| 2026 | Multilingual KokoroChat: A Multi-LLM Ensemble Translation Method for Creating a Multilingual Counseling Dialogue Dataset. | Ryoma Suzuki, Zhiyang Qi, Michimasa Inaba |
| 2026 | AnswerCarefully: Creating a Dataset for LLM Safety in Japanese. | Hisami Suzuki, Satoru Katsumata, Takashi Kodama, Tetsuro Takahashi, Kouta Nakayama, Satoshi Sekine |
| 2026 | Low-Rank Compression of Language Models via Differentiable Rank Selection. | Sidhant Sundrani, Francesco Tudisco, Pasquale Minervini |
| 2026 | An Exploration-Analysis-Disambiguation Reasoning Framework for Word Sense Disambiguation with Low-Parameter LLMs. | Deshan Koshala Sumanathilaka, Nicholas Micallef, Julian Hough |
| 2026 | Mapping Liberty Metaphors across Cultures and Time. | Sidney Suen, Rui Mao, Kenneth Kwok, Erik Cambria |
| 2026 | AURORA Model of Formant-to-tongue Inversion for Didactic and Clinical Applications. | Patrycja Strycharczuk, Sam Kirkham |
| 2026 | Śmigiel Dataset: Laying Foundations for Investigating Machine-Generated Text Detection in Polish. | Jakub Strebeyko, Alina Wrblewska, Piotr Przybyla |
| 2026 | Automatic Suggestions Help Extending Eventive Ontology: A Case Study on SynSemClass. | Jana Strakov, Eva Fuckov, Zdenka Uresov, Jan Hajic |
| 2026 | Spotlights and Blindspots: Evaluating Machine-Generated Text Detection. | Kevin Stowe, Kailash Patil |
| 2026 | Conversion of the Clark Hall Dictionary of Old English to TEI with RDF: An End-to-end Pipeline for Lexicographic Resource Retrodigitization. | Sergei Stoliarov, Maxim Ionov, Anas Fahad Khan, Marina Buzzoni, Francesca Frontini |
| 2026 | SlovKE: A Large-Scale Dataset and LLM Evaluation for Slovak Keyphrase Extraction. | David Stevank, Marek Suppa |
| 2026 | Modeling the Memory-Surprisal Trade-Off over Time: Communicative Efficiency Decreases with Lexico-Grammatical Change in Scientific English. | Julius Steuer, Marie-Pauline Krielke, Stefania Degaetano-Ortlieb, Elke Teich, Dietrich Klakow |
| 2026 | SEEM-CZ: Annotation and Classification of Epistemic Markers in Czech. | Barbora Stepnkov, Michal Novk, Toms Musil, Lucie Polkov |
| 2026 | Synthetic Instruction Generation for Low-Resource Nordic Languages: Viability and Limitations in LLM Instruction-Tuning. | Mathias Stenlund, Annika Simonsen, Lars Bungum, Jan Ebert, Jiangtao Wang, Oleg Filatov, Hemanadhan Myneni, Morris Riedel, Hafsteinn Einarsson |
| 2026 | STRUDEL: Unrolling a Benchmark for Evaluating Vision-Language Models on Structured Diagram Understanding across Domains. | Daniel Steinigen, Lucie Flek, Sebastian Houben |
| 2026 | Scoring the Translation: On Target Automatic Keyword-Based Evaluation of Machine Translation in the Sports Domain. | Steinr Steingrmsson, Einar Freyr Sigursson |
| 2026 | From CHAT to Coded CoNLL-U: A Reproducible Pipeline for the Syntactic Annotation and Querying of Child Language Data. | Achim Stein |
| 2026 | Cross-Corpus CEFR Classification through Artificial Learners Perplexities. | Bernardo Stearns, John P. McCrae, Thomas Gaillat |