| 2026 | LREC | Are Social Biases in LLMs Consistent across Generative Tasks? A Case Study for Basque. | Muitze Zulaika, Xabier Saralegi, Julia Shershneva, Lia Gonzalez, Arkaitz Fullaondo |
| 2025 | COLING | BasqBBQ: A QA Benchmark for Assessing Social Biases in LLMs for Basque, a Low-Resource Language. | Xabier Saralegi, Muitze Zulaika |
| 2025 | EMNLP | DIPLomA: Efficient Adaptation of Instructed LLMs to Low-Resource Languages via Post-Training Delta Merging. | Ixak Sarasua, Ander Corral, Xabier Saralegi |
| 2025 | NAACL | Pipeline Analysis for Developing Instruct LLMs in Low-Resource Languages: A Case Study on Basque. | Ander Corral, Ixak Sarasua, Xabier Saralegi |
| 2024 | COLING | How Well Can BERT Learn the Grammar of an Agglutinative and Flexible-Order Language? The Case of Basque. | Gorka Urbizu, Muitze Zulaika, Xabier Saralegi, Ander Corral |
| 2024 | EACL | Morphology Aware Source Term Masking for Terminology-Constrained NMT. | Ander Corral, Xabier Saralegi |
| 2024 | EAMT | MULTILINGTOOL, Development of an Automatic Multilingual Subtitling and Dubbing System. | Xabier Saralegi, Ander Corral, Igor Leturia, Xabier Sarasola, Josu Murua, Iker Manterola, Itziar Cortes |
| 2024 | NAACL | XNLIeu: a dataset for cross-lingual NLI in Basque. | Maite Heredia, Julen Etxaniz, Muitze Zulaika, Xabier Saralegi, Jeremy Barnes, Aitor Soroa |
| 2023 | ACL | Scaling Laws for BERT in Low-Resource Settings. | Gorka Urbizu, Iaki San Vicente, Xabier Saralegi, Rodrigo Agerri, Aitor Soroa |
| 2023 | ACL | Not Enough Data to Pre-train Your Language Model? MT to the Rescue! | Gorka Urbizu, Iaki San Vicente, Xabier Saralegi, Ander Corral |
| 2022 | LREC | TANDO: A Corpus for Document-level Machine Translation. | Harritxu Gete, Thierry Etchegoyhen, David Ponce, Gorka Labaka, Nora Aranberri, Ander Corral, Xabier Saralegi, Igor Ellakuria, Maite Martn |
| 2022 | LREC | BasqueGLUE: A Natural Language Understanding Benchmark for Basque. | Gorka Urbizu, Iaki San Vicente, Xabier Saralegi, Rodrigo Agerri, Aitor Soroa |
| 2021 | ECIR | Fine-Tuning BERT for COVID-19 Domain Ad-Hoc IR by Using Pseudo-qrels. | Xabier Saralegi, Iaki San Vicente |
| 2020 | LREC | Give your Text Representation Models some Love: the Case for Basque. | Rodrigo Agerri, Iaki San Vicente, Jon Ander Campos, Ander Barrena, Xabier Saralegi, Aitor Soroa, Eneko Agirre |
| 2020 | LREC | Building a Task-oriented Dialog System for Languages with no Training Data: the Case for Basque. | Maddalen Lopez de Lacalle, Xabier Saralegi, Iaki San Vicente |
| 2016 | LREC | Evaluating Translation Quality and CLIR Performance of Query Sessions. | Xabier Saralegi, Eneko Agirre, Iaki Alegria |
| 2016 | LREC | Polarity Lexicon Building: to what Extent Is the Manual Effort Worth? | Iaki San Vicente, Xabier Saralegi |
| 2013 | CICLING | Analyzing the Sense Distribution of Concordances Obtained by Web as Corpus Approach. | Xabier Saralegi, Pablo Gamallo |
| 2013 | CICLING | Cross-Lingual Projections vs. Corpora Extracted Subjectivity Lexicons for Less-Resourced Languages. | Xabier Saralegi, Iaki San Vicente, Irati Ugarteburu |
| 2012 | LREC | Building a Basque-Chinese Dictionary by Using English as Pivot. | Xabier Saralegi, Iker Manterola, Iaki San Vicente |
| 2011 | EMNLP | Analyzing Methods for Improving Precision of Pivot Based Bilingual Dictionaries. | Xabier Saralegi, Iker Manterola, Iaki San Vicente |
| 2010 | ECIR | Estimating Translation Probabilities from the Web for Structured Queries on CLIR. | Xabier Saralegi, Maddalen Lopez de Lacalle |
| 2010 | LREC | Dictionary and Monolingual Corpus-based Query Translation for Basque-English CLIR. | Xabier Saralegi, Maddalen Lopez de Lacalle |
| 2004 | LREC | A XML-Based Term Extraction Tool for Basque. | Iaki Alegria, Antton Gurrutxaga, P. Lizaso, Xabier Saralegi, Sahats Ugartetxea, Ruben Urizar |