Multimodal Latxa
Collection
Multimodal Large Language Models for Low-Resource Languages: A Case Study for Basque • 9 items • Updated
This model is an open Multimodal Large Language Model (MLLM) specifically developed for the Basque language. It adapts the Basque-instructed Latxa backbone to process both image and text inputs, enabling multimodal capabilities for a low-resource language.
clip-vit-large-patch14-336)This model is intended for general-purpose multimodal understanding and generation tasks in Basque and English. Typical use cases include:
The model was developed using a two-stage training procedure specifically adapted for low-resource language constraints.
If you use this model, please cite the following paper:
@misc{arana2025multimodallargelanguagemodels,
title={Multimodal Large Language Models for Low-Resource Languages: A Case Study for Basque},
author={Lukas Arana and Julen Etxaniz and Ander Salaberria and Gorka Azkune},
year={2025},
eprint={2511.09396},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2511.09396},
}