Team Ai
Datasetpublic

projecte-aina/ES-OC_Parallel_Corpus

Dataset Card for ES-OC Parallel Corpus Dataset Summary The ES-OC Parallel Corpus is a Spanish-Aranese dataset created to support the use of under-resourced languages from Spain, such as Aranese, in NLP tasks, specifically Machine Translation. Supported Tasks and Leaderboards The dataset can be used to train Bilingual Machine Translation models between Aranese and Spanish in any direction, as well as Multilingual Machine Translation models.… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/ES-OC_Parallel_Corpus.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
2likes61downloads
settings

This repository belongs to projecte-aina on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameES-OC_Parallel_Corpus
visibilitypublic
licencecc-by-sa-4.0
gatedno
ownerprojecte-aina
Account settings
projecte-aina/ES-OC_Parallel_Corpus · Team Ai