Team Ai
Datasetpublic

projecte-aina/CA-GL_Parallel_Corpus

Dataset Card for CA-GL Parallel Corpus Dataset Description Dataset Summary The CA-GL Parallel Corpus is a Catalan-Galician synthetic dataset parallel sentences created to support the use of co-official languages from Spain, such as Catalan and Galician, in NLP tasks, specifically Machine Translation. Supported Tasks and Leaderboards The dataset can be used to train Bilingual Machine Translation models between Galician and Catalan in… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/CA-GL_Parallel_Corpus.

sourceHugging Facecc-by-nc-sa-4.0updated 1y agoView on Hugging Face
1likes86downloads
16 commits on main
113d8e91y ago

Improve data source description

fdelucaf
4b7b4352y ago

Update README.md

mmarimon
5c4964a2y ago

Add parquet file description

fdelucaf
7cc69432y ago

Upload parquet

fdelucaf
ac6d97f2y ago

Update dataset card

fdelucaf
f516d433y ago

Update README.md

fdelucaf
5aea50a3y ago

Update README.md

fdelucaf
368b68b3y ago

Update README.md

fdelucaf
fe13d9d3y ago

Update README.md

fdelucaf
b1704513y ago

Upload 2 files

fdelucaf
877ccb63y ago

Update README.md

fdelucaf
5434d053y ago

Update README.md

fdelucaf
bfdbeb73y ago

Update README.md

fdelucaf
9825b293y ago

Update README.md

fdelucaf
3999f8e3y ago

Create README.md

fdelucaf
21175b03y ago

initial commit

fdelucaf