fahadqazi/sindhi-parallel-scripts
Sindhi Parallel Scripts Dataset Dataset Description This dataset is a parallel-script dataset for Sindhi, containing the same Sindhi sentences represented in multiple writing systems and transliteration schemes. The dataset contains four corresponding columns: devanagari: Sindhi sentences written in Devanagari script. This serves as the ground-truth source representation. persoarabic: Sindhi sentences represented in the Perso-Arabic script. khudabadi: Sindhi… See the full description on the dataset page: https://huggingface.co/datasets/fahadqazi/sindhi-parallel-scripts.
139
No card is published for this repository, or it could not be fetched from Hugging Face right now.
