Team Ai
Datasetpublic

DatarrX/Myanmar-Written-Spoken-Parallel-Corpus

Myanmar Written-Spoken Parallel Corpus (MWSPC) Dataset Description Myanmar Written-Spoken Parallel Corpus (MWSPC) is a high-quality open-source dataset designed to bridge the gap between formal written Burmese and daily spoken Burmese. This dataset is crucial for building natural-sounding AI models that understand the linguistic nuances of the Myanmar language. Curated by: Khant Sint Heinn (Kalix Louis) Organization: DatarrX | ဒေတာ-အက်စ် Language: Burmese… See the full description on the dataset page: https://huggingface.co/datasets/DatarrX/Myanmar-Written-Spoken-Parallel-Corpus.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
6likes46downloads

DatarrX/Myanmar-Written-Spoken-Parallel-Corpus · main · files are served by the source, never re-hosted here