Team Ai
Datasetpublic

legacy-datasets/multilingual_librispeech

Multilingual LibriSpeech (MLS) dataset is a large multilingual corpus suitable for speech research. The dataset is derived from read audiobooks from LibriVox and consists of 8 languages - English, German, Dutch, Spanish, French, Italian, Portuguese, Polish.

sourceHugging Facecc-by-4.0updated 2y agoView on Hugging Face
17likes204downloads

legacy-datasets/multilingual_librispeech · main · files are served by the source, never re-hosted here