datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Automatic-Speech-Recognition
Vietnamese ASR Collection
A consolidated Vietnamese speech collection for Automatic Speech Recognition (ASR), hosted at Tran1312/Automatic-Speech-Recognition.
The collection integrates audio–transcript pairs from multiple Vietnamese speech datasets into a unified metadata and storage format. Audio samples are normalized to the filename convention:
sample_<id>.wav
and described using a common JSONL schema.
This repository should not be interpreted as a newly recorded speech… See the full description on the dataset page: https://huggingface.co/datasets/Tran1312/Automatic-Speech-Recognition.AutomaticSpeechRecognition_LJSpeech
Dataset Card for "AutomaticSpeechRecognition_LJSpeech"
More Information needed
flock-demo-automatic-speech-recognition-sectionsAutomaticSpeechRecognition_LibriSpeech-TestOther
Dataset Card for "AutomaticSpeechRecognition_LibriSpeech-TestOther"
More Information needed
automatic-speech-recognition-checkpoint-downloadsAutomaticSpeechRecognition_LibriSpeech-TestClean
Dataset Card for "AutomaticSpeechRecognition_LibriSpeech-TestClean"
More Information needed
Kinyarwanda-Automatic-Speech-Recognition-Track-A
