datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
librispeech_asr_dummyaudiofolder_two_configs_in_metadataaudiofolder_single_config_in_metadataaudiofolder_no_configs_in_metadataaudiofolder_two_configs_in_metadata_with_defaultdummy-audio-sampleslibrispeech_asr_demoaudio-testdummy-flac-single-exampleashraq-esc50-1-dog-example
Dataset Card for "ashraq-esc50-1-dog-example"
More Information needed
dailytalk-dummyfixtures_common_voiceaudio-testing
audio-testing
Overview
This is a small, open dataset designed for quick validation of audio-related pipelines and applications, especially for Text-to-Speech (TTS) and Speech-to-Text (STT) systems.
It provides a few short, diverse audio clips and corresponding text transcripts, allowing developers to verify input/output handling, audio processing, and transcription logic without downloading large datasets.
Contents
3 short audio samples (.mp3, .wav)… See the full description on the dataset page: https://huggingface.co/datasets/JacobLinCool/audio-testing.Vietnamese_ASR_TestingDataVoca-Human-Testing
Voca-Human-Testing
Voca-Human-Testing is a 252-record, audio-backed evaluation subset covering four
top-level companion capabilities and 14 second-level capabilities.
The subset was sampled from the four MMVC task files at approximately 10% per
second-level capability. Sampling jointly balanced available categorical metadata,
including source-dataset, subtype, label, style, tts_style, and voice.
Contents
Voca-Human-Testing.json: 252 complete benchmark records.… See the full description on the dataset page: https://huggingface.co/datasets/imuxsh/Voca-Human-Testing.testingtestingAnn_Testing
Dataset Description
This dataset.....
Issues Encountered & Solution
I encountered....
model_testingtesting-datasetzen-audiodataset_e2b39e7b-e0b5-4828-9013-24a841493a24dataset_d68847fc-a199-4be0-9cbb-c322c0d69d67TestingDatasetF5-TTS-Small_audio_testing_datasetreservation-voice-dataset-testing-t15Vietnamese_ASR_TestingData_Old
About
This dataset is for only ASR testing in Vietnamese.
We collect data from various sources.
This is the first version of the dataset.
testing123dataset_5badddee-48d8-4201-b7d8-eaf2417d9be9testing
