SparkAudio/voxbox
VoxBox This dataset is a curated collection of bilingual speech corpora annotated clean transcriptions and rich metadata incluing age, gender, and emotion. Dataset Structure . ├── audios/ │ └── aishell-3/ # Audio files (organised by sub-corpus) │ └── ... └── metadata/ ├── aishell-3.jsonl ├── casia.jsonl ├── commonvoice_cn.jsonl ├── ... └── wenetspeech4tts.jsonl # JSONL metadata files Each JSONL file… See the full description on the dataset page: https://huggingface.co/datasets/SparkAudio/voxbox.
Upload speaker_ids/vctk.txt with huggingface_hub
Upload speaker_ids/tess.txt with huggingface_hub
Upload speaker_ids/savee.txt with huggingface_hub
Upload speaker_ids/ravdess.txt with huggingface_hub
Upload speaker_ids/ncssd_r_zh.txt with huggingface_hub
Upload speaker_ids/ncssd_r_en.txt with huggingface_hub
Upload speaker_ids/ncssd_c_zh.txt with huggingface_hub
Upload speaker_ids/ncssd_c_en.txt with huggingface_hub
Upload speaker_ids/msp-podcast.txt with huggingface_hub
Upload speaker_ids/mls_english.txt with huggingface_hub
Upload speaker_ids/meld.txt with huggingface_hub
Upload speaker_ids/mead.txt with huggingface_hub
Upload speaker_ids/magicdata.txt with huggingface_hub
Upload speaker_ids/m3ed.txt with huggingface_hub
Upload speaker_ids/libritts_r.txt with huggingface_hub
Upload speaker_ids/librispeech.txt with huggingface_hub
Upload speaker_ids/jlcorpus.txt with huggingface_hub
Upload speaker_ids/iemocap.txt with huggingface_hub
Upload speaker_ids/hq-conversations.txt with huggingface_hub
Upload speaker_ids/hifi_tts.txt with huggingface_hub
Upload speaker_ids/expresso.txt with huggingface_hub
Upload speaker_ids/esd.txt with huggingface_hub
Upload speaker_ids/emov-db.txt with huggingface_hub
Upload speaker_ids/emns.txt with huggingface_hub
Upload speaker_ids/emilia_zh.txt with huggingface_hub
Upload speaker_ids/emilia_en.txt with huggingface_hub
Upload speaker_ids/dailytalk.txt with huggingface_hub
Upload speaker_ids/cremad.txt with huggingface_hub
Upload speaker_ids/commonvoice_en.txt with huggingface_hub
Upload speaker_ids/commonvoice_cn.txt with huggingface_hub
Upload speaker_ids/casia.txt with huggingface_hub
Upload speaker_ids/aishell-3.txt with huggingface_hub
Merge branch 'main' of https://huggingface.co/datasets/SparkAudio/voxbox
upload done
Upload audios/wenetspeech4tts/wenetspeech4tts_0088.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0087.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0086.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0085.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0084.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0083.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0082.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0081.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0080.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0079.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0078.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0077.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0076.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0075.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0074.tar.gz with huggingface_hub
Upload audios/wenetspeech4tts/wenetspeech4tts_0073.tar.gz with huggingface_hub
