datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Processed_TTS_Multilingual_Data
Processed TTS Multilingual Data
Validated and quality-checked multilingual speech datasets for TTS training, covering 12+ Indian languages.
Datasets Included
Subset
Samples
Hours
Description
indic_voices_r
239,684
548.8h
Indic Voices_R — IVR recordings
rasa
201,509
361.2h
RASA — read speech (wiki, conv, book, news)
indictts_iitm
155,236
253.6h
Indic TTS (IIT Madras) — studio TTS recordings at 48kHz
Total
596,429
1,163.6h
Languages… See the full description on the dataset page: https://huggingface.co/datasets/PalakEngineerMaster/Processed_TTS_Multilingual_Data.ViSEC-processed
ViSEC Processed
Processed ViSEC speaker audio generated for the Meddies ASR collection.
Contents
processed_audio_by_id/: 147 WAV files named by speaker id.
metadata.csv: per-speaker metadata with duration, clip count, emotion coverage, and source-duration summary.
Schema
metadata.csv contains:
speaker_id: integer speaker identifier.
output_path: relative path to the processed WAV file.
duration_seconds: duration of the processed audio file.… See the full description on the dataset page: https://huggingface.co/datasets/Meddies/ViSEC-processed.
