AigizK/bashkort_tts_dataset
Bashkort TTS Dataset The largest open dataset for speech synthesis in the Bashkir language β featuring multi-speaker recordings and speaking styles. π Dataset Overview Total audio files: 62,852 Speakers: 7 female, 1 male Speaking styles: friendly, question, neutral Languages: Bashkir Format: MP3 audio + transcription text π How It Was Collected Initial recording: A female voice actor recorded ~15 hours of speech in Bashkir. Voice cloning: Usingβ¦ See the full description on the dataset page: https://huggingface.co/datasets/AigizK/bashkort_tts_dataset.
3103
Bashkort TTS Dataset
The largest open dataset for speech synthesis in the Bashkir language β featuring multi-speaker recordings and speaking styles.
π Dataset Overview
- Total audio files: 62,852
- Speakers: 7 female, 1 male
- Speaking styles:
friendly,question,neutral - Languages: Bashkir
- Format: MP3 audio + transcription text
π How It Was Collected
- Initial recording: A female voice actor recorded \~15 hours of speech in Bashkir.
- Voice cloning: Using ElevenLabs voice cloning, 8 additional synthetic voices were generated β including a male voice.
- Multi-style dataset: All voices were aligned with transcripts, with three distinct speaking styles:
friendlyβ warm, conversational tonequestionβ interrogative intonationneutralβ balanced, plain delivery
