Team Ai
Datasetpublic

AigizK/bashkort_tts_dataset

Bashkort TTS Dataset The largest open dataset for speech synthesis in the Bashkir language β€” featuring multi-speaker recordings and speaking styles. πŸ“Š Dataset Overview Total audio files: 62,852 Speakers: 7 female, 1 male Speaking styles: friendly, question, neutral Languages: Bashkir Format: MP3 audio + transcription text πŸŽ™ How It Was Collected Initial recording: A female voice actor recorded ~15 hours of speech in Bashkir. Voice cloning: Using… See the full description on the dataset page: https://huggingface.co/datasets/AigizK/bashkort_tts_dataset.

sourceHugging Facecc-by-4.0updated 1y agoView on Hugging Face
3likes103downloads
Dataset Card

Bashkort TTS Dataset

The largest open dataset for speech synthesis in the Bashkir language β€” featuring multi-speaker recordings and speaking styles.

πŸ“Š Dataset Overview

  • β€”Total audio files: 62,852
  • β€”Speakers: 7 female, 1 male
  • β€”Speaking styles: friendly, question, neutral
  • β€”Languages: Bashkir
  • β€”Format: MP3 audio + transcription text

πŸŽ™ How It Was Collected

  1. 1.Initial recording: A female voice actor recorded \~15 hours of speech in Bashkir.
  2. 2.Voice cloning: Using ElevenLabs voice cloning, 8 additional synthetic voices were generated β€” including a male voice.
  3. 3.Multi-style dataset: All voices were aligned with transcripts, with three distinct speaking styles:
  • β€”friendly β€” warm, conversational tone
  • β€”question β€” interrogative intonation
  • β€”neutral β€” balanced, plain delivery