Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01datapointai /text-to-speech-human-preferences-315kgated Text-to-speech human preferences: 315K votes across 15 models This gated dataset contains the evaluation record behind Datapoint Audio Bench: 315,000 eligible pairwise votes comparing 15 text-to-speech models in a complete round-robin over 300 English prompts. The prompt set covers eight practical voice-agent categories, and every generated sample is included as a typed audio record. The source evaluation collected 357,651 completed responses. The published benchmark excluded… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-to-speech-human-preferences-315k.audiotext-to-speech100K<n<1M39 likes418 downloads1mo agoHugging Face02Tamazight-NLP /Tamazight-Speech-to-Arabic-Text Tamazight-Arabic Speech Recognition Dataset This is the Tamazight-NLP organization-hosted version of the Tamazight-Arabic Speech Recognition Dataset. This dataset contains ~15.5 hours of Tamazight (Tachelhit dialect) speech paired with Arabic transcriptions, designed for automatic speech recognition (ASR) and speech-to-text translation tasks. Dataset Details Total Examples: 20,344 audio segments Training Set: 18,309 examples (~8.9GB) Test Set: 2,035 examples (~992MB)… See the full description on the dataset page: https://huggingface.co/datasets/Tamazight-NLP/Tamazight-Speech-to-Arabic-Text.audiotranslation10K<n<100K11 likes329 downloads2y agoHugging Face03Appenlimited /700h-tr-turkish-text-to-speechaudioautomatic-speech-recognition1K<n<10K21 likes328 downloads1y agoHugging Face04adiren7 /darija_speech_to_textaudioautomatic-speech-recognition10K<n<100K17 likes312 downloads2y agoHugging Face05lubobill1990 /speech_to_text_yixing_dialectaudio10K<n<100K0 likes289 downloads1y agoHugging Face06X-lord /Dataset-Text-To-Speech-Indonesia 🎵 Dataset Audio Bahasa Indonesia Dataset audio berkualitas tinggi untuk Text-to-Speech (TTS) bahasa Indonesia. Dibuat oleh : Muhammad Arief, S.Kom.Universitas Muhammadiyah SorongTeknik Informatika 2020 📊 Spesifikasi Teknis Parameter Nilai Satuan Total Durasi 16.38 jam Jumlah Segmen 4531 file Durasi Rata-rata 13.01 detik Sample Rate KHz 22 kHz Sample Rate Hz 22000 Hz Bit Depth PCM_16 PCM Format wav Lossless 🔄 Urutan Pengolahan… See the full description on the dataset page: https://huggingface.co/datasets/X-lord/Dataset-Text-To-Speech-Indonesia.audiotext-to-speech1K<n<10K7 likes250 downloads9mo agoHugging Face07danielrosehill /Speech-To-Text-System-Prompts-2 Speech To Text System Prompt Library This repository provides a collection of system prompts designed to transform and refine text captured using speech-to-text technologies. By passing STT outputs through large language models with these specialized prompts, you can achieve cleaner, more structured, and purpose-specific text formats. 📋 The Idea Here is the basic implementation. I don't pretend that this is the stuff of high AI engineering. But it does create quite… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/Speech-To-Text-System-Prompts-2.imagen<1K2 likes220 downloads1y agoHugging Face08crtvai /arabic_speech_to_text_20241218_144737_gkopimaudion<1K6 likes178 downloads2y agoHugging Face09andrewatef /Arabic-Text-to-Speechaudio10K<n<100K4 likes172 downloads1y agoHugging Face10crtvai /arabic_speech_to_text_20241219_205753_x4mhwqaudio1K<n<10K0 likes159 downloads2y agoHugging Face11BrunoHays /darija-speech-to-text Speech To Text Darija dataset Reupload of adiren7/darija_speech_to_text audioautomatic-speech-recognition1K<n<10K8 likes129 downloads2y agoHugging Face12EMINES /Tamazight-Speech-to-Arabic-Text Tamazight-Arabic Speech Recognition Dataset Overview This is the EMINES organization-hosted version of the Tamazight-Arabic Speech Recognition Dataset, synchronized with the original dataset. It contains ~15.5 hours of Tamazight speech (Tachelhit dialect) paired with Arabic transcriptions, designed for developing ASR and translation systems. Quick Start from datasets import load_dataset # Load the dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/EMINES/Tamazight-Speech-to-Arabic-Text.audioautomatic-speech-recognition10K<n<100K10 likes128 downloads2y agoHugging Face13adiren7 /darija_to_french_speech_to_textaudion<1K9 likes112 downloads2y agoHugging Face14crtvai /arabic_speech_to_text_20241219_203218_lp9vcvaudio1K<n<10K0 likes109 downloads2y agoHugging Face15baohuynhbk14 /vietnamese-speech-to-text-preprocessed-whisper-large-v31K<n<10K0 likes103 downloads3y agoHugging Face16baohuynhbk14 /vietnamese-speech-to-text-preprocessed-whisper-medium1K<n<10K1 likes91 downloads3y agoHugging Face17crtvai /arabic_speech_to_text_20241222_184338_sngqcgaudion<1K0 likes90 downloads2y agoHugging Face18crtvai /arabic_speech_to_text_20241224_092917_r6wod1audio1K<n<10K0 likes87 downloads2y agoHugging Face19crtvai /arabic_speech_to_text_20241224_134331_eiicpwaudio1K<n<10K1 likes84 downloads2y agoHugging Face20ANANDHU-SCT /Speech-to-texttextn<1K5 likes82 downloads3y agoHugging Face21the-vedantic-coder /text-to-speech-en-IN-checkpoint0 likes76 downloads10mo agoHugging Face22pujanpaudel /nepali_speech_to_text Nepali Speech-to-Text Dataset This repository contains a dataset for Automatic Speech Recognition (ASR) in the Nepali language. The dataset is designed for supervised learning tasks and includes audio files along with their corresponding transcriptions. The audio samples have been collected from various open-source platforms and other publicly available sources on the internet. Each audio file has an average length of 15 seconds and has been converted into a consistent WAV format… See the full description on the dataset page: https://huggingface.co/datasets/pujanpaudel/nepali_speech_to_text.audioautomatic-speech-recognition1K<n<10K1 likes71 downloads2y agoHugging Face23crtvai /arabic_speech_to_text_20241224_140146_q7hs7maudio1K<n<10K0 likes68 downloads2y agoHugging Face24Charif-Ayfarah /Afar-language-text-to-speech-TTS Usage This dataset is designed to support the development of Text-to-Speech (TTS) systems for the Afar language. It can be integrated into web applications, mobile apps, desktop software, or other platforms that require natural-sounding Afar voice synthesis or accurate spoken language recognition. For applications involving virtual avatars or voice personas, the following culturally appropriate voice names are recommended: Female Voices: Emeli, Hanaawi, Kareera, Laysani, Kulsuma… See the full description on the dataset page: https://huggingface.co/datasets/Charif-Ayfarah/Afar-language-text-to-speech-TTS.text-to-speech1 likes66 downloads10mo agoHugging Face25crtvai /arabic_speech_to_text_20241224_130013_z3rvjeaudio1K<n<10K0 likes57 downloads2y agoHugging Face26gunjan12 /speech-to-text9 likes54 downloads3y agoHugging Face27crtvai /arabic_speech_to_text_20241224_135643_72lw7raudio1K<n<10K0 likes54 downloads2y agoHugging Face28pavi1561 /Sinhala_speech_to_textaudioautomatic-speech-recognition10K<n<100K1 likes51 downloads2y agoHugging Face29crtvai /arabic_speech_to_text_20241218_174023_sd5hyyaudio1K<n<10K0 likes38 downloads2y agoHugging Face30phuhuyqhqb /speech-to-textaudio10K<n<100K0 likes36 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.