datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
audio_benchmarkswhisper-browser-benchmarks
whisper-browser-benchmarks
Measurements from a Whisper transcription pipeline running entirely inside a
browser tab: which audio and video containers the browser will actually decode,
how accurate the smallest usable Whisper size is on clean synthetic speech, how
long transcription takes relative to the length of the clip, what the first
load pulls over the wire, and what happens to clips longer than the model's
30-second window.
Everything here was measured, not quoted from a… See the full description on the dataset page: https://huggingface.co/datasets/ruanjiange/whisper-browser-benchmarks.telephony-benchmarks-databenchmark-sisdrBN_ASR_benchmarksecho-tts-en-benchmarks-v1
Echo-TTS English Benchmarks v1
Описание
Датасет содержит результаты бенчмарка модели Echo-TTS на 720 предложениях из Harvard Sentences.
Каждое предложение озвучено 14 спикерами — 5 встроенных голосов + 9 голосов клонированных из KaniTTS-2.
Структура
Поле
Тип
Описание
text
string
Текст предложения
echo_tts_af_bella
audio
Голос AF Bella (44100 Hz)
echo_tts_af_heart
audio
Голос AF Heart (44100 Hz)
echo_tts_am_fenrir
audio
Голос AM Fenrir (44100… See the full description on the dataset page: https://huggingface.co/datasets/data-lab-voice/echo-tts-en-benchmarks-v1.benchmark-speech-f0benchmarks
