Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01hf-audio /open-asr-leaderboard-resultstabularn<1K2 likes13k downloads1d agoHugging Face02egcortes /asr-jargon-specialized-vocabulary A Dataset for Evaluating ASR on Specialized Vocabulary Novel synthetic datasets from the paper "A Dataset for Evaluating ASR on Specialized Vocabulary" (LREC 2026). Code and reproduction scripts: https://github.com/eduardogc8/ASR-Jargon-Dataset-Code Configs Config Language Description synthetic_terms_en English Utterances embedding entirely novel, 100% OOV, LLM-generated technical terms synthetic_terms_pt Portuguese Portuguese equivalent… See the full description on the dataset page: https://huggingface.co/datasets/egcortes/asr-jargon-specialized-vocabulary.audioautomatic-speech-recognition10K<n<100K0 likes763 downloads3mo agoHugging Face03ysdede /asr_benchmark_storetabularn<1K1 likes414 downloads3mo agoHugging Face04Shramadeepd /uyghur-ASR-dataset Uyghur ASR Corpus (Latin Transliteration) A speech corpus for Uyghur automatic speech recognition, with transcriptions in a case-sensitive Latin transliteration scheme. Approximately 23 hours of audio across 9,468 clips. Uyghur is a Turkic language spoken by roughly 10–12 million people. It is severely under-represented in open speech datasets, and this corpus is intended to support ASR research for the language. Dataset summary Language Uyghur (ug)… See the full description on the dataset page: https://huggingface.co/datasets/Shramadeepd/uyghur-ASR-dataset.audioautomatic-speech-recognition1K<n<10K0 likes312 downloads24d agoHugging Face05pavanmaddula /ASRD-Datasetgated Adversarial Surface-Form Robustness Dataset (ASRD) 💻 GitHub · 🤗 Dataset · 📦 Zenodo · 📝 Cite · 🛡️ Responsible Use News [2026/09] 🎉 Accepted at EvoRobust @ NeurIPS 2026, the NeurIPS 2026 Workshop on Self-Evolving Diversity-Driven Search for Robust AI Systems (Sydney, Australia). [2026/10] 📦 v1.0.0 released and archived on Zenodo (10.5281/zenodo.23103902). Paper: Quad-State Safety Evaluation of Open-Weight Large Language Models on Non-Canonical… See the full description on the dataset page: https://huggingface.co/datasets/pavanmaddula/ASRD-Dataset.texttext-generation1K<n<10K0 likes183 downloads8d agoHugging Face06sophia8888 /clipquill-asr-benchmark Measuring whisper-tiny vs whisper-base in a browser tab Word error rate, wall-clock timing, transfer size and peak memory for two quantised Whisper tiers running entirely client-side in a real Chrome window, with the scripts that produced every number. If you are building an in-browser transcription page, the two results worth knowing before you pick a model tier: On clean synthetic audio the two tiers tie. If that is all you test, you will conclude the tier does not matter… See the full description on the dataset page: https://huggingface.co/datasets/sophia8888/clipquill-asr-benchmark.tabularautomatic-speech-recognitionn<1K0 likes152 downloads21d agoHugging Face07s512757 /polish-tedx-asr-eval Polish-TEDx-ASR-Eval A dataset for evaluating automatic speech recognition (ASR) systems for Polish in the domain of TEDx public talks. Contains audio segments from Polish TEDx talks available on YouTube (CC BY-NC-ND 4.0) and synthetic speech generated with KugelAudio (MIT), with manually created and cross-verified transcriptions. Created as part of the course "Workshops on Evaluation of Speech Recognition Systems" (ZWESUI, AMU 2026) by Group 1. Statistics… See the full description on the dataset page: https://huggingface.co/datasets/s512757/polish-tedx-asr-eval.audioautomatic-speech-recognitionn<1K0 likes94 downloads4mo agoHugging Face08almaz-nlp /almaz-asr-roster The ALMAZ ASR Roster Archived at Zenodo: 10.5281/zenodo.22761871 (concept DOI, always resolves to the latest version). A curated catalog of Azerbaijani speech-to-text artifacts: corpora, models, services, benchmarks and tools. Companion to the ALMAZ Resource Roster, which does the same for text. Schema matches the text roster so the two join, plus three columns speech needs and text does not: hours, condition, and verified. The verified column The standard way a… See the full description on the dataset page: https://huggingface.co/datasets/almaz-nlp/almaz-asr-roster.tabularn<1K0 likes62 downloads25d agoHugging Face09bengaliAI /ben10-asr-results Ben-10 Regional ASR — public results Score rows for the maintainer-run Ben-10 regional dialect ASR leaderboard. Field Meaning model_id Hub id or slug model_url Link to weights / paper wer Corpus Word Error Rate on private ben-10-test (lower better) wer_by_region JSON map region → WER backend Decode stack used by maintainers scorer_commit / decode_commit Git SHAs in BengaliAI/reg-speech-aacl evaluated_at ISO date requested_by Who asked, or maintainer if… See the full description on the dataset page: https://huggingface.co/datasets/bengaliAI/ben10-asr-results.tabularautomatic-speech-recognitionn<1K0 likes61 downloads3mo agoHugging Face10molamin /Kinyarwanda_Engligh_Multilingual_ASRThis dataset was created from Mozilla's Common Voice dataset for the purposes of Multilingual ASR on Kinyarwanda and English. The dataset contains 3000 hours of multilingual training samples, 300 hours of validation samples and 200 of testing samples. text100K<n<1M0 likes53 downloads4y agoHugging Face11witcheer /sovereign-asr-bench Sovereign ASR Bench — RTX 5090 Local, self-hosted automatic speech recognition benchmarks on one RTX 5090 32GB. Part of the WITCHEER local-AI rig. Methodology that matters: load-once measurement (so RTFx times transcription, not model load), one shared text normalizer applied to every model output and reference, and micro-averaged WER (total errors / total reference words — the LibriSpeech standard). The board lives as data in board.csv (shown in the viewer). Board —… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/sovereign-asr-bench.tabularn<1K0 likes51 downloads4mo agoHugging Face12IbrahimDayax /somali-asr-synthetic-youtube Somali ASR Synthetic YouTube Dataset A Somali-language speech dataset derived from YouTube audio, intended for training and evaluating automatic speech recognition (ASR) and speech-to-text (STT) models. Transcriptions were generated synthetically (silver-standard) via ASR bootstrapping. Dataset Summary Split Samples train ~4,393 validation 200 test 100 Total ~4,693 Language: Somali (so) Audio format: WAV, 16 kHz, mono, 16-bit PCM Total… See the full description on the dataset page: https://huggingface.co/datasets/IbrahimDayax/somali-asr-synthetic-youtube.audioautomatic-speech-recognition1K<n<10K2 likes50 downloads4mo agoHugging Face13bijaykumarsingh /assamese-qwen-asr-benchmark Assamese Qwen ASR Benchmark & Reproduction Package Contains prediction outputs, character/word error metrics, duration-stratified analyses, generated figures, and bootstrap confidence intervals for: Architectural Inductive Bias Trumps Parameter Scale: Adapting Large Audio-Language Models to Low-Resource Assamese ASR Summary Table Model Architecture Parameters WER (%) 95% Bootstrap CI CER (%) Training Time Qwen3-ASR-1.7B ~2.0B 25.05 [23.87, 26.14] 9.95… See the full description on the dataset page: https://huggingface.co/datasets/bijaykumarsingh/assamese-qwen-asr-benchmark.tabular10K<n<100K0 likes49 downloads7d agoHugging Face14kvest /Swedia-ASR-Dataset Swedia ASR Dataset This repository contains a small Swedish ASR evaluation dataset based on speech transcriptions from Swedia 2000. It was assembled to compare automatic speech-recognition output against manually corrected reference transcriptions for Swedish dialectal speech. The dataset is useful for quick experiments with Swedish ASR systems, especially when you want to inspect recognition quality on spontaneous speech from different regions, speakers, ages, and genders.… See the full description on the dataset page: https://huggingface.co/datasets/kvest/Swedia-ASR-Dataset.tabularautomatic-speech-recognitionn<1K1 likes46 downloads5mo agoHugging Face15sidleal /CORAA-MUPE-ASRtabular100K<n<1M0 likes40 downloads2y agoHugging Face16holygleb /asr-correction-rutext1K<n<10K0 likes32 downloads14d agoHugging Face17uam-wmi-asr-eval-labs /2026-dwesui-g02-kulinarna DWESUI 2026 - Grupa 2 - kulinarna (PIEROGA) Robocza/archiwalna kopia zbioru ewaluacyjnego ASR zbudowanego przez studentow kursu Warsztaty z ewaluacji systemow rozpoznawania mowy (UAM WMI), edycja 2026, tryb dzienny. Zespol (atrybucja): Grupa 2 (DWESUI 2026) Zrodlo oryginalne: https://huggingface.co/datasets/s479246/dwesui-grupa-2-kulinarna Domena: kulinarna Licencja zrodla: nagrania YouTube CC-BY/CC-BY-SA + TTS Status: kopia publiczna w organizacji kursowej (zespół opublikował… See the full description on the dataset page: https://huggingface.co/datasets/uam-wmi-asr-eval-labs/2026-dwesui-g02-kulinarna.audioautomatic-speech-recognitionn<1K0 likes30 downloads2mo agoHugging Face18u042e /asr-spell-correction-synthetictext1K<n<10K0 likes29 downloads14d agoHugging Face19Haradrick228 /gendl-hw1-asr-correction-rutext1K<n<10K0 likes23 downloads6d agoHugging Face20SabrinaSadiekh /responses-and-asr-labels-small-models LLM Responses and ASR Labels — Small Models Model responses to harmful prompts, labelled by 4 LLM-as-judge guards.Companion dataset for the master's thesis ASR Signal Geometry: Dense Representations vs. SAE Features (HSE, 2025). Dataset composition N = 4 326 prompts per model, (no adversarial suffix). Two sources: Source N Description JailbreakBench () 100 Curated harmful behaviours Anthropic HH-RLHF red-team-attempts () 4 226 Red-team conversations… See the full description on the dataset page: https://huggingface.co/datasets/SabrinaSadiekh/responses-and-asr-labels-small-models.tabulartext-classification10K<n<100K0 likes22 downloads4mo agoHugging Face21raiyan007 /asr_bangla_2024audio100K<n<1M0 likes19 downloads2y agoHugging Face22archanatikayatray /ASRS-ChatGPTgated Dataset Summary The dataset contains a total of 9984 incident records and 9 columns. Some of the columns contain ground truth values whereas others contain information generated by ChatGPT based on the incident Narratives. The creation of this dataset is aimed at providing researchers with columns generated by using ChatGPT API which is not freely available. Dataset Structure The column names present in the dataset and their descriptions are provided below: Column… See the full description on the dataset page: https://huggingface.co/datasets/archanatikayatray/ASRS-ChatGPT.textzero-shot-classification1K<n<10K8 likes14 downloads3y agoHugging Face23Subu19 /Devnagari-ASRaudio1K<n<10K0 likes13 downloads2y agoHugging Face24asr-malayalam /Flores-subsettext1K<n<10K0 likes13 downloads2y agoHugging Face25Ephrem /Amharic_ASR_Datasettextn<1K0 likes12 downloads4y agoHugging Face26v4xsh /indian-names-asrtext100K<n<1M0 likes12 downloads5mo agoHugging Face27sylviali /EDEN_ASR_Data EDEN ASR Dataset A subset of this data was used to support the development of empathetic feedback modules in EDEN and its prior work. The dataset contains audio clips of native Mandarin speakers. The speakers conversed with a chatbot hosted on an English practice platform. 3081 audio clips from 613 conversations and 163 users remained after filtering. The filtering process removes audio clips containing only Mandarin, duplicates, and a subset of self-introductions from the users.… See the full description on the dataset page: https://huggingface.co/datasets/sylviali/EDEN_ASR_Data.text1K<n<10K2 likes11 downloads2y agoHugging Face28awacke1 /ASRLive.csvtextn<1K1 likes10 downloads4y agoHugging Face29rahmostafijurrahman /asr_bangla_2024audio100K<n<1M0 likes10 downloads9mo agoHugging Face30yulia774 /chuvash-asr-final Chuvash ASR1 Небольшой аудиодатасет для распознавания чувашской речи. Structure audio/ — аудиофайлы train.csv — метаданные с колонками: file_name text_chv client_id Columns file_name — относительный путь к аудиофайлу text_chv — расшифровка на чувашском языке client_id — идентификатор говорящего Example file_name,text_chv,client_id audio/00001.ogg,Салам,spk01 audio/00002.ogg,Ырӑ ир,spk01 Web demo (single page) Файл web_app.py… See the full description on the dataset page: https://huggingface.co/datasets/yulia774/chuvash-asr-final.audioautomatic-speech-recognition1K<n<10K0 likes10 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.