Team Ai
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Myrtle /CAIMAN-ASR-BackgroundNoise Dataset Card for Myrtle/CAIMAN-ASR-BackgroundNoise This dataset provides background noise audio, suitable for noise augmentation while training Myrtle.ai's CAIMAN-ASR models. Dataset Details Dataset Description Curated by: Myrtle.ai License: Myrtle.ai's modifications to the source data are licensed under the CC BY 4.0 license. Some of the original data is under the CC BY 3.0 license; the rest is in the public domain. Please see the Source Data section… See the full description on the dataset page: https://huggingface.co/datasets/Myrtle/CAIMAN-ASR-BackgroundNoise.audio1K<n<10K11 likes1.3k downloads3y agoHugging Face02Elfsong /musicai-background-music-audio-llm-benchmark Does Background Music Matter to Speech in Pre-trained Language Models The completed September 2026 study covers 8 model families, 55 instrumental recordings, and 10 evaluation settings. It studies how adding background music to the same spoken question changes model responses. Latest release and artifact guide Technical report PDF Complete LaTeX project LaTeX GitHub repository Matrices, figures, and supporting data Regenerated speech and mixtures: 550 archives / 250,800… See the full description on the dataset page: https://huggingface.co/datasets/Elfsong/musicai-background-music-audio-llm-benchmark.audio0 likes695 downloads22d agoHugging Face03JackyHoCL /common_voice_22_yue_w_background_captionMerged JackyHoCL/urban-noise-uganda-61k-caption, OpenSound/AudioCaps TODO: convert to MP3, reduce size audio100K<n<1M0 likes161 downloads6mo agoHugging Face04ayush0912 /backgroundmusicaudion<1K0 likes51 downloads11mo agoHugging Face05danielrosehill /ASR-WPM-And-Background-Noise-Eval ASR WPM and Background Noise Evaluation Dataset A dataset of annotated audio recordings for evaluating how different factors affect Whisper (and other ASR/STT systems) transcription accuracy. Purpose This dataset provides controlled audio samples with annotations to evaluate ASR performance across: Speaking pace (fast, normal, slow, mumbled, whispered, weird voices) Background noise (cafe, music, conversations in various languages, traffic, sirens, etc.) Microphone… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/ASR-WPM-And-Background-Noise-Eval.audioautomatic-speech-recognitionn<1K1 likes42 downloads10mo agoHugging Face06kakadong2018 /background-noise-detection-dataset Speech-Free Background Noise Dataset — Real-World, Non-Synthetic (50+ Hours) Dataset summary 50+ hours of real-world urban environmental/ambient background noise (field recordings) without intelligible speech (speech-free), from three scenes: airport, street, subway. The dataset is non-synthetic and intended for speech enhancement via noise augmentation and sound event detection (SED) as “clean background”/negative class Full version of dataset is… See the full description on the dataset page: https://huggingface.co/datasets/kakadong2018/background-noise-detection-dataset.audion<1K1 likes35 downloads3mo agoHugging Face07sdialog /background Background Noise Dataset This dataset contains 3 audio recordings of 2 different background noise classes. Dataset Statistics Total Audio Files: 3 Total Classes: 2 Format: WAV (AudioFolder with metadata.jsonl) Classes and Descriptions The dataset covers the following background noises: Label Description fan_noise Fan noise background white_noise White noise background Structure The dataset is organized in a folder structure… See the full description on the dataset page: https://huggingface.co/datasets/sdialog/background.audioaudio-classificationn<1K0 likes30 downloads8mo agoHugging Face08AxonData /background-noise-detection-dataset Speech-Free Background Noise Dataset — Real-World, Non-Synthetic (50+ Hours) Dataset summary 50+ hours of real-world urban environmental/ambient background noise (field recordings) without intelligible speech (speech-free), from three scenes: airport, street, subway. The dataset is non-synthetic and intended for speech enhancement via noise augmentation and sound event detection (SED) as “clean background”/negative class Full version of dataset is availible… See the full description on the dataset page: https://huggingface.co/datasets/AxonData/background-noise-detection-dataset.audion<1K1 likes29 downloads8mo agoHugging Face09Bear3 /background_noiseaudio0 likes10 downloads1y agoHugging Face10NashAli /background-noise-detection-dataset Speech-Free Background Noise Dataset — Real-World, Non-Synthetic (50+ Hours) Dataset summary 50+ hours of real-world urban environmental/ambient background noise (field recordings) without intelligible speech (speech-free), from three scenes: airport, street, subway. The dataset is non-synthetic and intended for speech enhancement via noise augmentation and sound event detection (SED) as “clean background”/negative class Purpose and usage scenarios Speech… See the full description on the dataset page: https://huggingface.co/datasets/NashAli/background-noise-detection-dataset.audion<1K0 likes10 downloads9mo agoHugging Face11kdcyberdude /background_noiseaudio1 likes5 downloads2y agoHugging Face12xmynscnq /background-movingaudion<1K0 likes3 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.