Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ArlingtonCL2 /Barkopedia_Dog_Sex_Classification_Dataset 📦 Dataset Description This dataset is part of the Barkopedia Challenge: https://uta-acl2.github.io/barkopedia.html Check training data on Hugging Face: 👉 ArlingtonCL2/Barkopedia_Dog_Sex_Classification_Dataset This challenge provides a dataset of labeled dog bark audio clips: 29,345 total clips of vocalizations from 156 individual dogs across 5 breeds: Shiba Inu Husky Chihuahua German Shepherd Pitbull Training set: 26,895 clips 13,567 female13,328 male Test set: 2,450… See the full description on the dataset page: https://huggingface.co/datasets/ArlingtonCL2/Barkopedia_Dog_Sex_Classification_Dataset.audioaudio-classification10K<n<100K0 likes5.5k downloads1y agoHugging Face02ArlingtonCL2 /Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Check Training Data here: ArlingtonCL2/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Description This dataset is for Dog Age Group Classification and contains dog bark audio clips. The data is split into training, public test, and private test sets. Training set: 17888 audio clips. Test set: 4920 audio clips, further divided into: Test Public (~40%): 1966 audio clips for live leaderboard updates. Test Private (~60%): 2954 audio clips for final evaluation. You will… See the full description on the dataset page: https://huggingface.co/datasets/ArlingtonCL2/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K2 likes4.5k downloads1y agoHugging Face03hlx1021 /Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Check Training Data here: ArlingtonCL2/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET Dataset Description This dataset is for Dog Age Group Classification and contains dog bark audio clips. The data is split into training, public test, and private test sets. Training set: 17888 audio clips. Test set: 4920 audio clips, further divided into: Test Public (~40%): 1966 audio clips for live leaderboard updates. Test Private (~60%): 2954 audio clips for final evaluation. You… See the full description on the dataset page: https://huggingface.co/datasets/hlx1021/Barkopedia_DOG_AGE_GROUP_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K0 likes1.1k downloads3mo agoHugging Face04vnahata /AfriMCQA-category-classification Afri-MCQA cross-modal cultural category classification (MTEB) Classify the cultural category of an entry from its photograph and the question about it spoken by a native speaker, across 16 African languages. Labels index this list: geography, building, and landmarks public figure and pop culture cooking and food objects, materials, clothing tranditions, art, and history brands, products, and companies plants and animals people, and everyday life vehicles and transportation… See the full description on the dataset page: https://huggingface.co/datasets/vnahata/AfriMCQA-category-classification.audioaudio-classification1K<n<10K0 likes961 downloads1mo agoHugging Face05mesolitica /Zeroshot-Audio-Classification-Instructions Zeroshot-Audio-Classification-Instructions Convert audio classification dataset into zero-shot format speech instructions, support both single label and multi-label, VGGSound FSD50k Nonspeech7k urbansound8K VocalSound Emotion Gender ESD Emotion Age Language TAU Urban Acoustic Scenes 2022 CochlScene BirdCLEF_2021 EmoBox AudioSet We also converted huge WAV files into MP3 16k sample rate to reduce storage size.To prevent leakage, please do not include test set in training session.… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Zeroshot-Audio-Classification-Instructions.audio1M<n<10M7 likes743 downloads1y agoHugging Face06mteb /Vehicle_sounds_classification_datasetaudio1K<n<10K1 likes632 downloads8mo agoHugging Face07ArlingtonCL2 /Barkopedia_DOG_BREED_CLASSIFICATION_DATASET 📦 Dataset Description This dataset is part of the Barkopedia Challenge 🔗 https://uta-acl2.github.io/barkopedia.html Check Training Data here:👉 ArlingtonCL2/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET This dataset contains 29,347 audio clips of dog barks labeled by dog breed. The audio samples come from 156 individual dogs across 5 dog breeds: shiba inu husky chihuahua german shepherd pitbull 📊 Per-Breed Summary Breed Train Public TestPrivate Test Test… See the full description on the dataset page: https://huggingface.co/datasets/ArlingtonCL2/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K0 likes597 downloads1y agoHugging Face08dolphinteam /OpenWhistle-Classification-Finetuning OpenWhistle Classification Finetuning Dataset dolphinteam/OpenWhistle-Classification-Finetuning is the public classification finetuning dataset used for dolphin whistle identity classification. It contains short whistle clips, whistle-level metadata, fundamental-frequency tracks, rendered F0 spectrograms, and integer class labels. The main reviewer-facing subset is the balanced balanced config. It contains six classes: NSW_1 (label=0) SW_Luna (label=1) SW_Nana (label=2) SW_Neo… See the full description on the dataset page: https://huggingface.co/datasets/dolphinteam/OpenWhistle-Classification-Finetuning.audioaudio-classification10K<n<100K2 likes576 downloads10d agoHugging Face09rpmon /fma-genre-classification FMA Genre Classification Dataset The FMA Genre Classification Dataset is a subset of the Free Music Archive (FMA), containing audio samples and genre labels for music classification tasks. This version uses the "small" subset of FMA, which contains 8,000 tracks of 30 seconds each, evenly distributed across 8 genres. Dataset Description Dataset Summary This dataset consists of 8,000 audio tracks from the Free Music Archive (FMA), each 30 seconds in length… See the full description on the dataset page: https://huggingface.co/datasets/rpmon/fma-genre-classification.audio1K<n<10K3 likes441 downloads2y agoHugging Face10mesolitica /Classification-Speech-Instructions Classification Speech Instructions Speech instructions for emotion, gender, age and language audio classification. Source code Source code at https://github.com/mesolitica/malaysian-dataset/tree/master/llm-instruction/speech-classification-instructions audioaudio-classification100K<n<1M1 likes258 downloads2y agoHugging Face11hamsaai /Recorrected_Classification_Data_filtered_trainaudio10K<n<100K0 likes236 downloads3mo agoHugging Face12vhands /audio-event-classification-post-public audio-event-classification-post-public Sound-event and acoustic-scene classification annotations: ESC-50 (environmental), UrbanSound8K, FSD50k (50k+ events), TUT-Acoustic-Scenes-2017, DCASE-2025, NonSpeech7k (vocal sounds), VocalSound (laugh/cough/sigh). Useful for training audio LLMs on the perception substrate underneath higher-level reasoning. Audio is not bundled in this repo. See download.sh and per-dataset data/<name>.info.json for the fetch recipe; run postlink_audio.py… See the full description on the dataset page: https://huggingface.co/datasets/vhands/audio-event-classification-post-public.textaudio-classification100K<n<1M1 likes204 downloads3mo agoHugging Face13vocsim /mouse-identity-classification-benchmark VocSim — Mouse Identity Classification A companion dataset for the VocSim benchmark that tests whether audio embeddings preserve individual identity in mouse ultrasonic vocalizations (USVs). It contains pre-segmented USV syllables from multiple individual mice (the speaker field), sampled at the native 250 kHz, derived from recordings by Van Segbroeck et al. (2017). Basha, M., Zai, A. T., Stoll, S., & Hahnloser, R. H. R. VocSim: A Training-free Benchmark for Zero-shot Content… See the full description on the dataset page: https://huggingface.co/datasets/vocsim/mouse-identity-classification-benchmark.audio10K<n<100K0 likes156 downloads5mo agoHugging Face14danki2meme /Audio_for_age_classification_Trainaudio1K<n<10K6 likes134 downloads8mo agoHugging Face15kudukudu /building_floor_classificationDataset_chunked_5 : chunks of 05 seconds obtained from expert samples Dataset_chunked_10 : chunks of 10 seconds obtained from expert samples Dataset_expanded : chunks of 10 seconds obtained from whole samples Data.zip : original dataset audio1K<n<10K0 likes128 downloads4y agoHugging Face16PuneettArora /Barkopedia_DOG_BREED_CLASSIFICATION_DATASET 📦 Dataset Description This dataset is part of the Barkopedia Challenge 🔗 https://uta-acl2.github.io/barkopedia.html Check Training Data here:👉 ArlingtonCL2/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET This dataset contains 29,347 audio clips of dog barks labeled by dog breed. The audio samples come from 156 individual dogs across 5 dog breeds: shiba inu husky chihuahua german shepherd pitbull 📊 Per-Breed Summary Breed Train Public Test Private Test… See the full description on the dataset page: https://huggingface.co/datasets/PuneettArora/Barkopedia_DOG_BREED_CLASSIFICATION_DATASET.audioaudio-classification10K<n<100K0 likes120 downloads1mo agoHugging Face17danilotpnta /GTZAN_genre_classificationaudion<1K0 likes102 downloads2y agoHugging Face18DynamicSuperb /Vehicle_sounds_classification_datasetaudio1K<n<10K1 likes86 downloads2y agoHugging Face19danki2meme /Audio_for_age_classification_Evalaudion<1K2 likes84 downloads8mo agoHugging Face20hamsaai /Recorrected_Classification_Data_filtered_train_22audio10K<n<100K0 likes84 downloads3mo agoHugging Face21MUGEN-Benchmark /Genre_Classificationaudion<1K0 likes78 downloads8mo agoHugging Face22laion /vocal-burst-classification-v2 Vocal Burst Classification V2 — laion/vocal-burst-classification-v2 The V2 training corpus for the Vocal Burst Classifier V2: a single-label dataset over an 83-class vocal-burst taxonomy (82 non-speech human vocalizations no_burst, index 82), shipped as precomputed VoiceCLAP-commercial embeddings plus the raw vocal-bursts-clean audio. The vocal-burst clips were generated with various synthetic text-to-audio models such as DramaBox, then annotated and filtered as described… See the full description on the dataset page: https://huggingface.co/datasets/laion/vocal-burst-classification-v2.audioaudio-classification10K<n<100K0 likes75 downloads2mo agoHugging Face23vocsim /mouse-strain-classification-benchmark VocSim — Mouse Strain Classification A companion dataset for the VocSim benchmark that tests whether audio embeddings preserve strain identity in mouse ultrasonic vocalizations (USVs). It contains pre-segmented USV syllables from C57BL/6J (C57) and DBA/2J (DBA) mice, sampled at the native 250 kHz so high-frequency structure is preserved. Basha, M., Zai, A. T., Stoll, S., & Hahnloser, R. H. R. VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source… See the full description on the dataset page: https://huggingface.co/datasets/vocsim/mouse-strain-classification-benchmark.audio10K<n<100K0 likes67 downloads5mo agoHugging Face24Thanushs25 /tamil-audio-emotion-classificationaudioaudio-classification1B<n<10B2 likes64 downloads2y agoHugging Face25vnahata /CAMEO-emotion-classification CAMEO multilingual speech emotion classification (MTEB) Speech emotion recognition across five languages, drawn from the CAMEO collection. Labels index this list: anger fear happiness neutral sadness surprise Source: amu-cai/CAMEO at revision 38e9e96, cc-by-nc-sa-4.0. Split by speaker so no speaker appears in both train and test. Only the six emotions common to every included language are kept. Audio is 16 kHz Opus. Built by scripts/data/cameo_emotion/create_data.py in the… See the full description on the dataset page: https://huggingface.co/datasets/vnahata/CAMEO-emotion-classification.audioaudio-classification10K<n<100K0 likes63 downloads1mo agoHugging Face26hr16 /ViSpeech-Gender-Dialect-Classificationimport datasets as hugDS import pandas as pd import os os.environ["HF_HUB_ENABLE_HF_TRANSFER"] = "1" from df.io import resample from df.enhance import enhance, init_df import torch import warnings df_model, df_state, _ = init_df() SAMPLING_RATE = 16_000 def normalize_vietmed(example): global vietmed_info example["gender"] = vietmed_info[vietmed_info["Speaker ID"] == example["Speaker ID"]]["Gender"].values[0].lower() example["dialect"] = vietmed_info[vietmed_info["Speaker ID"] ==… See the full description on the dataset page: https://huggingface.co/datasets/hr16/ViSpeech-Gender-Dialect-Classification.audioaudio-classification10K<n<100K2 likes61 downloads2y agoHugging Face27greenarcade /wav2vec2-vd-bird-sound-classification-datasetaudioaudio-classification1K<n<10K3 likes59 downloads2y agoHugging Face28kuross /dl-proj-classificationaudio1K<n<10K0 likes57 downloads11mo agoHugging Face29Arulpandi /audio_classification_dataset2audion<1K1 likes54 downloads2y agoHugging Face30assoni2002 /jailbreak_classification_2000audio1K<n<10K0 likes49 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.