Team Ai
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01pre-view /CS50-rawaudio10K<n<100K0 likes134 downloads2y agoHugging Face02TigreGotico /synthetic-wakeword-view_glass synthetic-wakeword-view_glass Synthetic wake-word audio for training and benchmarking OVOS wake-word plugins, covering the phrase "view glass". Every clip is machine-generated by text-to-speech, and no human recording is included. The dataset is published under CC BY 4.0, free to use, redistribute and build on, including for model training. Produced with support from the NGI0 Commons Fund. Layout Every clip is in train/ (1000 clips), inside a folder named after… See the full description on the dataset page: https://huggingface.co/datasets/TigreGotico/synthetic-wakeword-view_glass.audioaudio-classification1K<n<10K0 likes79 downloads8d agoHugging Face03polinaeterna /test_audio_vieweraudion<1K0 likes24 downloads3y agoHugging Face04htdung167 /vivos-preprocessed-vieweraudio10K<n<100K0 likes20 downloads3y agoHugging Face05kittinol /fleurs-th-vieweraudio1K<n<10K0 likes16 downloads1mo agoHugging Face06RemiFabre /marionette-viewer-test marionette viewer test • Reachy Mini Moves Temporary test dataset — verifying that an audiofolder configs: block makes the Hugging Face dataset viewer play audio for Marionette community datasets. Files copied from Anne-Charlotte/reachy-songs. Safe to delete. Community-contributed Marionette recordings captured on Reachy Mini. Files live under data/, each move ships as a JSON trajectory plus an optional audio sidecar. audioroboticsn<1K0 likes9 downloads3mo agoHugging Face07Thanarit /Thai-Voice-Test-Viewer-Fix Thanarit/Thai-Voice Combined Thai audio dataset from multiple sources Dataset Details Total samples: 120 Total duration: 0.13 hours Language: Thai (th) Audio format: 16kHz mono WAV Volume normalization: -20dB Sources Processed 1 datasets in streaming mode Source Datasets GigaSpeech2: Large-scale multilingual speech corpus Usage from datasets import load_dataset # Load with streaming to avoid downloading everything dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Thanarit/Thai-Voice-Test-Viewer-Fix.audion<1K0 likes8 downloads1y agoHugging Face08ClaudeChen /test_viewer [doc] audio dataset 10 This dataset contains four audio files, two in the /train directory (one in the cat/ subdirectory and one in the dog/ subdirectory), and two in the test/ directory (same distribution in subdirectories). The label column is not created because the configuration contains drop_labels: true. audion<1K0 likes6 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.