Team Ai
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01aoxo /asmr-yt-chaptersaudio0 likes18k downloads3h agoHugging Face02noxwano /ASMR-Archive-Processed-SFW ASMR-Archive-Processed-SFW Overview This dataset is an “educational” subset of the original OmniAICreator/ASMR-Archive-Processed dataset. We filtered the original dataset to include only records where the nsfw metadata flag is false. To maintain the randomness and anonymity of the entries, multiple directories were combined and shuffled. The nsfw tag in the original dataset is inherited from the tags of the original audio works before they were passed through the… See the full description on the dataset page: https://huggingface.co/datasets/noxwano/ASMR-Archive-Processed-SFW.audioautomatic-speech-recognition1M<n<10M9 likes3.5k downloads6mo agoHugging Face03OmniAICreator /ASMR-Archive-Processed ASMR-Archive-Processed (WIP) Update (2026-04-03): This dataset has reached the Hugging Face Public Storage Limit. After contacting support, we were informed that the only option is to pay for a storage expansion. Consequently, updates to this dataset are now suspended. Work in Progress — expect breaking changes while the pipeline and data layout stabilize. This dataset contains ASMR audio data sourced from DeliberatorArchiver/asmr-archive-data-01 and… See the full description on the dataset page: https://huggingface.co/datasets/OmniAICreator/ASMR-Archive-Processed.imageautomatic-speech-recognition98 likes2.2k downloads6mo agoHugging Face04DeliberatorArchiver /asmr-archive-data-01 ASMR Media Archive Storage This repository contains an archive of ASMR works. All data in this repository is uploaded for educational and research purposes only. All use is at your own risk. [!IMPORTANT] This repository contains >= 64 TiB of files.Git LFS consumes twice as much disk space because of the way it works, so git clone is not recommended. Hugging Face CLI or Python libraries allow you to select and download only a subset of files. >>> CLICK HERE or on the IMAGE BELOW… See the full description on the dataset page: https://huggingface.co/datasets/DeliberatorArchiver/asmr-archive-data-01.n>1T15 likes698 downloads2y agoHugging Face05DeliberatorArchiver /asmr-archive-data-02 ASMR Media Archive Storage This repository contains an archive of ASMR works. All data in this repository is uploaded for educational and research purposes only. All use is at your own risk. [!IMPORTANT] This repository contains >= 64 TiB of files.Git LFS consumes twice as much disk space because of the way it works, so git clone is not recommended. Hugging Face CLI or Python libraries allow you to select and download only a subset of files. >>> CLICK HERE or on the IMAGE BELOW… See the full description on the dataset page: https://huggingface.co/datasets/DeliberatorArchiver/asmr-archive-data-02.n>1T7 likes641 downloads2y agoHugging Face06jasonfan /asmr-zh-r18 asmr-zh-r18 Chinese R18 ASMR audio dataset for TTS/voice cloning fine-tuning. Works: 5876 RJ-coded works Total size: ~1.1TB raw (10 compressed parts) Format: MP3/WAV/FLAC, 3 tracks per work Source: asmr.one API, Chinese R18 category Extract for f in packs/*.tar.zst; do zstd -d "$f" --stdout | tar -xf - done 2 likes280 downloads7mo agoHugging Face07nyuuzyou /asmr Dataset Card for ASMR Audio Dataset Dataset Summary This dataset contains a large collection of ASMR (Autonomous Sensory Meridian Response) audio clips with corresponding machine-generated transcriptions. The dataset includes approximately 283,132 audio segments totaling over 307 hours of content, with an average duration of 3.92 seconds per clip. All audio files are provided in WAV format at 24 kHz sampling rate, making them suitable for various audio processing and… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/asmr.textautomatic-speech-recognition100K<n<1M6 likes95 downloads1y agoHugging Face08urcate /japanese_asmr14 likes78 downloads3y agoHugging Face09Hachiki /asmr-mami1 likes72 downloads3y agoHugging Face10LuckPr4yx /ASMR1 likes61 downloads3y agoHugging Face11wangdezihao /japanese_asmr1 likes49 downloads3mo agoHugging Face12DeliberatorArchiver /asmr-api-backup-output1 likes45 downloads2y agoHugging Face13OOPPEENN /ASMR_Datasetcrawl progress: 57% 121249/213330 pipeline: Manual cleaning, The following situations will be discarded: no Chinese subtitles timestamps of each sentence in the subtitles are connected at the end timestamps are not aligned mel band roformer remove SE, BGM, saliva sound, etc. parsec subtitle & cut audio Split the left and right channels and use anime-whisper to identify the loudest channel(beta version, maybe there is a better way) create index 16 likes38 downloads2y agoHugging Face14telecomadm1145 /asmr_archive_qwentts_encoded0 likes21 downloads2mo agoHugging Face15kontextox /uk_UA-ASMR Ukrainian ASMR TTS Dataset A Ukrainian text-to-speech dataset for training single-speaker ASMR-style voice models using Piper. Dataset Details Property Value Language Ukrainian (uk_UA) Speakers 1 Segments 7,318 Audio Format 16-bit WAV, 22050 Hz, Mono License CC0 Dataset Structure Prerequisites # Install Piper training dependencies git clone https://github.com/kontextox/piper1-gpl.git cd piper1-gpl python3 -m venv .venv source… See the full description on the dataset page: https://huggingface.co/datasets/kontextox/uk_UA-ASMR.text1K<n<10K0 likes20 downloads6mo agoHugging Face16DeliberatorArchiver /asmr-archive-data-meta3 likes17 downloads2y agoHugging Face17ruliad /asm_rgatedaudio1K<n<10K0 likes13 downloads1mo agoHugging Face18cmeraki /youtube_en_asmr_raw0 likes11 downloads2y agoHugging Face19rafihmd21 /humanoid-asmr-datatextn<1K0 likes7 downloads9mo agoHugging Face20cmeraki /youtube_en_asmr0 likes6 downloads2y agoHugging Face21Rvcmodel /whisper-asmrgated0 likes3 downloads1y agoHugging Face22kikouousya /personal-asmr0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.