Team Ai
Datasetpublic

aoxo/t2a-daddy

t2a-daddy Male-voice ASMR corpus for the text2asmr project. Previously published as aoxo/audios3. Companion repos: aoxo/t2a-mommy (female voice), aoxo/t2a-audios-v1 (the original v1 corpus). Layout path what <creator>/<title>.m4a source audio, 48 kHz AAC, one folder per creator <creator>/<title>.json word-level Whisper large-v3 alignment ([] = skipped: near-silent or undecodable) labels/qwen3omni.jsonl non-speech ontology labels for gap clips… See the full description on the dataset page: https://huggingface.co/datasets/aoxo/t2a-daddy.

sourceHugging Faceapache-2.0updated 8d agoView on Hugging Face
0likes8.7kdownloads
Dataset Card

t2a-daddy

Male-voice ASMR corpus for the text2asmr project. Previously published as aoxo/audios3.

Companion repos: `aoxo/t2a-mommy` (female voice), `aoxo/t2a-audios-v1` (the original v1 corpus).

Layout

pathwhat
<creator>/<title>.m4asource audio, 48 kHz AAC, one folder per creator
<creator>/<title>.jsonword-level Whisper large-v3 alignment ([] = skipped: near-silent or undecodable)
labels/qwen3omni.jsonlnon-speech ontology labels for gap clips, Qwen3-Omni-30B-A3B-Instruct (945 k+ rows)
labels/pending_candidates.jsonlgap clips cut but not yet labeled

Label rows are {uid, raw, label, labeler}, where uid is <source>.m4a_<start_ms>.

Ontology: whispering, normal speech, breathing, mouth sounds, moaning, kissing, silence, plus the physical tail (tapping, scratching, crinkling, brushing, liquid, page turning, other sound). A label covers at most 3–4 s of audio.

Audio is sourced from soundgasm; each creator keeps their own folder, so attribution and creator-split evaluation are preserved.

aoxo/t2a-daddy · Team Ai