Team Ai
3 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Reubencf /adaption-music-style-prompts This dataset is a remastered version of Reubencf/fma-labeled prepared using Adaption's Adaptive Data platform. music_style_prompts This dataset contains a collection of descriptive text prompts designed to generate diverse musical tracks across various genres, including pop, techno, ambient, and rock. Each entry details specific instrumentation, rhythmic patterns, atmospheric qualities, and emotional tones to guide audio synthesis. The content serves as a resource Intended for… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/adaption-music-style-prompts.audio1K<n<10K4 likes507 downloads6mo agoHugging Face02Reubencf /Adaption-multilingual-speech This dataset is a remastered version of Reubencf/multilingual-synthetic-tts prepared using Adaption's Adaptive Data platform. Multilingual Speech (Adaption) 10,274 audio + text rows selected from the original 68,677-clip multilingual synthetic speech corpus, with Adaption-sharpened enhanced_prompt and enhanced_completion columns. Every row carries the synthesised audio, the ground-truth text, and language/style/voice metadata — ready for speech SFT. Original dataset (for… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/Adaption-multilingual-speech.audiotext-to-speech10K<n<100K0 likes58 downloads6mo agoHugging Face03Reubencf /Adaption-low-resource-audio Adaption Low-Resource Audio A low-resource-language subset of Reubencf/PolyglotAudio, remastered with Adaption's Adaptive Data platform. Each row carries the original Tatoeba-derived audio clip alongside sharpened enhanced_prompt / enhanced_completion columns so the data is ready for speech-model fine-tuning and evaluation on languages that are typically under-represented in open ASR/TTS corpora. Dataset size 3,704 rows of paired audio + text, spanning 10 languages… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/Adaption-low-resource-audio.audioautomatic-speech-recognition1K<n<10K1 likes36 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.