Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01manavtabbly /hindi_audio_dataset_testaudion<1K0 likes2.6k downloads1y agoHugging Face02pipecat-ai /smart-turn-data-v3.2-testTesting dataset for Smart Turn v3.2. Thank you to the following contributors whose audio samples are included in this dataset: The Pipecat team Liva AI: https://www.theliva.ai/ Midcentury: https://www.midcentury.xyz/ MundoAI: https://mundoai.world/ Also, thank you to the following people for the CC-0 background noise sample data which has been used in this dataset: https://freesound.org/people/4team/sounds/214995/ https://freesound.org/people/tomhannen/sounds/698090/… See the full description on the dataset page: https://huggingface.co/datasets/pipecat-ai/smart-turn-data-v3.2-test.audio10K<n<100K3 likes1.2k downloads11d agoHugging Face03argmaxinc /whisperkit-test-dataaudion<1K0 likes893 downloads5mo agoHugging Face04tsinghua-ee /ELLSA_test_data ELLSA: End-to-end Listen, Look, Speak and Act The first end-to-end model that unifies vision, speech, text and actionin a streaming full-duplex framework, enabling joint multimodal perception and concurrent generation. 🧪 Highlights Full-Duplex Multimodal Interaction: unifies listening, looking, speaking, and acting in a single end-to-end architecture, enabling simultaneous… See the full description on the dataset page: https://huggingface.co/datasets/tsinghua-ee/ELLSA_test_data.audio1K<n<10K0 likes746 downloads6mo agoHugging Face05masumtechnonext /test-data-set-Arabic-letteraudio10K<n<100K0 likes585 downloads2mo agoHugging Face06chunking-ai /chunking-test-dataaudion<1K0 likes507 downloads1y agoHugging Face07RareConcepts /suno-reggae-test-dataset Suno Patois Reggae Test Set 299 patois-language reggae and dancehall tracks with style captions and structured lyrics, laid out for SimpleTuner's textfile audio caption strategy. Built as a small, high-consistency probe set for text-to-audio training runs — not a general-purpose music corpus. Rights and provenance Every track here was generated by a third-party Suno user, and rights in the audio and lyrics remain with those creators. Nothing in this repository is… See the full description on the dataset page: https://huggingface.co/datasets/RareConcepts/suno-reggae-test-dataset.audiotext-to-audion<1K0 likes396 downloads2mo agoHugging Face08alexandrainst /audio_test_dataset Dataset Card for "audio_test_dataset" This dataset consists of the first 5 samples of mozilla-foundation/common_voice_13_0 and is only used for unit testing. audion<1K0 likes372 downloads3y agoHugging Face09pipecat-ai /smart-turn-data-v3.1-testTesting dataset for Smart Turn v3.1. Thank you to the following contributors whose audio samples are included in this dataset: The Pipecat team Liva AI: https://www.theliva.ai/ Midcentury: https://www.midcentury.xyz/ MundoAI: https://mundoai.world/ License This dataset is licensed under the Creative Commons Attribution 4.0 International License (CC BY 4.0). See the LICENSE file for the full license text. audio10K<n<100K2 likes230 downloads11d agoHugging Face10pipecat-ai /smart-turn-data-v3-testTesting dataset for Smart Turn v3. License This dataset is licensed under the Creative Commons Attribution 4.0 International License (CC BY 4.0). See the LICENSE file for the full license text. audio10K<n<100K8 likes181 downloads11d agoHugging Face11speako /wav2vec2-test-datasetaudio10K<n<100K0 likes169 downloads1y agoHugging Face12Narsil /test_dataaudion<1K0 likes148 downloads2y agoHugging Face13MTNRA-Amazigh-IA /test_dataaudio1K<n<10K0 likes134 downloads17d agoHugging Face14BSYM-25 /Test_Audio_Generate_Dataset Hinglish Audio Dataset Generated by Sarvam AI. audio1K<n<10K0 likes118 downloads9mo agoHugging Face15akashgoel-id /test_TTS_data_hindi_v2audio0 likes102 downloads10mo agoHugging Face16CYenHua /Test_data Title Test Test Test audio1K<n<10K0 likes98 downloads9mo agoHugging Face17xiaofff /omnievalkit-data-test OmniEvalKit Evaluation Datasets Evaluation datasets for OmniEvalKit, a comprehensive evaluation framework for omni-modal (audio + video + image + text) models. Overview Total subsets: 89 Total samples: 353,610 Total size: 352.3 GB (Parquet with embedded audio/image, no video) Subsets requiring video download: 42 Note: Video files are NOT embedded in the Parquet files due to size constraints. Usage from datasets import load_dataset ds =… See the full description on the dataset page: https://huggingface.co/datasets/xiaofff/omnievalkit-data-test.audioaudio-classification10K<n<100K0 likes78 downloads7mo agoHugging Face18MaHaWo /iSparrow_test_dataaudion<1K1 likes72 downloads3y agoHugging Face19speaches-ai /realtime-turn-detection-test-data Realtime speech test recordings Synthetic speech recordings for black-box Realtime API behavior tests in Speaches. Each WAV file is the unmodified output of OpenAI text-to-speech. Tests are responsible for adding silence, combining recordings, and choosing streaming chunk boundaries for their scenarios. metadata.jsonl follows the Hugging Face AudioFolder layout. Each record contains the generation inputs, file digest, expected text, transcription, and word/speech intervals from… See the full description on the dataset page: https://huggingface.co/datasets/speaches-ai/realtime-turn-detection-test-data.audion<1K0 likes54 downloads2mo agoHugging Face20chanchungkit /IMDA-NSC-datasets-testaudio10K<n<100K0 likes49 downloads1y agoHugging Face21SarahUssama /sada-arabic-test-dataset-sample 🗣️ Arabic Dialect Segmented Speech Dataset (SADA2022 Subset) This dataset contains segmented Arabic speech samples from the SADA2022 corpus, annotated by dialect, gender, age group, speaking rate, environmental condition, and includes ground truth transcriptions. It is intended to support research and applications in Arabic dialect classification, automatic speech recognition (ASR), and spoken language understanding. 📁 Dataset Structure Audio segments are stored… See the full description on the dataset page: https://huggingface.co/datasets/SarahUssama/sada-arabic-test-dataset-sample.audion<1K0 likes41 downloads1y agoHugging Face22uriel /audio_data_kaggle_test_taskcaudio1K<n<10K0 likes39 downloads1y agoHugging Face23MoneerProject /qaloon_dataset_testaudio1K<n<10K0 likes38 downloads1y agoHugging Face24swatt-TRF /Test-Dataset3gated Test Dataset 3 This is a test preview of The Rights Foundry Collection One Dataset Summary A licensed 10-track music dataset for non-commercial AI research, reproducible benchmarking, model evaluation, and music-information-retrieval research. This test preview of The Rights Foundry Collection One contains 10 fully rights-cleared commercial music tracks from independent artists, provided as 16-bit, 44.1 kHz stereo WAV audio with structured descriptive metadata.… See the full description on the dataset page: https://huggingface.co/datasets/swatt-TRF/Test-Dataset3.audioaudio-classificationn<1K0 likes38 downloads4d agoHugging Face25MaryWambo /Test_model_dataaudion<1K0 likes37 downloads6mo agoHugging Face26bayartsogt /test-audio-datasetaudion<1K0 likes36 downloads4y agoHugging Face27salmanshahid /test_audio_datasetaudio10K<n<100K0 likes35 downloads2y agoHugging Face28uriel /audio_data_kaggle_test_taskb_audio1K<n<10K0 likes35 downloads1y agoHugging Face29Zetaphor /test-datasetaudion<1K0 likes33 downloads1y agoHugging Face30leeoxiang /test-audio-dataset Test Audio Dataset 这是一个用于测试的音频数据集,包含 100 条伪造的 WAV 格式音频文件。 Dataset Structure audio-dataset/ ├── data/ │ ├── audio_0000.wav │ ├── audio_0001.wav │ └── ... ├── metadata.csv └── README.md Data Fields file_name: 音频文件路径 transcription: 转录文本 speaker_id: 说话人ID duration: 音频时长(秒) sample_rate: 采样率 Usage from datasets import load_dataset dataset = load_dataset("your-username/test-audio-dataset") License MIT License audioautomatic-speech-recognitionn<1K0 likes33 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.