Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RidheshBhati /Codemixed_New Codemixed ASR Dataset Unified collection of code-mixed ASR datasets. audioautomatic-speech-recognition100K<n<1M2 likes6.1k downloads5mo agoHugging Face02Codec-SUPERB /fluent_speech_commands_synth Dataset Card for "fluent_speech_commands_synth" More Information needed audio100K<n<1M1 likes1.8k downloads3y agoHugging Face03rogertseng /CodecFake CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems Paper, Code, Project Page Interspeech 2024 TL;DR: We show that better detection of deepfake speech from codec-based TTS systems can be achieved by training models on speech re-synthesized with neural audio codecs. This dataset is released for this purpose. See our paper and Github for more details on using our dataset. Acknowledgement… See the full description on the dataset page: https://huggingface.co/datasets/rogertseng/CodecFake.audio100K<n<1M5 likes1.6k downloads2y agoHugging Face04besimple-ai /voice-code-bench VoiceCodeBench VoiceCodeBench is a test-only benchmark for evaluating whether automatic speech recognition (ASR) systems preserve exact structured values in English workplace speech. Paper: VoiceCodeBench: Evaluating Exact Structured-Token Recovery in Automatic Speech Recognition The benchmark targets cases where a transcript is software input: callback numbers, email addresses, command-line flags, file paths, URLs, account identifiers, dates, measurements, and similar values… See the full description on the dataset page: https://huggingface.co/datasets/besimple-ai/voice-code-bench.audioautomatic-speech-recognitionn<1K14 likes1.4k downloads8d agoHugging Face05code-lover-ai /WSC-Evalaudio0 likes1.3k downloads3mo agoHugging Face06Perle-ai /ASR_Code_Switch ASR Code-Switching Benchmark A curated benchmark of 1,200 code-switching utterances (300 per language pair) for evaluating commercial ASR systems on multilingual speech with intra-sentential language switching. Paper Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German arXiv link Language pairs Split Language pair Samples Scripts egyptian_arabic_english Egyptian Arabic–English 300 Arabic + Latin… See the full description on the dataset page: https://huggingface.co/datasets/Perle-ai/ASR_Code_Switch.audioautomatic-speech-recognition1K<n<10K12 likes1k downloads5mo agoHugging Face07Codec-SUPERB /crema_d_synth Dataset Card for "crema_d_synth" More Information needed audio100K<n<1M0 likes653 downloads3y agoHugging Face08Codec-SUPERB /maestro_synth Dataset Card for "maestro_synth" More Information needed audio1K<n<10K0 likes618 downloads3y agoHugging Face09ajaykarthick /codecfake-audio Codecfake Dataset Overview The Codecfake dataset is a large-scale dataset designed for the detection of Audio Language Model (ALM)-based deepfake audio. This dataset includes millions of audio samples across two languages and various test conditions, tailored specifically for ALM-based audio detection. Conversion The original dataset was downloaded from Zenodo and converted to FLAC format to maintain audio quality while reducing file size. The dataset has been… See the full description on the dataset page: https://huggingface.co/datasets/ajaykarthick/codecfake-audio.audioaudio-classification100K<n<1M1 likes539 downloads2y agoHugging Face10Codec-SUPERB /voxceleb1_synthaudio10K<n<100K3 likes532 downloads3y agoHugging Face11DhruvBhatia0 /smash-battlefield-fox-codec-20fps Smash Battlefield Fox — codec-ready videos Lossless 20 FPS preprocessing of DhruvBhatia0/smash-battlefield-fox at revision 2c94351b82c2c65a31fb39fe52a34ff905b6abcf. This dataset contains 512 complete replays selected deterministically with seed 28. Every third decoded frame is resized to 252×208 with PyAV's training-time resize, then stored losslessly as RGB FFV1 in a streaming NUT container. Decoding the processed files reproduces the preprocessed RGB tensors bit-for-bit.… See the full description on the dataset page: https://huggingface.co/datasets/DhruvBhatia0/smash-battlefield-fox-codec-20fps.audion<1K0 likes532 downloads3mo agoHugging Face12Codec-SUPERB /vocal_imitation_synth Dataset Card for "vocal_imitation_synth" More Information needed audio10K<n<100K1 likes514 downloads3y agoHugging Face13Codec-SUPERB /opensinger_synthaudio10K<n<100K0 likes503 downloads3y agoHugging Face14CodecSR /torgo_synthaudio100K<n<1M0 likes503 downloads2y agoHugging Face15CodecSR /speech_accent_archive_synthaudio10K<n<100K0 likes488 downloads2y agoHugging Face16Codec-SUPERB /fsd50k_synthaudio100K<n<1M0 likes479 downloads3y agoHugging Face17CodecSR /librispeech_asr_test_48k_synthaudio100K<n<1M0 likes476 downloads3y agoHugging Face18ghanaopenai /Ghana_English-Twi_Code-switching_Speech Dataset Card for KasaSpeech Dataset Summary KasaSpeech is a large-scale English–Twi code-switching speech dataset developed to advance research in speech technologies for English and Twi. The dataset comprises 54,855 transcribed speech recordings collected from speakers across Ghana and is designed to capture natural code-switching between English and Twi across a diverse range of everyday topics and communication scenarios With over 95 hours of manually… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/Ghana_English-Twi_Code-switching_Speech.audiotext-to-speech10K<n<100K5 likes471 downloads24d agoHugging Face19CodecSR /vox_lingua_top10_synthaudio10K<n<100K0 likes458 downloads3y agoHugging Face20Codec-SUPERB /sample_100audion<1K0 likes456 downloads2y agoHugging Face21CodecSR /vocalset_synthaudio10K<n<100K0 likes449 downloads3y agoHugging Face22CodecSR /speech_accent_archive_englishaudio10K<n<100K1 likes445 downloads2y agoHugging Face23CodecSR /fluent_speech_commands_femaleaudio10K<n<100K1 likes422 downloads2y agoHugging Face24CodecSR /librispeech_asr_test_synthaudio100K<n<1M0 likes421 downloads3y agoHugging Face25Codec-SUPERB /vocalset_synth Dataset Card for "vocalset_synth" More Information needed audio10K<n<100K0 likes393 downloads3y agoHugging Face26Codec-SUPERB /snips_test_valid_synth Dataset Card for "snips_test_valid_synth" More Information needed audio100K<n<1M0 likes387 downloads3y agoHugging Face27CodecSR /speech_accent_archive_otheraudio10K<n<100K1 likes374 downloads2y agoHugging Face28CodecSR /easycall_synthaudio100K<n<1M0 likes366 downloads2y agoHugging Face29Codec-SUPERB /librispeech_synth Dataset Card for "librispeech_synth" More Information needed audio1M<n<10M1 likes355 downloads3y agoHugging Face30Codec-SUPERB /cv_13_zh_tw_synthaudio100K<n<1M0 likes351 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.