Team Ai
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01leungtianle /AgentChat-Test Test Set Description This directory contains the test set used for tool-use evaluation. The JSON files under Test-JSON/ are organized by task type: SingleTaskProcessing/tool-select_test.json: single-tool selection tasks. ParallelProcessing/parallel-call_test.json: parallel tool-call tasks. ProactiveSeeking/searchTools_test_predictions_kept.json: proactive tool-search tasks. TaskDecomposition/muti-tool-select_test.json: multi-tool task decomposition tasks.… See the full description on the dataset page: https://huggingface.co/datasets/leungtianle/AgentChat-Test.audion<1K0 likes504 downloads3mo agoHugging Face02ICTNLP /MutiEmo-Test MultiEmo-Test MultiEmo-Test is an English evaluation set for instruction-following multi-emotion text-to-speech synthesis. It accompanies HybridEmo, a system for modeling sequential emotion trajectories and simultaneous emotion blending within an utterance. The dataset is intended for evaluation only. It contains synthesis text, natural-language emotion instructions, emotion annotations, and prompt audio for speaker-timbre conditioning. It does not contain target synthesized… See the full description on the dataset page: https://huggingface.co/datasets/ICTNLP/MutiEmo-Test.audion<1K1 likes108 downloads1mo agoHugging Face03NVVSpeech-Challenge /NVVSpeech-Challenge-Track1-Test-Set NVVSpeech Challenge Track 1 Test Set Track 1 test set for the NVVSpeech Challenge at ISCSLP 2026. Task Given a speech recording, produce a transcript that contains the spoken content and the non-verbal vocalization (NVV) tags at their corresponding positions. Dataset Summary Language Samples Chinese 985 English 961 Total 1,946 Files . ├── README.md ├── SUBMISSION_GUIDE.txt ├── test.jsonl ├── ground_truth.jsonl ├──… See the full description on the dataset page: https://huggingface.co/datasets/NVVSpeech-Challenge/NVVSpeech-Challenge-Track1-Test-Set.audioautomatic-speech-recognition1K<n<10K0 likes106 downloads12d agoHugging Face04NVVSpeech-Challenge /NVVSpeech-Challenge-Track2-Test-Set NVVSpeech Challenge Track 2 Test Set Track 2 test set for the NVVSpeech Challenge at ISCSLP 2026. Task Given a transcript containing one or more non-verbal vocalization (NVV) tags, synthesize speech that naturally realizes the requested NVVs while preserving intelligibility, naturalness, and audio quality. Dataset Summary Language Samples Chinese 800 English 800 Total 1,600 Files . ├── README.md ├──… See the full description on the dataset page: https://huggingface.co/datasets/NVVSpeech-Challenge/NVVSpeech-Challenge-Track2-Test-Set.texttext-to-speech1K<n<10K0 likes63 downloads12d agoHugging Face05Codyfederer /test321 test321 This is a merged speech dataset containing 118 audio segments from 2 source datasets. Dataset Information Total Segments: 118 Speakers: 4 Languages: tr Emotions: happy, angry, sad, neutral Original Datasets: 2 Dataset Structure Each example contains: audio: Audio file (WAV format, 16kHz sampling rate) text: Transcription of the audio speaker_id: Unique speaker identifier (made unique across all merged datasets) emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/test321.audioautomatic-speech-recognitionn<1K0 likes37 downloads1y agoHugging Face06eturok-weizmann /laser-vibrations-test Laser Vibrations Dataset of laser speckle vibration recordings used to locate objects hidden inside a cardboard box. A 10×10 grid of lasers shines on the side of a box containing an object; as loudspeakers excite the box, the speckle patterns shift in proportion to the local surface vibration. Per-sample metadata is in data/metadata.jsonl; full signal data and media files live in per-sample subdirectories. Dataset Viewer Columns Column Type Description… See the full description on the dataset page: https://huggingface.co/datasets/eturok-weizmann/laser-vibrations-test.audion<1K1 likes22 downloads6mo agoHugging Face07DaveLoay /Nsynth_Test_Split_Tango_Formataudio1K<n<10K0 likes16 downloads3y agoHugging Face08Codyfederer /testtr43 testtr43 This is a merged speech dataset containing 2655 audio segments from 3 source datasets. Dataset Information Total Segments: 2655 Speakers: 13 Languages: tr Emotions: angry, happy, neutral Original Datasets: 3 Dataset Structure Each example contains: audio: Audio file (WAV format, original sampling rate preserved) text: Transcription of the audio speaker_id: Unique speaker identifier (made unique across all merged datasets) emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/testtr43.audioautomatic-speech-recognition1K<n<10K0 likes16 downloads1y agoHugging Face09Adanmohh /hafidh-test-fixturesaudion<1K0 likes15 downloads2mo agoHugging Face10007ask /testdatasetgated NeMo Tarred Dataset Generated from Test3. Train rows: 40338 · Test rows: 422 Shards: 5 · Codec: flac · Sample rate: 16000 Hz mono Primary text: text · target_lang: ta-IN is_tarred: true tarred_audio_filepaths: .../audio__OP_0..4_CL_.tar manifest_filepath: .../train_manifest.json textautomatic-speech-recognition10K<n<100K0 likes13 downloads2mo agoHugging Face11MohamedHussienOmar /whisper-finetune-audio_test2audion<1K0 likes10 downloads1y agoHugging Face12MohamedHussienOmar /whisper-finetune-audio_test3audion<1K0 likes4 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.