Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Multilingual-Multimodal-NLP /IfEvalCode-testsettextn<1K2 likes4.9k downloads1y agoHugging Face02allenai /preference-test-sets Preference Test Sets Very few preference datasets have heldout test sets for validation of reward model accuracy results. In this dataset, we curate the test sets from popular preference datasets into a common schema for easy loading and evaluation. Anthropic HH (Helpful & Harmless Agent and Red Teaming), test set in full is 8552 samples Anthropic HHH Alignment (Helpful, Honest, & Harmless), formatted from Big Bench for standalone evaluation. Learning to summarize, downsampled from… See the full description on the dataset page: https://huggingface.co/datasets/allenai/preference-test-sets.textsummarization10K<n<100K28 likes3.8k downloads3y agoHugging Face03ASLP-lab /Easy-Turn-Testset Easy Turn: Integrating Acoustic and Linguistic Modalities for Robust Turn-Taking in Full-Duplex Spoken Dialogue Systems Guojian Li1, Chengyou Wang1, Hongfei Xue1, Shuiyuan Wang1, Dehui Gao1, Zihan Zhang2, Yuke Lin2, Wenjie Li2, Longshuai Xiao2, Zhonghua Fu1,╀, Lei Xie1,╀ 1 Audio, Speech and Language Processing Group (ASLP@NPU), Northwestern Polytechnical University 2 Huawei Technologies, China 🎤 Demo Page 🤖 Easy Turn Model 📑 Paper 🌐 Huggingface… See the full description on the dataset page: https://huggingface.co/datasets/ASLP-lab/Easy-Turn-Testset.automatic-speech-recognition8 likes1.7k downloads1y agoHugging Face04mideind /gec-test-setTest data for Icelandic spell and grammar checking, created as part of the Icelandic Language Technology Programme. The test data is divided into three different formats, type 1, 2 and 3. For every original file corrected, three files are included in the test data when possible: _original, _corrected and _metadata. The original and metadata files are always .txt files, but the format of the corrected file differs between types. Texts corrected are from the News2 subcorpus of the Icelandic… See the full description on the dataset page: https://huggingface.co/datasets/mideind/gec-test-set.0 likes1.5k downloads2y agoHugging Face05manycore-research /SpatialLM-Testset SpatialLM Testset Project page | Paper | Code We provide a test set of 107 preprocessed point clouds and their corresponding GT layouts, point clouds are reconstructed from RGB videos using MASt3R-SLAM. SpatialLM-Testset is quite challenging compared to prior clean RGBD scan datasets due to the noises and occlusions in the point clouds reconstructed from monocular RGB videos. Folder Structure Outlines of the dataset files:… See the full description on the dataset page: https://huggingface.co/datasets/manycore-research/SpatialLM-Testset.3dn<1K60 likes1.2k downloads1y agoHugging Face06junchaoh-cs /SolarWM-Data_test-set-v1gated SolarWM Standalone Test Set v1 This repository contains the complete, self-contained SolarWM test set without the training shards. It includes 1,300 clips from 13 source views (100 per view), packaged as 60 uncompressed WebDataset tar files totaling approximately 77.4 GB. The test identities are the current accepted SolarWM standalone evaluation split, excluding MIND. The release contains 757 xhigh and 543 high samples. All selected samples have non-empty captions and finite… See the full description on the dataset page: https://huggingface.co/datasets/junchaoh-cs/SolarWM-Data_test-set-v1.video1K<n<10K0 likes1.2k downloads26d agoHugging Face07MiniMaxAI /TTS-Multilingual-Test-Set Overview To assess the multilingual zero-shot voice cloning capabilities of TTS models, we have constructed a test set encompassing 24 languages. This dataset provides both audio samples for voice cloning and corresponding test texts. Specifically, the test set for each language includes: 100 distinct test sentences. Audio samples from two speakers (one male and one female) carefully selected from the Mozilla Common Voice (MCV) dataset, intended for voice cloning. Researchers can… See the full description on the dataset page: https://huggingface.co/datasets/MiniMaxAI/TTS-Multilingual-Test-Set.audiotext-to-speechn<1K47 likes1.2k downloads1y agoHugging Face08sunday-hao /vindr-cxr-testsetimage1K<n<10K0 likes764 downloads3mo agoHugging Face09Iceclear /StableSR-TestSets StableSR TestSets Card These test sets are used associated with the StableSR, available here. Data Details Developed by: Jianyi Wang Data type: Synthetic and real-world test sets for image super-resolution License: S-Lab License 1.0 Data Description: The test sets are used to reproduce the metric results shown in Paper. Resources for more information: GitHub Repository. Cite as: @InProceedings{wang2023exploiting, author = {Wang, Jianyi and Yue, Zongsheng and… See the full description on the dataset page: https://huggingface.co/datasets/Iceclear/StableSR-TestSets.imageimage-to-image1K<n<10K3 likes575 downloads3y agoHugging Face10videoSALMONN2 /video-SALMONN_2_testset video-SALMONN 2 Benchmark Generate the caption corresponding to the video and the audio with video_salmonn2_test.json Organize your results in the format like the following example: [ { "id": ["0.mp4"], "pred": "Generated Caption" } ] Replace res_file in eval.py with your result file. Run python3 eval.pytextn<1K3 likes566 downloads1y agoHugging Face11argilla-internal-testing /test_import_dataset_from_hub_using_settings_with_recordsFalse Dataset Card for test_import_dataset_from_hub_using_settings_with_recordsFalse This dataset has been created with Argilla. As shown in the sections below, this dataset can be loaded into your Argilla server as explained in Load with Argilla, or used directly with the datasets library in Load with datasets. Using this dataset with Argilla To load with Argilla, you'll just need to install Argilla as pip install argilla --upgrade and then use the following code: import… See the full description on the dataset page: https://huggingface.co/datasets/argilla-internal-testing/test_import_dataset_from_hub_using_settings_with_recordsFalse.n<1K0 likes531 downloads2y agoHugging Face12facebook /emu_edit_test_set Dataset Card for the Emu Edit Test Set Dataset Summary To create a benchmark for image editing we first define seven different categories of potential image editing operations: background alteration (background), comprehensive image changes (global), style alteration (style), object removal (remove), object addition (add), localized modifications (local), and color/texture alterations (texture). Then, we utilize the diverse set of input images from the MagicBrush… See the full description on the dataset page: https://huggingface.co/datasets/facebook/emu_edit_test_set.image1K<n<10K47 likes529 downloads3y agoHugging Face13gauravparajuli /coco_test_set_pybboxes COCO Test Set This is a coco test set which is used for unit testing in pybboxes library. imagen<1K0 likes484 downloads2y agoHugging Face14CocoBro /MMEdit-TestSet MMEdit Test Set A paired audio editing test set for text-guided audio manipulation evaluation, released with MMEdit. Overview This dataset contains 3,317 aligned triplets: Component Description raw/ Source audio before editing target/ Target audio after editing content.jsonl Editing instruction (caption) keyed by audio_id Each sample is linked by a shared audio_id. For example, sample add_017221 corresponds to: raw/add_017221.wav — original… See the full description on the dataset page: https://huggingface.co/datasets/CocoBro/MMEdit-TestSet.audioaudio-to-audio1K<n<10K0 likes470 downloads4mo agoHugging Face15bandad /asr-testset-kw-ja-v1 日本語 ASR アノテーション v1 重要:評価結果を報告する際の規約 本テストセットで評価結果を報告する際は、事前学習を含む学習にYouTubeの音声を使用したかどうかと、次の評価区分を必ず明記してください。 使用した場合:in-domain評価 使用していない場合:out-of-domain評価 この区分は、音声認識の評価結果を比較する上で重要です。最終的な追加学習だけでなく、使用するモデルの事前学習も含めて判断してください。 元データセットの音声に、人手で区間ごとの転記・タグ・採否を付けたデータです。音声の内容、ディレクトリ構成、ファイル名は元のままです。 ファイルと表示 **annotations.jsonl**:提出済みの全結果。1行が1音声です。 **metadata.jsonl**:HF表示用に自動生成したデータ。音声全体が不使用の行を除き、audio を file_name に置き換えています。 **data/**:採用した音声ファイル。 HFのビューアーは… See the full description on the dataset page: https://huggingface.co/datasets/bandad/asr-testset-kw-ja-v1.audioautomatic-speech-recognitionn<1K0 likes460 downloads2d agoHugging Face16tsinghua-ee /video-SALMONN_2_testset video-SALMONN 2 Benchmark Github Link Paper Link Generate the caption corresponding to the video and the audio with video_salmonn2_test.json Organize your results in the format like the following example: [ { "id": ["0.mp4"], "pred": "Generated Caption" } ] Replace res_file in eval.py with your result file. Run python3 eval.py videovideo-text-to-text0 likes457 downloads1y agoHugging Face17Isamu136 /indexed-open-image-v4-test-set Dataset Card for "indexed-open-image-v4-test-set" More Information needed image100K<n<1M1 likes432 downloads4y agoHugging Face18masumtechnonext /test-data-set-Arabic-letteraudio10K<n<100K0 likes414 downloads2mo agoHugging Face19Eureka-Leo /MCABSA_testsetaudion<1K2 likes335 downloads1y agoHugging Face20argilla-internal-testing /test_import_dataset_from_hub_using_settings_with_recordsTrue Dataset Card for test_import_dataset_from_hub_using_settings_with_recordsTrue This dataset has been created with Argilla. As shown in the sections below, this dataset can be loaded into your Argilla server as explained in Load with Argilla, or used directly with the datasets library in Load with datasets. Using this dataset with Argilla To load with Argilla, you'll just need to install Argilla as pip install argilla --upgrade and then use the following code: import… See the full description on the dataset page: https://huggingface.co/datasets/argilla-internal-testing/test_import_dataset_from_hub_using_settings_with_recordsTrue.textn<1K0 likes308 downloads2y agoHugging Face21allegrolab /testset_piqatext1K<n<10K0 likes295 downloads1y agoHugging Face22ganchengguang /MMM-datasets-TestsetMultilingual Mutual Reinforcement Effect Mix Datasets This is a Training set of OIELLM. This Train set already formatted by OIELLM's format. The test set is in the another page in huggingface. The MMM support 3 languages (English, Chinese and Japanese). And you must use task instruct words to define kind of task. Mutual Reinforcement Effect. OIELLM's input and output MMM Dataset The following is input and output format: { "input": "In 1953, filming of "On the Waterfront" starring… See the full description on the dataset page: https://huggingface.co/datasets/ganchengguang/MMM-datasets-Testset.text100K<n<1M1 likes294 downloads2y agoHugging Face23etiennebamas /inference-test-set1 likes277 downloads2mo agoHugging Face24Tengpaz /WorldRenderer-Testsetimage10K<n<100K0 likes273 downloads27d agoHugging Face25fanrui00 /FakeVV_testset_videovideon<1K0 likes263 downloads11mo agoHugging Face261xg /Easy-Turn-Testset Easy Turn: Integrating Acoustic and Linguistic Modalities for Robust Turn-Taking in Full-Duplex Spoken Dialogue Systems Guojian Li1, Chengyou Wang1, Hongfei Xue1, Shuiyuan Wang1, Dehui Gao1, Zihan Zhang2, Yuke Lin2, Wenjie Li2, Longshuai Xiao2, Zhonghua Fu1,╀, Lei Xie1,╀ 1 Audio, Speech and Language Processing Group (ASLP@NPU), Northwestern Polytechnical University 2 Huawei Technologies, China 🎤 Demo Page 🤖 Easy Turn Model 📑 Paper 🌐 Huggingface… See the full description on the dataset page: https://huggingface.co/datasets/1xg/Easy-Turn-Testset.automatic-speech-recognition0 likes249 downloads2mo agoHugging Face27swordhealth /MindGuard-testsetgated MindGuard-testset: Expert-Annotated Evaluation Data for Mental Health AI Safety MindGuard-testset is a clinically grounded benchmark dataset for evaluating safety classifiers in mental health AI systems. This dataset was developed by Sword Health in collaboration with licensed clinical psychologists to address the critical need for contextually appropriate safety measures in therapeutic AI applications. Overview MindGuard-testset contains 1,134 annotated user turns… See the full description on the dataset page: https://huggingface.co/datasets/swordhealth/MindGuard-testset.text1K<n<10K4 likes241 downloads8mo agoHugging Face28allegrolab /testset_popqatext1K<n<10K0 likes232 downloads1y agoHugging Face29allegrolab /testset_mmlutext1K<n<10K0 likes219 downloads1y agoHugging Face30allegrolab /testset_winogrande-infilltext1K<n<10K0 likes204 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.