Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01LPY /BridgeVLA_COLOSSEUM_EVAL_DATAarxiv: https://arxiv.org/abs/2506.07961 image1M<n<10M0 likes1.7k downloads1y agoHugging Face02LongVideo-Reason /longvideo_eval_videos Long-RL: Scaling RL to Long Sequences (Evaluation Dataset - for research only) Data Distribution We strategically construct a high-quality dataset with CoT annotations for long video reasoning, named LongVideo-Reason. Leveraging a powerful VLM (NVILA-8B) and a leading open-source reasoning LLM, we develop a dataset comprising 52K high-quality Question-Reasoning-Answer pairs for long videos. We use 18K high-quality samples for Long-CoT-SFT to initialize… See the full description on the dataset page: https://huggingface.co/datasets/LongVideo-Reason/longvideo_eval_videos.text1K<n<10K1 likes1.5k downloads1y agoHugging Face03DL3DV /DL3DV-Evaluationgated DL3DV Testing Split Download Instructions This repo contains all 55 scenes for evaluation. Note: it is an independent dataset, and none of its scenes overlap with those in DL3DV-10K. Have a galance on the preview page: https://dl3dv-10k.github.io/DL3DV-Testing-Split-Preview/. Download As the whole benchmark dataset is ~500G, a python script to download and untar files. Environment Setup The download script relies on huggingface hub, tqdm. You can download by… See the full description on the dataset page: https://huggingface.co/datasets/DL3DV/DL3DV-Evaluation.image100K<n<1M6 likes863 downloads1y agoHugging Face04ASLP-lab /WSYue-ASR-eval WSYue-ASR-eval: Cantonese ASR Benchmark To address the unique linguistic characteristics of Cantonese in speech recognition, we propose WSYue-ASR-eval, a benchmark specifically designed for evaluating Cantonese ASR systems. It is tailored to assess model performance across diverse lengths, domains, and linguistic phenomena of Cantonese speech. The test set annotations are provided by Beijing AISHELL Technology Co., Ltd. Key features: Annotated through multiple rounds of manual… See the full description on the dataset page: https://huggingface.co/datasets/ASLP-lab/WSYue-ASR-eval.text1K<n<10K4 likes443 downloads1y agoHugging Face05luuuulinnnn /Nurec_eval Nurec_eval — NRE Closed-Loop Eval Set with HD Map 59 OOD-longtail scenes for AlpaSim closed-loop evaluation, exported from NVIDIA NRE 26.02 and augmented with HD-map artifacts (lane graph, road boundaries, crosswalks) so models can be evaluated with full map context — not just geometry. What's inside each pai_<uuid>.usdz File Purpose default.usda USDZ entry point checkpoint.ckpt Neural reconstruction (NRE-26.02 gsplat) weights parsed_config.yaml NRE… See the full description on the dataset page: https://huggingface.co/datasets/luuuulinnnn/Nurec_eval.3droboticsn<1K0 likes313 downloads5mo agoHugging Face06semi-truths /Semi-Truths-Evalset Semi-Truths: The Evaluation Sample Recent efforts have developed AI-generated image detectors claiming robustness against various augmentations, but their effectiveness remains unclear. Can these systems detect varying degrees of augmentation? To address these questions, we introduce Semi-Truths, featuring 27,600 real images, 245,300 masks, and 850,200 AI-augmented images featuring varying degrees of targeted and localized edits, created using diverse augmentation methods… See the full description on the dataset page: https://huggingface.co/datasets/semi-truths/Semi-Truths-Evalset.image10K<n<100K3 likes258 downloads2y agoHugging Face07LPY /BridgeVLA_RLBench_EVAL_DATAarxiv: https://arxiv.org/abs/2506.07961 image1M<n<10M0 likes164 downloads1y agoHugging Face08wyhhey /twoframe-eval-artifacts-20260505 TwoFrame Eval Artifacts 2026-05-05 Generated image artifacts for TwoFrame image-editing evaluation. The large image payload is stored as tar archives under archives/ to avoid uploading tens of thousands of loose PNG files. Layout archives/single_ref.tar: all complete single-reference outputs. archives/multiref_part*.tar: complete K=2/K=3 multi-reference runs, sharded by run name. archives/metadata.tar: README, index, manifests, and metrics as an archive.… See the full description on the dataset page: https://huggingface.co/datasets/wyhhey/twoframe-eval-artifacts-20260505.textimage-to-imagen<1K0 likes142 downloads5mo agoHugging Face09haodongli /DA-2-Evaluation DA2: Depth Anything in Any Direction DA2 predicts dense, scale-invariant distance from a single 360° panorama in an end-to-end manner, with remarkable geometric fidelity and strong zero-shot generalization. 🎮 Usage Please see here. 🎓 Citation If you find these datasets useful, please consider citing 🌹: @article{li2025depth, title={DA$^{2}$: Depth Anything in Any Direction}, author={Li, Haodong and Zheng, Wangguangdong and He, Jing and Liu, Yuhao and… See the full description on the dataset page: https://huggingface.co/datasets/haodongli/DA-2-Evaluation.imagedepth-estimation1K<n<10K5 likes123 downloads6mo agoHugging Face10zszhong /Lyra-Evalaudio10K<n<100K1 likes105 downloads2y agoHugging Face11PixArt-alpha /PixArt-Eval-30Kimage10K<n<100K4 likes91 downloads2y agoHugging Face12aidealab /aidealab-videojp-eval AIdeaLab VideoJP 評価再現用データ はじめに このリポジトリはAIdeaLab VideoJPのFVDを測定するためのデータを 集めました。再現手順を次のとおりに示します。 評価方法 まず、評価用ライブラリをダウンロードします。 git clone https://github.com/JunyaoHu/common_metrics_on_video_quality ダウンロードできたら、ライブラリのインストール手順を踏んで、インストールします。 インストールしたら、同じディレクトリに次のファイルをコピーしてください evaluate_videos.py videos.tar gen_ja.tar コピーできたら、videos.tarとgen_ja.tarを展開します。 tar xf videos.tar tar xf gen_ja.tar 最後にevaluate_videos.pyを実行すると、FVDが表示されるはずです。 おまけ: 評価用映像の作り方… See the full description on the dataset page: https://huggingface.co/datasets/aidealab/aidealab-videojp-eval.texttext-to-video1K<n<10K0 likes88 downloads1y agoHugging Face13Helios1208 /Kling-Audio-Eval-cachetext10K<n<100K1 likes72 downloads5mo agoHugging Face14Ray2333 /Offline_Evaluationimage10K<n<100K0 likes39 downloads8mo agoHugging Face15GUI-Libra /Offline_Evaluationimage10K<n<100K0 likes33 downloads8mo agoHugging Face16rishitdagli /see-2-sound-evalWe sample images from Laion400M and the web to construct this small evaluation set. imagen<1K1 likes28 downloads2y agoHugging Face17Infektyd /syntra-testing-evals-v2textn<1K0 likes25 downloads8mo agoHugging Face18LEGO-Eval /object_images Overview This dataset contains 2D rendered images generated from 3D assets originally sourced from Objaverse and Objathor.The purpose of this dataset is to provide a large-scale collection of photo-realistic renderings for research on vision, multimodal learning, and text-to-3D understanding. Following prior works such as Diffusion4D and Stable-Zero123,we release only the rendered 2D images (not the original 3D assets) to facilitate efficient experimentation while preserving the… See the full description on the dataset page: https://huggingface.co/datasets/LEGO-Eval/object_images.image100K<n<1M0 likes23 downloads11mo agoHugging Face19jayw /t2v-gen-evaltextn<1K4 likes18 downloads3y agoHugging Face20vrsp-mcq /cls-evaluation-datasetimage10K<n<100K0 likes18 downloads1y agoHugging Face21grenoble /worldarena_832_480_and_abot_eval_ESRGAN.tarimage1K<n<10K0 likes11 downloads5mo agoHugging Face22grenoble /ABot-PhysWorld_eval_640_480_HYPIR_vace_skeleton_3250steptext1K<n<10K0 likes11 downloads5mo agoHugging Face23yslan /GaussianAnything-evalimage100K<n<1M0 likes10 downloads1y agoHugging Face24ynchen11 /PAPO-Evaltext10K<n<100K0 likes10 downloads10mo agoHugging Face25shmublu /membership-task-evaltext1K<n<10K0 likes8 downloads9mo agoHugging Face26nerako /ori_coco_evalimage1K<n<10K0 likes7 downloads1y agoHugging Face27xuhaorran /WSYue-ASR-eval WSYue-ASR-eval: Cantonese ASR Benchmark To address the unique linguistic characteristics of Cantonese in speech recognition, we propose WSYue-ASR-eval, a benchmark specifically designed for evaluating Cantonese ASR systems. It is tailored to assess model performance across diverse lengths, domains, and linguistic phenomena of Cantonese speech. The test set annotations are provided by Beijing AISHELL Technology Co., Ltd. Key features: Annotated through multiple rounds of manual… See the full description on the dataset page: https://huggingface.co/datasets/xuhaorran/WSYue-ASR-eval.text1K<n<10K0 likes7 downloads10mo agoHugging Face28nerako /relo_coco_evalimage1K<n<10K0 likes6 downloads1y agoHugging Face29shmublu /membership-evaltext10K<n<100K0 likes6 downloads9mo agoHugging Face30Infektyd /syntra-testing-evalstextn<1K0 likes5 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.