datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Imagenet-1k_validationCOCO_captions_validation
Dataset Card for "COCO_captions_validation"
More Information needed
VQAv2_validation
Dataset Card for "VQAv2_validation"
More Information needed
Places365-ValidationTextVQA_validation
Dataset Card for "TextVQA_validation"
More Information needed
VQAv2_sample_validation
Dataset Card for "VQAv2_sample_validation"
More Information needed
getting-started-labeled-validation
Dataset Card for validation_photos
This is a FiftyOne dataset with 143 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("TheSteve0/getting-started-labeled-validation")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/getting-started-labeled-validation.getting-started-validation-clip-pred
Dataset Card for labeled_validation_predicted_clip
This is a FiftyOne dataset with 143 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("TheSteve0/getting-started-validation-clip-pred")
# Launch the App
session =… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/getting-started-validation-clip-pred.imagenet-1k-validation-subsetsImagenet1k_sample_validation
Dataset Card for "Imagenet1k_sample_validation"
More Information needed
ValidationDataSetVizWiz_validation
Dataset Card for "VizWiz_validation"
More Information needed
apriltag-validation-data
Official README
See the README.txt
Dataset Sources
A non-sharepoint hosting of the "ICRA 2020 - Determining and Improving the Localization Accuracy of AprilTag Detection" dataset
All rights belong to the original authors.
Repository: https://rzunibw-my.sharepoint.com/personal/thorsten_luettel_rzunibw_onmicrosoft_com/_layouts/15/onedrive.aspx?id=%2Fpersonal%2Fthorsten%5Fluettel%5Frzunibw%5Fonmicrosoft%5Fcom%2FDocuments%2Fdatasets%2Ficra2020%2Dapriltag%2Ddataset&ga=1… See the full description on the dataset page: https://huggingface.co/datasets/NoeFontana/apriltag-validation-data.watermarks-validationVQAv2_minival_validation_vprevious
Dataset Card for "VQA_minival_validation"
More Information needed
HUD-UI-Validation-Pilot
HUD/UI annotation validation pilot
This public artifact compares 10 gameplay screenshots across four columns:
raw frame;
GPT-5.6 Sol X-High final reference;
GPT-5.6 Terra High refined annotation;
confidence-routed final annotation.
Files:
analysis.md: aggregate metrics and per-game error table;
per_sample_metrics.csv: machine-readable sample metrics;
overlay_comparison_contact_sheet.jpg: full comparison sheet.
The 0--100 confidence value is a conservative pipeline routing… See the full description on the dataset page: https://huggingface.co/datasets/GameWorldData/HUD-UI-Validation-Pilot.VPPO_MMK12_validation
Dataset Card for VPPO_MMK12_validation
Dataset Details
Dataset Description
This dataset is the official validation split used to fine-tune the VPPO-7B and VPPO-32B models presented in our paper, "Spotlight on Token Perception for Multimodal Reinforcement Learning".
This is a direct copy of the test split of FanqingM/MMK12 dataset. We have isolated it here to ensure the exact version used in our experiments is publicly available, guaranteeing reproducibility for… See the full description on the dataset page: https://huggingface.co/datasets/chamber111/VPPO_MMK12_validation.medical_records_parsing_validation_set
Medical Records Parsing Validation Set
Dataset Composition and Clinical Relevance
The Eka Medical Records Parsing Dataset empowers evaluation of AI systems designed to extract structured information from unstructured medical documents, enabling true digitisation of healthcare data while maintaining clinical accuracy.
The dataset comprise 288 carefully selected images of laboratory reports and prescriptions representing diverse formats and templates encountered in Indian… See the full description on the dataset page: https://huggingface.co/datasets/ekacare/medical_records_parsing_validation_set.alt-text-validationThis dataset contains images and alt text from various sources.
It is used to control the quality of https://huggingface.co/Mozilla/distilvit using the https://github.com/mozilla/checkvite application
This application let users try out the model on the images and classify them. The dataset is then updated.
When an image is marked as need_training it will be use to fine-tune the model to fix some of its inaccuracies
bdd100k-validation-onlyMMMU-Reasoning-Distill-Validation中文版本
Description
MMMU-Reasoning-Distill-Validation is a Multi-Modal reasoning dataset that contains 839 image descriptions and natural language inference data samples. This dataset is built upon the validation set of MMMU. The construction process begins with using Qwen2.5-VL-72B-Instruct for image understanding and generating detailed image descriptions, followed by generating reasoning conversations using the DeepSeek-R1 model. Its main features are as follows:
Use the… See the full description on the dataset page: https://huggingface.co/datasets/modelscope/MMMU-Reasoning-Distill-Validation.OmniEdit-validation-datasetVQAv2_minival_validation
Dataset Card for "VQAv2_minival_validation_v2"
More Information needed
dentex-validation-imagesImagenette_validation
Dataset Card for "Imagenette_validation"
More Information needed
marvin-home-thumb-100-30hz-validation
Marvin home approach and thumb curl
10 validation episodes, 2337 synchronized rows at 30 Hz. Deterministic 90/10 episode split of 100 successful simulations. Overhead and wrist RGB are both stored; pi05_marvin_home_ego uses overhead only. Actions are absolute actuator targets in hand22+arm7 order; state is arm7+hand22. The OpenPI adapter aligns state order and applies stock arm delta transforms. Tactile channels are synthetic contact proxies, not real hand measurements. See… See the full description on the dataset page: https://huggingface.co/datasets/preload-bionics/marvin-home-thumb-100-30hz-validation.commtr-real-video-validation
COMMTR Real-Video Validation Dataset (Round 3 · Comment 4)
This package accompanies the manuscript "Agent-Pilot" (submitted to
Communications in Transportation Research, manuscript ID COMMTR-2026-0064)
and contains the data and code for the real-world, fixed-view, video-based
validation of the Agent-Pilot framework, added in Round 3 in response to
reviewer Comment 4 ("simulation alone is insufficient ... add real-world
video-based validation").
1. Overview of the… See the full description on the dataset page: https://huggingface.co/datasets/ChenglinLiuChris/commtr-real-video-validation.Chest_Xray_validation_N_Hotlatent_upscale_validationacouslic_ai_validation
