datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MPII_Human_Pose_Dataset
Dataset Card for MPII Human Pose
MPII Human Pose dataset is a state of the art benchmark for evaluation of articulated human pose estimation.
The dataset includes around 25K images containing over 40K people with annotated body joints.
The images were systematically collected using an established taxonomy of every day human activities.
Overall the dataset covers 410 human activities and each image is provided with an activity label.
Each image was extracted from a YouTube… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/MPII_Human_Pose_Dataset.700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3
NOTE: A newer version of this dataset is available Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Preference_Dataset
Rapidata Image Generation Preference Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
Link to the Text-2-Image Alignment dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3.Cabin-Human-Behavior-Dataset
全球最大的智能座舱多模态开源高质量数据集来啦!
一. 数据集摘要 (Dataset Summary)
「CyberData塞塔」智能座舱用户行为数据集是一个专为加速智能座舱感知算法开发而设计的高质量、程序化生成的图像数据集。随着 C-NCAP、EU GSR 等全球汽车安全法规对驾驶员监控系统 (DMS) 和乘客监控系统 (OMS) 提出更高要求,安全、合规、多样化的训练数据变得至关重要。本数据集通过合成方式,旨在解决真实世界数据采集面临的隐私风险、高昂成本和长尾场景覆盖不足等核心挑战。
该数据集包含 5,000 张 由 XAI Lab 自主研发的数据集生成引擎合成的高保真座舱内用户行为图像,每张图像都附带丰富的、100% 精确的标注信息。
核心特点:
丰富的场景多样性: 涵盖不同年龄、性别、种族和衣着风格的虚拟人模型,以及多种驾驶与乘坐行为(如使用手机、喝水、疲劳、手势)和面部表情。
专为座舱感知优化: 数据集可直接用于智能座舱端侧视觉模型,尤其是 DMS/OMS 算法的训练、微调与验证,帮助模型精准理解座舱内复杂的交互与状态。… See the full description on the dataset page: https://huggingface.co/datasets/OpenSparX/Cabin-Human-Behavior-Dataset.human_parsing_dataset
Dataset Card for Human parsing data (ATR)
Dataset Summary
This dataset has 17,706 images and mask pairs. It is just a copy of
Deep Human Parsing ATR dataset. The mask labels are:
"0": "Background",
"1": "Hat",
"2": "Hair",
"3": "Sunglasses",
"4": "Upper-clothes",
"5": "Skirt",
"6": "Pants",
"7": "Dress",
"8": "Belt",
"9": "Left-shoe",
"10": "Right-shoe",
"11": "Face",
"12": "Left-leg",
"13": "Right-leg",
"14":… See the full description on the dataset page: https://huggingface.co/datasets/mattmdjaga/human_parsing_dataset.Flux_SD3_MJ_Dalle_Human_Alignment_Dataset
NOTE: A newer version of this dataset is available Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Alignment_Dataset
Rapidata Image Generation Alignment Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
Link to the Preference dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Alignment_Dataset.Cabin-Human-ABNORMAL-Behavior-Dataset
全球最大的智能座舱多模态开源高质量数据集来啦!
一. 数据集摘要 (Dataset Summary)
「CyberData塞塔」智能座舱用户行为数据集是一个专为加速智能座舱感知算法开发而设计的高质量、程序化生成的图像数据集。随着 C-NCAP、EU GSR 等全球汽车安全法规对驾驶员监控系统 (DMS) 和乘客监控系统 (OMS) 提出更高要求,安全、合规、多样化的训练数据变得至关重要。本数据集通过合成方式,旨在解决真实世界数据采集面临的隐私风险、高昂成本和长尾场景覆盖不足等核心挑战。
该数据集包含 5,000 张 由 XAI Lab 自主研发的数据集生成引擎合成的高保真座舱内用户行为图像,每张图像都附带丰富的、100% 精确的标注信息。
数据格式
数据集以JSON格式提供,包含以下字段:
image_id: 图像ID
image_path: 图像路径
category: 行为类别
tags: 行为标签
behaviors: 包含左右乘客行为描述的对象
left_passenger: 左侧乘客行为描述… See the full description on the dataset page: https://huggingface.co/datasets/OpenSparX/Cabin-Human-ABNORMAL-Behavior-Dataset.Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
NOTE: A newer version of this dataset is available: Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Coherence_Dataset
Rapidata Image Generation Coherence Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Preference dataset: https://huggingface.co/datasets/Rapidata/700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3
Link to the Text-2-Image Alignment dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset.my-human-datasethuman_ego_data
Human Ego Data
Example RGB-D episodes collected with the
ego_realsense_tool for the
Ego2Sim benchmark.
Only tasks with completed successful recordings on the collection machine are
included. Each currently available task contributes one raw example episode.
Layout
Txx_task_name/
suite/
session_xxx/
raw/
frames/
*_color.png
*_depth.png
frames.csv
imu.csv
intrinsics.json
session.json… See the full description on the dataset page: https://huggingface.co/datasets/baiyu858/human_ego_data.text-2-image-human-preferences-2m
Text-to-image human preferences: 2M votes across 30 models
This dataset contains the complete voting record behind the
Datapoint Image Bench
leaderboard: 2,161,160 validated pairwise votes — exactly 10 for each of
216,116 image pairs. The votes compare 30 text-to-image models in a complete
round-robin on 500 prompts, judged by annotators from over 200 countries.
Every vote includes the annotator's trust score at the time the vote was
cast.
Built on the Datapoint annotation… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-human-preferences-2m.human_parsing_fashion_datasetswimsuit-human-segmentation-dataset
People Clothing Segmentation Dataset
Dataset comprises 14,358 high-quality photos of 7,179 people of diverse genders wearing bathing suits, each paired with detailed segmentation masks for precise body-parts segmentation. It designed for semantic and instance segmentation tasks, this large-scale collection offers manually annotated labels, enabling robust training for deep learning models in human body analysis.
By leveraging this dataset, researchers can train high-precision… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/swimsuit-human-segmentation-dataset.human_protein_atlas_cells_datasetai-vs-human-generated-datasetChatGPT-Gemini-Claude-Perplexity-Human-Evaluation-Multi-Aspects-Review-Dataset
ChatGPT Gemini Claude Perplexity Human Evaluation Multi Aspect Review Dataset
Introduction
Human evaluation and reviews with scalar score of AI Services responses are very usefuly in LLM Finetuning, Human Preference Alignment, Few-Shot Learning, Bad Case Shooting, etc, but extremely difficult to collect.
This dataset is collected from DeepNLP AI Service User Review panel (http://www.deepnlp.org/store), which is an open review website for users to give reviews and upload… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/ChatGPT-Gemini-Claude-Perplexity-Human-Evaluation-Multi-Aspects-Review-Dataset.human-faces-dataset
Human Faces Dataset
This repository contains an image classification dataset comparing AI-Generated face images and Real face images.
Total Images: 9,630
Dataset Structure:
ai_generated/: 4,630 images (split into two parts to avoid timeouts):
part_1/: 2,315 images
part_2/: 2,315 images
real/: 5,000 images (split into two parts to avoid timeouts):
part_1/: 2,500 images
part_2/: 2,500 images
Format: Image Classification folder structure (subfolders named after classes… See the full description on the dataset page: https://huggingface.co/datasets/shravya11/human-faces-dataset.uiclip_human_data_hfimage-2-video-human-preferences-large
I2V Human Preferences (Large)
Human preference dataset for image-to-video (I2V) generation quality. Each row contains a reference image, two generated videos (one from Pika and one from CogVideoX), and 10 human preference annotations aggregated via majority vote.
This is the large (3,000-row) subset — the complete dataset. See also: small (1,000 rows), medium (2,000 rows).
Dataset Summary
Metric
Value
Total rows
3,000
Annotations per row
10
Total… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/image-2-video-human-preferences-large.RBY1_human_data_0401_v8_egoengineThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "rby1",
"total_episodes": 93,
"total_frames": 13116,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": [
82,
51,
52,
88,
78,
65,
33,
8,
37… See the full description on the dataset page: https://huggingface.co/datasets/Daniel233/RBY1_human_data_0401_v8_egoengine.uiclip_human_data-paired_hftext-2-video-human-preferences-motion
Human Preferences for AI-Generated Video: Motion Quality
29,283 pairwise human preference labels comparing 4 frontier video generation models on human motion across 3 quality dimensions, collected from 4,349 real annotators via Datapoint AI.
This is the largest publicly available human preference dataset focused specifically on human motion in AI-generated video.
Why This Dataset
Video generation models are improving fast, but evaluating human motion remains… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-video-human-preferences-motion.text-2-image-dpo-human-preferences-full
Text-2-Image DPO Human Preferences (Full)
The complete human preference dataset for text-to-image generation. 416,360 pairwise judgments from ~20,000 annotators comparing AI-generated images across two evaluation dimensions: prompt alignment and overall preference.
This is the full, unfiltered version with uniform vote weights. For quality-filtered subsets with calibrated annotator weighting, see:
datapointai/text-2-image-dpo-human-preferences (5,000 pairs, trust-weighted)… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-dpo-human-preferences-full.human_skin_parsing_datasetLabels:
0: Background
1: Skin
2: Other
image-2-video-human-preferences-medium
I2V Human Preferences (Medium)
Human preference dataset for image-to-video (I2V) generation quality. Each row contains a reference image, two generated videos (one from Pika and one from CogVideoX), and 10 human preference annotations aggregated via majority vote.
This is the medium (2,000-row) subset. See also: small (1,000 rows), large (3,000 rows).
Dataset Summary
Metric
Value
Total rows
2,000
Annotations per row
10
Total annotations
20,000
Unique… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/image-2-video-human-preferences-medium.chris_human_datasetThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "panda",
"total_episodes": 3,
"total_frames": 925,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:3"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/brandonyang/chris_human_dataset.ai-vs-human-generated-dataset-sampleimage-2-video-human-preferences-small
I2V Human Preferences (Small)
Human preference dataset for image-to-video (I2V) generation quality. Each row contains a reference image, two generated videos (one from Pika and one from CogVideoX), and 10 human preference annotations aggregated via majority vote.
This is the small (1,000-row) subset. See also: medium (2,000 rows), large (3,000 rows).
Dataset Summary
Metric
Value
Total rows
1,000
Annotations per row
10
Total annotations
10,000
Unique… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/image-2-video-human-preferences-small.human-scan-dataset
tags:
- 3d
- human
- photogrammetry
- avatar
- pose-estimation
- digital-human
- body-scan
- dataset
pretty_name: HumanAlloy 3D Human Scan Dataset
size_categories:
- 1K<n<10K
HumanAlloy 3D Human Scan Dataset
Dataset Description
HumanAlloy is a large-scale library of real human subjects captured with a professional 180-camera photogrammetry rig in controlled studio conditions. Each subject produces a clean, consistent full-body 3D mesh with realistic… See the full description on the dataset page: https://huggingface.co/datasets/humanalloy/human-scan-dataset.signature-dataset-real-vs-forgedhuman_parsing_dataset_plus_neck
