datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MPII_Human_Pose_Dataset
Dataset Card for MPII Human Pose
MPII Human Pose dataset is a state of the art benchmark for evaluation of articulated human pose estimation.
The dataset includes around 25K images containing over 40K people with annotated body joints.
The images were systematically collected using an established taxonomy of every day human activities.
Overall the dataset covers 410 human activities and each image is provided with an activity label.
Each image was extracted from a YouTube… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/MPII_Human_Pose_Dataset.human_data_for_openvlaHSTLI_A-Dataset-of-Human-Semen-Time-Lapse-Images
HSTLI: A Dataset of Human Semen Time Lapse Images
Dataset Details
HSTLI contains 3,266 time-lapse microscopy videos of human sperm.Clips were recorded from two imaging modalities:
CASA system (Sperm Class Analyzer)
Optical microscope (Swift M10DB-MP + Fujifilm X-T30)
A subset of videos was manually annotated with bounding boxes around each visible sperm head.
The dataset supports detection, tracking and motility computation.
Total contents:
34… See the full description on the dataset page: https://huggingface.co/datasets/DFL-KamLab/HSTLI_A-Dataset-of-Human-Semen-Time-Lapse-Images.LA_dataset_human_made_franka_eef_Lerobotv21_260511700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3
NOTE: A newer version of this dataset is available Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Preference_Dataset
Rapidata Image Generation Preference Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
Link to the Text-2-Image Alignment dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3.Cabin-Human-Behavior-Dataset
全球最大的智能座舱多模态开源高质量数据集来啦!
一. 数据集摘要 (Dataset Summary)
「CyberData塞塔」智能座舱用户行为数据集是一个专为加速智能座舱感知算法开发而设计的高质量、程序化生成的图像数据集。随着 C-NCAP、EU GSR 等全球汽车安全法规对驾驶员监控系统 (DMS) 和乘客监控系统 (OMS) 提出更高要求,安全、合规、多样化的训练数据变得至关重要。本数据集通过合成方式,旨在解决真实世界数据采集面临的隐私风险、高昂成本和长尾场景覆盖不足等核心挑战。
该数据集包含 5,000 张 由 XAI Lab 自主研发的数据集生成引擎合成的高保真座舱内用户行为图像,每张图像都附带丰富的、100% 精确的标注信息。
核心特点:
丰富的场景多样性: 涵盖不同年龄、性别、种族和衣着风格的虚拟人模型,以及多种驾驶与乘坐行为(如使用手机、喝水、疲劳、手势)和面部表情。
专为座舱感知优化: 数据集可直接用于智能座舱端侧视觉模型,尤其是 DMS/OMS 算法的训练、微调与验证,帮助模型精准理解座舱内复杂的交互与状态。… See the full description on the dataset page: https://huggingface.co/datasets/OpenSparX/Cabin-Human-Behavior-Dataset.tokenization-multiplicity-data
Dataset: Tokenization Multiplicity Leads to Arbitrary Price Variation in LLM-as-a-service
This dataset contains the official experiment inference traces for the paper Tokenization Multiplicity Leads to Arbitrary Price Variation in LLM-as-a-service by Ivi Chatzi, Nina Corvelo Benz, Stratis Tsirtsis and Manuel Gomez-Rodriguez.
📂 Dataset Structure
The dataset is organized into folders as follows:
.\{model}\{task}\{lang}\{seed}_{10*temperature}.jsonl
where {model}… See the full description on the dataset page: https://huggingface.co/datasets/Human-Centric-Machine-Learning/tokenization-multiplicity-data.human_parsing_dataset
Dataset Card for Human parsing data (ATR)
Dataset Summary
This dataset has 17,706 images and mask pairs. It is just a copy of
Deep Human Parsing ATR dataset. The mask labels are:
"0": "Background",
"1": "Hat",
"2": "Hair",
"3": "Sunglasses",
"4": "Upper-clothes",
"5": "Skirt",
"6": "Pants",
"7": "Dress",
"8": "Belt",
"9": "Left-shoe",
"10": "Right-shoe",
"11": "Face",
"12": "Left-leg",
"13": "Right-leg",
"14":… See the full description on the dataset page: https://huggingface.co/datasets/mattmdjaga/human_parsing_dataset.Flux_SD3_MJ_Dalle_Human_Alignment_Dataset
NOTE: A newer version of this dataset is available Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Alignment_Dataset
Rapidata Image Generation Alignment Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Coherence dataset: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
Link to the Preference dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Alignment_Dataset.text-to-speech-human-preferences-315k
Text-to-speech human preferences: 315K votes across 15 models
This gated dataset contains the evaluation record behind Datapoint Audio
Bench: 315,000 eligible pairwise votes comparing 15 text-to-speech
models in a complete round-robin over 300 English prompts. The prompt set
covers eight practical voice-agent categories, and every generated sample is
included as a typed audio record.
The source evaluation collected 357,651 completed responses. The published
benchmark excluded… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-to-speech-human-preferences-315k.Flux_SD3_MJ_Dalle_Human_Coherence_Dataset
NOTE: A newer version of this dataset is available: Imagen3_Flux1.1_Flux1_SD3_MJ_Dalle_Human_Coherence_Dataset
Rapidata Image Generation Coherence Dataset
This Dataset is a 1/3 of a 2M+ human annotation dataset that was split into three modalities: Preference, Coherence, Text-to-Image Alignment.
Link to the Preference dataset: https://huggingface.co/datasets/Rapidata/700k_Human_Preference_Dataset_FLUX_SD3_MJ_DALLE3
Link to the Text-2-Image Alignment dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/Flux_SD3_MJ_Dalle_Human_Coherence_Dataset.ai-vs-human-rubric-companion-data
Companion dataset for the AI-vs-human rubric study
This dataset is the data side of an anonymous submission. It pairs with a separate anonymous code repository that contains the runnable scripts, validators, and documentation. The two artifacts together reproduce every paper-facing headline number without re-running any API-backed stage. The code URL for review is https://anonymous.4open.science/r/codereviewer-47F3/README.md.
Paper sections and where their data are… See the full description on the dataset page: https://huggingface.co/datasets/forreview43/ai-vs-human-rubric-companion-data.dopabase-human-data009_blender_vault_dataset_11000_11999human_performance_dataset
Human performance dataset
Piano MIDI dataset used for training and evaluation. Layout and processing:
Source: MAESTRO (or similar) with a prompt (or user_prompt) per piece.
Augmentation: dataset_preprocess/augment_dataset.py adds pitch/time/velocity variants; output CSV lists originals and augmented files.
Grouping: dataset_preprocess/group_dataset.py copies MIDIs into grouped/ as composer/genre/filename.mid (and augmented/composer/genre/ for augmented). Genre is taken from the… See the full description on the dataset page: https://huggingface.co/datasets/cnmat/human_performance_dataset.Human-Like-DPO-Dataset
Enhancing Human-Like Responses in Large Language Models
🤗 Models | 📊 Dataset | 📄 Paper
📢 The paper associated with this dataset has been accepted to the AAAI-26 Workshop on Personalization in the Era of Large Foundation Models (PerFM).
Human-Like-DPO-Dataset
This dataset was created as part of research aimed at improving conversational fluency and engagement in large language models. It is suitable for formats like Direct Preference Optimization (DPO) to guide… See the full description on the dataset page: https://huggingface.co/datasets/HumanLLMs/Human-Like-DPO-Dataset.Cabin-Human-ABNORMAL-Behavior-Dataset
全球最大的智能座舱多模态开源高质量数据集来啦!
一. 数据集摘要 (Dataset Summary)
「CyberData塞塔」智能座舱用户行为数据集是一个专为加速智能座舱感知算法开发而设计的高质量、程序化生成的图像数据集。随着 C-NCAP、EU GSR 等全球汽车安全法规对驾驶员监控系统 (DMS) 和乘客监控系统 (OMS) 提出更高要求,安全、合规、多样化的训练数据变得至关重要。本数据集通过合成方式,旨在解决真实世界数据采集面临的隐私风险、高昂成本和长尾场景覆盖不足等核心挑战。
该数据集包含 5,000 张 由 XAI Lab 自主研发的数据集生成引擎合成的高保真座舱内用户行为图像,每张图像都附带丰富的、100% 精确的标注信息。
数据格式
数据集以JSON格式提供,包含以下字段:
image_id: 图像ID
image_path: 图像路径
category: 行为类别
tags: 行为标签
behaviors: 包含左右乘客行为描述的对象
left_passenger: 左侧乘客行为描述… See the full description on the dataset page: https://huggingface.co/datasets/OpenSparX/Cabin-Human-ABNORMAL-Behavior-Dataset.009_blender_vault_dataset_10000_10999009_blender_vault_dataset_0_999strategic-ttc-data
Dataset: Strategic Test-Time Compute (TTC)
This dataset contains the official experiment inference traces for the paper "Test-Time Compute Games" (arXiv:2601.21839).
It includes full model generations, token counts, and correctness verifications for various Large Language Models (LLMs) across three major reasoning benchmarks: GSM8K, AIME, and GPQA.
This data allows researchers to analyze the relationship between test-time compute and model performance without needing to re-run… See the full description on the dataset page: https://huggingface.co/datasets/Human-Centric-Machine-Learning/strategic-ttc-data.playworld_human_demo_data
PlayWorld Human Demonstration Data
Human teleoperation demonstrations for robotic manipulation tasks.
Dataset Description
This dataset contains human demonstrations collected via teleoperation for robotic manipulation tasks.
Each episode includes:
Multi-view RGB videos (3 cameras: exterior_1, exterior_2, wrist)
Pre-encoded video latents (PyTorch .pt files)
Robot state trajectories (7-DOF: x, y, z, roll, pitch, yaw, gripper)
Task descriptions and success labels… See the full description on the dataset page: https://huggingface.co/datasets/tennyyyin/playworld_human_demo_data.human_data_contact
human_data 접촉 라벨 (3 dataset · 259 episode · 104,818 step)
사람이 원격조작한 teleoperation dataset 세 벌의 frame 마다 손(gripper)이 물체에 닿아 있는가를
0/1 로 적은 라벨이다. 같은 형식의 LIBERO 판이 TTaekwan/libero_contact, RoboCasa 판이
TTaekwan/robocasa_contact 에 있다.
폴더 셋
폴더
episode
step
fps
contact 비율
라벨을 만든 방법
openarm_contact
107
66,203
30
0.0919
사람이 두 egocentric camera 영상을 보고 구간을 직접 찍었다
pnp_task_contact
101
14,902
15
0.4054
tactile(오른손) Otsu threshold + gripper 열림으로 끝을 잡고, 33 episode 는… See the full description on the dataset page: https://huggingface.co/datasets/TTaekwan/human_data_contact.human-telemetry-driving-dataset-lite-version
Dataset Card for 15 Laps of 30Hz NGSIM-Style Telemetry
This is a Lite Version of a larger research dataset focusing on human driving signatures in high-fidelity simulations. It includes 15 full laps of telemetry captured at 30Hz within Unreal Engine 5, specifically formatted to match NGSIM standards.
Dataset Details
Dataset Description
This Lite Version dataset contains 15 laps of high-fidelity human driving telemetry. It is intended for researchers and… See the full description on the dataset page: https://huggingface.co/datasets/AtlasBuiltIt/human-telemetry-driving-dataset-lite-version.furry_dataset_e621_captions_claims_human_reviewedAbout 13,000 human reviewed captions. ~5000 of them were reviewed by me, the rest by independent workers.
I cannot confirm they will have caught all of the mistakes (there will still be some small mistakes). But the accuracy of these captions is higher than what any VLM can produce.
The "claims" are one-liner claims about an image, and given a truth value. Mostly machine-verified, but the ~10000 human-reviewed ones are a very valuable set of sex-related items or items which powerful VLMs had… See the full description on the dataset page: https://huggingface.co/datasets/furproxy/furry_dataset_e621_captions_claims_human_reviewed.human_ego_data
Human Ego Data
Example RGB-D episodes collected with the
ego_realsense_tool for the
Ego2Sim benchmark.
Only tasks with completed successful recordings on the collection machine are
included. Each currently available task contributes one raw example episode.
Layout
Txx_task_name/
suite/
session_xxx/
raw/
frames/
*_color.png
*_depth.png
frames.csv
imu.csv
intrinsics.json
session.json… See the full description on the dataset page: https://huggingface.co/datasets/baiyu858/human_ego_data.my-human-dataset009_blender_vault_dataset_9000_9999human-movement-pov-dataset-for-robotics
VLA Egocentric Video Dataset — 100+ Hours of 4K Human Manipulation
100+ hours of first-person (egocentric POV) video of real humans performing real-life hand tasks for training vision-language-action (VLA) models, imitation learning policies, and embodied AI systems
Key Highlights
100+ hours of real-world egocentric video
Continuous, uncut - full task arcs (preparation → execution → result)
Head-mounted POV - matches robot wrist/head sensor geometry better than… See the full description on the dataset page: https://huggingface.co/datasets/AxonData/human-movement-pov-dataset-for-robotics.human-microbe-dataset
Human Microbe Evidence Dataset
Multi-source microbial evidence dataset focused on taxa associated
with human infection, disease, clinical literature, and pathogen
genomic evidence.
Current release
This is an intermediate checkpoint of the pipeline.
Current stage:
PubMed discovery + NCBI Taxonomy + disease evidence + ChEMBL +
NCBI Pathogen Detection indexing
Important interpretation
PubMed literature signals represent evidence of association and
do… See the full description on the dataset page: https://huggingface.co/datasets/prakhya15/human-microbe-dataset.009_blender_vault_dataset_2000_2999
