Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lmarena-ai /arena-human-preference-55kDataset for Kaggle competition on predicting human preference on Chatbot Arena battles. The training dataset includes over 55,000 real-world user and LLM conversations and user preferences across over 70 state-of-the-art LLMs, such as GPT-4, Claude 2, Llama 2, Gemini, and Mistral models. Each sample represents a battle consisting of 2 LLMs which answer the same question, with a user label of either prefer model A, prefer model B, tie, or tie (both bad). Citation Please cite the… See the full description on the dataset page: https://huggingface.co/datasets/lmarena-ai/arena-human-preference-55k.tabulartext-classification10K<n<100K159 likes2.4k downloads2y agoHugging Face02Humanbased-AI /MM-Food-100K Overview This project aims to introduce and release a comprehensive food image dataset designed specifically for computer vision tasks, particularly food recognition, classification, and nutritional analysis. We hope this dataset will provide a reliable resource for researchers and developers to advance the field of food AI. By publishing on Hugging Face, we expect to foster community collaboration and accelerate innovation in applications such as smart recipe recommendations… See the full description on the dataset page: https://huggingface.co/datasets/Humanbased-AI/MM-Food-100K.imageimage-classification100K<n<1M79 likes463 downloads1y agoHugging Face03RicardoRei /wmt-mqm-human-evaluation Dataset Summary This dataset contains all MQM human annotations from previous WMT Metrics shared tasks and the MQM annotations from Experts, Errors, and Context. The data is organised into 8 columns: lp: language pair src: input text mt: translation ref: reference translation score: MQM score system: MT Engine that produced the translation annotators: number of annotators domain: domain of the input text (e.g. news) year: collection year You can also find the original data here.… See the full description on the dataset page: https://huggingface.co/datasets/RicardoRei/wmt-mqm-human-evaluation.tabular100K<n<1M1 likes420 downloads4y agoHugging Face04AtlasBuiltIt /human-telemetry-driving-dataset-lite-version Dataset Card for 15 Laps of 30Hz NGSIM-Style Telemetry This is a Lite Version of a larger research dataset focusing on human driving signatures in high-fidelity simulations. It includes 15 full laps of telemetry captured at 30Hz within Unreal Engine 5, specifically formatted to match NGSIM standards. Dataset Details Dataset Description This Lite Version dataset contains 15 laps of high-fidelity human driving telemetry. It is intended for researchers and… See the full description on the dataset page: https://huggingface.co/datasets/AtlasBuiltIt/human-telemetry-driving-dataset-lite-version.tabulartabular-classification1K<n<10K0 likes267 downloads5mo agoHugging Face05NoeFlandre /landuse-sentence-relevance-golden-human-set Land-use sentence relevance golden human set This release contains the final 300-row V3 benchmark in English plus one parallel CSV for each of the 84 non-English project-provided sat-3l-sm language codes. There are 85 language files in total. Files Every file is at data/translations/<iso>/v3-final-<iso>.csv. The nine columns are: sentence, label, polygon_name, h3_cell, latitude, longitude, source, region, source_url. The Dataset Viewer exposes these files as 85… See the full description on the dataset page: https://huggingface.co/datasets/NoeFlandre/landuse-sentence-relevance-golden-human-set.tabulartext-classification10K<n<100K0 likes227 downloads19d agoHugging Face06RicardoRei /wmt-da-human-evaluation Dataset Summary This dataset contains all DA human annotations from previous WMT News Translation shared tasks. The data is organised into 8 columns: lp: language pair src: input text mt: translation ref: reference translation score: z score raw: direct assessment annotators: number of annotators domain: domain of the input text (e.g. news) year: collection year You can also find the original data for each year in the results section https://www.statmt.org/wmt{YEAR}/results.html… See the full description on the dataset page: https://huggingface.co/datasets/RicardoRei/wmt-da-human-evaluation.tabular1M<n<10M10 likes217 downloads4y agoHugging Face07OliverUrbann /HumanoidRobotSoccer Fall Prediction Dataset for Humanoid Robots Dataset Summary This dataset consists of 37.9 hours of real-world sensor data collected from 20 Nao humanoid robots over the course of one year in various test environments, including RoboCup soccer matches. The dataset includes 18.3 hours of walking data, featuring 2519 falls. It captures a wide range of activities such as omni-directional walking, collisions, standing up, and falls on various surfaces like artificial turf and… See the full description on the dataset page: https://huggingface.co/datasets/OliverUrbann/HumanoidRobotSoccer.tabular10M<n<100M2 likes212 downloads1y agoHugging Face08contralabs /HumanCreativityBenchmark The Human Creativity Benchmark (HCB) Expert evaluations of AI-generated creative work, built to separate two signals that single-score benchmarks collapse: convergence, where professionals align around shared, checkable standards, and divergence, where creative taste legitimately differs. Each AI output is judged by domain professionals through three complementary lenses — forced-choice pairwise comparisons, 1-5 scalar ratings on prompt adherence, usability, and visual appeal… See the full description on the dataset page: https://huggingface.co/datasets/contralabs/HumanCreativityBenchmark.imagetext-to-image1K<n<10K2 likes209 downloads4mo agoHugging Face09humanlong /emotion-negotiation-benchmarks Emotion-Aware LLM Negotiation Benchmarks Four high-stakes, edge-deployable negotiation benchmarks — the official evaluation suite for our research program on emotion-aware LLM agents. Each benchmark targets a distinct domain where (a) LLM-vs-LLM negotiation has real-world consequences, and (b) on-device deployment of small language models matters for privacy and latency. The benchmarks were originally introduced with EmoMAS (ACL 2026 Main, top 9% of 12,148 submissions) and are… See the full description on the dataset page: https://huggingface.co/datasets/humanlong/emotion-negotiation-benchmarks.tabulartext-generationn<1K0 likes151 downloads4mo agoHugging Face10latkes /humaneval-rerun-scorestabular100K<n<1M0 likes142 downloads4mo agoHugging Face11arcprize /arc_agi_2_human_testing ARC-AGI-2 Human testing data This file contains data from human testing sessions on ARC-AGI tasks. Each row represents a single test attempt by a human participant on a specific task-test pair in the "Public Train" or "Public Eval" ARC-AGI-2 datasets. Not all tasks in the released "Public Train" sets were tested, so these results are not comprehensive. This data does not include tasks from "Semi Private Evaluation" or "Private Evaluation" Column Descriptions… See the full description on the dataset page: https://huggingface.co/datasets/arcprize/arc_agi_2_human_testing.tabular1K<n<10K9 likes131 downloads1y agoHugging Face12ClarusC64 /autonomous-driving-human-vehicle-coupling-coherence-scoring-v0.1What this dataset tests Whether a system can score coherence between driver state, vehicle behavior, and scene context. This is not crash prediction. It is coupling integrity. Required outputs coupling_coherence_score overassertive_flag underassertive_flag trust_stability_index takeover_risk_score recovery_margin Scoring conventions all scores range 0 to 1 flags are 0 or 1 takeover risk estimates likelihood of manual override in the next window Use case Layer two of… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/autonomous-driving-human-vehicle-coupling-coherence-scoring-v0.1.tabulartabular-classificationn<1K0 likes123 downloads8mo agoHugging Face13sparklessszzz /InstaArt-HumanAI Instagram AI Art vs Human Art: Engagement & Comment Dataset Dataset Summary This dataset was created and contributed by Akshaya, Cynthia, Grace, and Soham as part of a project at UC San Diego. This dataset supports research into how audiences engage with AI-generated art versus human-made art on Instagram, with a specific focus on comment sentiment, reaction types, and engagement patterns. It consists of 40 matched pairs of Instagram posts - one human art post and one… See the full description on the dataset page: https://huggingface.co/datasets/sparklessszzz/InstaArt-HumanAI.tabulartext-classificationn<1K1 likes111 downloads7mo agoHugging Face14Shaow /humanbreast_xenium_janesick # Human Breast Cancer Xenium · Sample 1 Rep1+Rep2 Curated, ready-to-load spatial transcriptomics dataset. ## Source - Paper: [Janesick et al., Nat. Commun. 2023](https://www.nature.com/articles/s41467-023-43458-x) - Canonical download: cf.10xgenomics.com/samples/xenium/1.0.1/Xenium_FFPE_Human_Breast_Cancer_Rep{1,2} ## Scale | Property | Value | |---|---| | Technology | 10x Genomics Xenium (313-gene panel) | | Species | Homo sapiens | | Tissue |… See the full description on the dataset page: https://huggingface.co/datasets/Shaow/humanbreast_xenium_janesick.tabular100K<n<1M0 likes111 downloads5mo agoHugging Face15openaging /human_methylation_bench_ver1_train Human DNA Methylation Dataset ver1 This dataset is a benchmark dataset for predicting the aging clock, curated from publicly available DNA methylation data. The original benchmark dataset was published by Dmitrii Kriukov et al. (2024) by integrating data from 65 individual studies. To improve usability, we ensured unique sample IDs (excluding duplicate data, GSE118468 and GSE118469) and randomly split the data into training and testing subsets (train : test = 7 : 3) to… See the full description on the dataset page: https://huggingface.co/datasets/openaging/human_methylation_bench_ver1_train.tabular10K<n<100K1 likes71 downloads2y agoHugging Face16ABSTRACTION-ERC /subCat-human SubCat: A Dataset of Subordinate Categories in Human Mind and LLMs for the Italian Language A psycholinguistic italian dataset released with the paper How Humans and LLMs Organize Conceptual Knowledge: Exploring Subordinate Categories in Italian. It contains a list of subordiante categories, or exemplars, for 187 concrete words or, basic-level categories. Dataset Creation The dataset was created to study how Italian L1 speakers generate exemplars for common… See the full description on the dataset page: https://huggingface.co/datasets/ABSTRACTION-ERC/subCat-human.tabular1K<n<10K0 likes64 downloads1y agoHugging Face17RicardoRei /wmt-sqm-human-evaluation Dataset Summary In 2022, several changes were made to the annotation procedure used in the WMT Translation task. In contrast to the standard DA (sliding scale from 0-100) used in previous years, in 2022 annotators performed DA+SQM (Direct Assessment + Scalar Quality Metric). In DA+SQM, the annotators still provide a raw score between 0 and 100, but also are presented with seven labeled tick marks. DA+SQM helps to stabilize scores across annotators (as compared to DA). The data is… See the full description on the dataset page: https://huggingface.co/datasets/RicardoRei/wmt-sqm-human-evaluation.tabular100K<n<1M1 likes59 downloads4y agoHugging Face18Ragab-Adel /privacy-preserving-real-world-human-motion-sample Privacy-Preserving Real-World Human Motion Sample A market-validation sample of anonymous 2D skeleton/pose observations derived from a real-world indoor CCTV stream. Why this sample exists We are validating demand for continuously collected, privacy-oriented real-world human-motion data before expanding to multi-camera releases. Current public sample 750 public observations derived pose/skeleton data anonymous track identifiers no raw RGB video no… See the full description on the dataset page: https://huggingface.co/datasets/Ragab-Adel/privacy-preserving-real-world-human-motion-sample.tabularothern<1K1 likes58 downloads2mo agoHugging Face19SingleBicycle /pico_human_demo_v1 PICO 4 Ultra 真人第一视角示范数据 v1 PICO 4 Ultra 头显、人裸手操作桌面物体的第一视角数据。每条包含:第一视角 RGB 视频、头部 6DoF 轨迹、 PICO 原生裸手追踪的双手 26 关节(位置 + 朝向 + 有效标记)、每一帧的相机内参和相机位姿,以及逐帧对应表。 没有机器人数据。 下载 pip install -U huggingface_hub hf download SingleBicycle/pico_human_demo_v1 --repo-type dataset --local-dir pico_human_demo_v1 cd pico_human_demo_v1 && sha256sum -c checksums.sha256 # 核对文件完整(macOS:shasum -a 256 -c checksums.sha256) 目录 README.md 本文件 manifest.csv… See the full description on the dataset page: https://huggingface.co/datasets/SingleBicycle/pico_human_demo_v1.imageroboticsn<1K0 likes58 downloads4d agoHugging Face20roboterradar /humanoid-robot-radarscore Roboterradar Humanoid & Quadruped Robot Dataset Curated editorial assessments of 21 commercially relevant humanoid robots (16) and quadruped robots (5), with a frozen scoring methodology, evidence grades and a complete source register. This Hugging Face repository is a versioned distribution mirror. The canonical, citeable publication is the Zenodo release: Version 1.0.0 DOI: https://doi.org/10.5281/zenodo.21797689 Concept DOI for all versions:… See the full description on the dataset page: https://huggingface.co/datasets/roboterradar/humanoid-robot-radarscore.tabularn<1K0 likes57 downloads2mo agoHugging Face21AdControlCenter /ad-creative-quality-human-vs-llm Human Expert vs LLM Judge: Facebook Ad Creative Quality 500 real Facebook ads from 253 advertisers, each rated for creative quality by a human ad expert AND by a vision LLM — with the LLM's full reasoning. The headline finding baked into this data: the human and the LLM agree on image quality only 26.8% of the time. The LLM judge rates 71.8% of ads "good"; the human expert rates only 20% "good". If you are using an LLM as a judge of ad creative (or any subjective visual quality)… See the full description on the dataset page: https://huggingface.co/datasets/AdControlCenter/ad-creative-quality-human-vs-llm.tabularimage-classificationn<1K2 likes55 downloads2mo agoHugging Face22mohitraiyani27 /Human_Gut_Microbiome_Data Human Gut Microbiome Dataset (40-class) Pre-processed, train/val/test-split human gut microbiome dataset used to train and evaluate MicrobiomeFM, a Transformer-based foundation model for multi-class disease classification from shotgun metagenomic data. This release contains the 40-class filtered version of the dataset that the published results were obtained on (test accuracy ≈ 96.05%, weighted F1 ≈ 0.950, macro F1 ≈ 0.691). It is derived from the curatedMetagenomicData (cMD)… See the full description on the dataset page: https://huggingface.co/datasets/mohitraiyani27/Human_Gut_Microbiome_Data.tabulartabular-classification10K<n<100K1 likes54 downloads3mo agoHugging Face23rhan721 /human-jury-afd4ba human-jury-afd4ba Synthetic sensors test data: 36 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/rhan721/human-jury-afd4ba.tabularn<1K0 likes54 downloads1mo agoHugging Face24airt-ml /twitter-human-botstabular10K<n<100K4 likes50 downloads4y agoHugging Face25jaxaht /lbf-human-datatabularn<1K0 likes50 downloads16d agoHugging Face26openaging /human_methylation_bench_ver1_test Human DNA Methylation Dataset ver1 This dataset is a benchmark dataset for predicting the aging clock, curated from publicly available DNA methylation data. The original benchmark dataset was published by Dmitrii Kriukov et al. (2024) by integrating data from 65 individual studies. To improve usability, we ensured unique sample IDs (excluding duplicate data, GSE118468 and GSE118469) and randomly split the data into training and testing subsets (train : test = 7 : 3) to… See the full description on the dataset page: https://huggingface.co/datasets/openaging/human_methylation_bench_ver1_test.tabular1K<n<10K0 likes49 downloads2y agoHugging Face27wty-yy /humanoid_retargeting_tools_resources Humanoid Retargeting Data The humanoid retargeting data used in humanoid_retargeting_tools. # make sure hf CLI is installed curl -LsSf https://hf.co/cli/install.sh | bash # download the dataset hf download wty-yy/humanoid_retargeting_tools_resources --repo-type=dataset --local-dir resources resources/data/g1/lafan1/: Converted LAFAN1 G1 retargeting dataset in NPZ format. resources/data/g1/kimodo/: Generate data from kimodo_fock. resources/data/g1/dailylife/: DailyLife dataset… See the full description on the dataset page: https://huggingface.co/datasets/wty-yy/humanoid_retargeting_tools_resources.3dn<1K0 likes45 downloads5mo agoHugging Face28Kenneth-Gonzalez /human-direction-63a8a0 human-direction-63a8a0 Synthetic sensors test data: 53 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/Kenneth-Gonzalez/human-direction-63a8a0.tabularn<1K0 likes39 downloads1mo agoHugging Face29bengusu80 /humanoid-object-pushing-dataset-v1Dataset for controlled pushing and object relocation. Description Object weight and applied force signals mapped to pushing behaviors. Task Description Helps humanoid robots move objects safely by adapting push strength and body posture. tabularroboticsn<1K0 likes37 downloads8mo agoHugging Face30lifescore /fda-peptide-human-evidence Seven FDA-reviewed peptides: claims coded for identity, administration, outcome and replication A claim-level comparison of the seven peptide pairs reviewed at the FDA Pharmacy Compounding Advisory Committee meeting of 23-24 July 2026. The data separates molecular identity, administration to people, claimed outcomes and independent replication. Read the evidence-led article: https://lifesco.re/edge/which-peptide-claims-have-actually-been-tested-in-people/ Archived version and… See the full description on the dataset page: https://huggingface.co/datasets/lifescore/fda-peptide-human-evidence.textn<1K0 likes33 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.