datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
assetsliberoThis dataset was created using LeRobot.
Dataset Description
This dataset combines four individual Libero datasets: Libero-Spatial, Libero-Object, Libero-Goal and Libero-10.
All datasets were taken from here and converted into LeRobot format.
Homepage: https://libero-project.github.io
Paper: https://arxiv.org/abs/2306.03310
License: CC-BY 4.0
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "panda",
"total_episodes": 1693… See the full description on the dataset page: https://huggingface.co/datasets/physical-intelligence/libero.Waste-Dumpsites-DroneImagery
Dataset for Waste/Dumpsite Detection using drone imagery
Contains 2115 drone images of illegal waste dumpsites
1280 x 1280 px resolution
Nadir perspective (camera pointing straight down at a 90-degree angle to the ground)
Annotations and Images
train | valid | test
actual images
COCO - annotations_coco.json files in each split directory
.parquet files in data directory with embeded images
The dataset was collected as part of the [ Raven Scan ] project, more… See the full description on the dataset page: https://huggingface.co/datasets/INS-IntelligentNetworkSolutions/Waste-Dumpsites-DroneImagery.Hausa
Hausa Ajami OCR Dataset
Ce dataset contient des paires image/transcription de manuscrits haoussa en écriture ajami (écriture arabe adaptée au haoussa).
Contenu
Chaque ligne du fichier data/train/metadata.jsonl correspond à une ligne de texte ajami segmentée, avec :
file_name : nom du fichier image correspondant (image de la ligne, recadrée)
transcript : translittération en écriture latine de la ligne
source : identifiant du manuscrit d'origine (voir tableau… See the full description on the dataset page: https://huggingface.co/datasets/IntelligenceResearchLab/Hausa.pd12m
PD12M
This is a curated PD12M dataset for use with the II-Commons project.
Dataset Details
Dataset Description
This dataset comprises a curated Public Domain 12M image collection, refined by filtering for active image links. EXIF data was extracted, and images underwent preprocessing and feature extraction using SigLIP 2. All vector embeddings are normalized 16-bit half-precision vectors optimized for L2 indexing with vectorchord.… See the full description on the dataset page: https://huggingface.co/datasets/Intelligent-Internet/pd12m.aloha_pen_uncap_diverseThis dataset was created using LeRobot.
Dataset Description
This dataset is a lerobot conversion of the aloha_pen_uncap_diverse subset of BiPlay.
BiPlay contains 9.7 hours of bimanual data collected with an aloha robot at the RAIL lab @ UC Berkeley, USA. It contains 7023 clips, 2000 language annotations and 326 unique scenes.
Paper: https://huggingface.co/papers/2410.10088 Code: https://github.com/sudeepdasari/dit-policy If you use the dataset please cite:… See the full description on the dataset page: https://huggingface.co/datasets/physical-intelligence/aloha_pen_uncap_diverse.syntheticDocQA_artificial_intelligence_test_beirBEIR version of vidore/syntheticDocQA_artificial_intelligence_test.
intel-image-classification
Intel Image Classification
The Intel Image Classification dataset contains images of natural scenes categorized into six classes:
Buildings
Forest
Glacier
Mountain
Sea
Street
📆 Content
The dataset contains ~25,000 images of size 150x150 pixels.
Images are evenly distributed across 6 categories:
{'buildings' -> 0,
'forest' -> 1,
'glacier' -> 2,
'mountain' -> 3,
'sea' -> 4,
'street' -> 5 }
It is divided into three parts:
Training set: ~14… See the full description on the dataset page: https://huggingface.co/datasets/sfarrukhm/intel-image-classification.SocialCounterfactualssyntheticDocQA_artificial_intelligence_test_beirBEIR version of vidore/syntheticDocQA_artificial_intelligence_test.
syntheticDocQA_artificial_intelligence_test
Dataset Description
This dataset is part of a topic-specific retrieval benchmark spanning multiple domains, which evaluates retrieval in more realistic industrial applications.
It includes documents about the Artificial Intelligence.
Data Collection
Thanks to a crawler (see below), we collected 1,000 PDFs from the Internet with the query ('artificial intelligence'). From these documents, we randomly sampled 1000 pages.
We associated these with 100 questions and answers… See the full description on the dataset page: https://huggingface.co/datasets/vidore/syntheticDocQA_artificial_intelligence_test.synthetic-lapse-surrender-study-2023-2025
Synthetic Term & Whole Life Lapse and Surrender Study, 2023–2025
10,000 synthetic policyholders · 27,692 VM-51 policy-year records · 25,776 policy-years of exposure · ten markets
This is synthetic data. Every record was generated, not extracted. It is a documented,
reproducible generative prior about policyholder behaviour, not experience. It carries no
information about any real portfolio or person, and no valuation, pricing or reserving
assumption should be set from it.… See the full description on the dataset page: https://huggingface.co/datasets/intelligentactuaries/synthetic-lapse-surrender-study-2023-2025.blogThis is for images in HuggingFace blogs. Please do not delete it!
Emotion.Intelligencedocqa_artificial_intelligence_beirThis is a copy of https://huggingface.co/datasets/jinaai/docqa_artificial_intelligence reformatted into the BEIR format. For any further information like license, please refer to the original dataset.
Disclaimer
This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/docqa_artificial_intelligence_beir.ocr-bn-datagen-v1INTEL-TAUeie-earth-intelligence-engine
Introduction
This dataset contains eight subdatasets to study segmentation-guided image-to-image (im2im) translation in Earth Observation. The subdatasets are divided into flood, reforestation, and Arctic sea ice melt events. As displayed in Fig. 1, the subdatasets contain image triplets with two HD (1024x1024 px) satellite images that were taken before and after the event, such as flooding, occured. The two images are paired with a segmentation-mask that outlines the area where… See the full description on the dataset page: https://huggingface.co/datasets/blutjens/eie-earth-intelligence-engine.ComfyAgent-David
ComfyAgent-David
Final experiment image artifacts grouped by method.
comfyclaw/: final ComfyClaw runs from comfy_agent_experiments_output
comfygems/: final ComfyGEMS runs from cemfygems_comparison_output
baseline/: final baseline runs from comfy_agent_baseline_experiments_output
Included files are limited to image outputs under results/images/ and detailed/.
Smoke tests, caches, and earlier trial runs are excluded.
SK-VQA
Dataset Card for SQ-VQA
Dataset Summary
SK-VQA is a large-scale synthetic multimodal dataset containing over 2 million visual question-answer pairs, each paired with context documents that contain the information needed to answer the questions.
The dataset is designed to address the critical need for training and evaluating multimodal LLMs (MLLMs) in context-augmented generation settings, particularly for retrieval-augmented generation (RAG) systems. It enables training… See the full description on the dataset page: https://huggingface.co/datasets/Intel/SK-VQA.piper_uncap_penThis dataset was created using LeRobot.
About Task
Task Objective: Pick up the pen from the desk and uncap the pen.
Operational Objects: Marker Pen
Operation Duration: Each operation takes approximately 15 to 20 seconds.
Recording Frequency: 15 Hz.
Robot Type: 7-DOF dual-arm Agilex 2 Pipers Desktop Robot.
End Effector: Gripper.
Dual-Arm Operation: Yes.
Image Resolution: 640x480.
Camera Positions: High; Low; Left; Right.
Data Content: • Robot's current state. •… See the full description on the dataset page: https://huggingface.co/datasets/io-intelligence/piper_uncap_pen.Intel_Robotic_Welding_Multimodal_Dataset
Dataset Card for the Intel Robotic Welding Multimodal Dataset
This dataset was collected to enable multimodal welding defect detection research. The dataset contains over 4000 annotated samples and was collected in an automotive production floor setting in collaboration with a supplier with access to such facilities. Each sample contains a video, associated audio, a time-series from welding sensors, and five post-weld images for a particular weld. A separately licensed… See the full description on the dataset page: https://huggingface.co/datasets/IntelLabs/Intel_Robotic_Welding_Multimodal_Dataset.objaverse.data.intelligence
Objaverse Data Intelligence: Scene-Level Structural Analysis and Decomposition Metadata
Large-scale scene-level geometry, structure, and material analysis of ~680K Objaverse assets, designed for ML-ready filtering, structural decomposition, and dataset curation.
🎥 Watch Reel
Quick Navigation
Overview
Key Features
Dataset Structure
Additional Files
Dataset Statistics
Classification & Detection Quality
Attribution
Citation
License
Acknowledgements… See the full description on the dataset page: https://huggingface.co/datasets/milimlee-synth3d/objaverse.data.intelligence.see-world-1-CGDsee-world-1-TVCIntelligentVBenchfunctional-grasp-demos
Selected Grasp Demos
Contents
Each case folder contains:
grasp_data.npz — Top-1 ranked grasp (pre_grasp_dofs, grasp_target_dofs, reward, z_lift)
image_grasp.png — AI-generated grasp image (input to perception pipeline)
debug_retarget.png — WujiHand retarget visualization (if available)
Cases
Case
Object
Scale
Reward
z_lift
paper_coffee_cup_rot090__grasp_06
Paper Coffee Cup
1.0
0.742
0.196
paper_cup_8_oz_rot000__grasp_06
Paper Cup 8 Oz
1.0
0.696… See the full description on the dataset page: https://huggingface.co/datasets/Genesis-Intelligence/functional-grasp-demos.Intel-Image-Classificationsee-world-1-LACONVphysical-intelligencev
