datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
visualizationWildDet3D-visualization-source
WildDet3D Visualization Data
This repository hosts the visualization data for the WildDet3D-Bench benchmark — a human-annotated evaluation set for monocular 3D object detection in the wild.
Dataset Overview
WildDet3D-Bench is a validation set of 2,470 images drawn from three source datasets, with 9,256 human-verified 3D bounding box annotations across 2,196 images.
Source
Images
Description
COCO Val
424
MS-COCO 2017 validation
LVIS Train
1,113
LVIS v1.0 (COCO… See the full description on the dataset page: https://huggingface.co/datasets/allenai/WildDet3D-visualization-source.itw_pipeline_visualization
ITW Pipeline Visualization — assets
Purpose of this repository
This repository exists for one reason: to serve media files to the
visualization site at
https://silicon23.github.io/itw_pipeline_visualization/.
A static site cannot host its own heavy media, so the frames, depth maps and
overlay renders it streams live here.
It is not published for redistribution, and it is not a dataset to train or
evaluate on. It is the asset backing of a figure — the equivalent… See the full description on the dataset page: https://huggingface.co/datasets/Silicon23/itw_pipeline_visualization.factorio-blueprint-visualizations
Dataset Description
This dataset is a collection of visualizations of Factorio Blueprints using this Factorio Visualization Tool: https://github.com/piebro/factorio-blueprint-visualizer. The Blueprints are collected from https://www.factorio.school/.
Examples
Dataset Structure
"svg_original": The svg downloaded like this from the website
"svg_rect": The svg reshaped to a rect and a slightly bigger border
"png_1024x1024": The svg_rect images exported as… See the full description on the dataset page: https://huggingface.co/datasets/piebro/factorio-blueprint-visualizations.SOC-Training-Data-Visualization
Paper Link
SOS: Synthetic Object Segments Improve Detection, Segmentation, and Grounding
Code repo
Code for Generation
Citation
@misc{huang2025sossyntheticobjectsegments,
title={SOS: Synthetic Object Segments Improve Detection, Segmentation, and Grounding},
author={Weikai Huang and Jieyu Zhang and Taoyang Jia and Chenhao Zheng and Ziqi Gao and Jae Sung Park and Ranjay Krishna},
year={2025},
eprint={2510.09110},
archivePrefix={arXiv}… See the full description on the dataset page: https://huggingface.co/datasets/weikaih/SOC-Training-Data-Visualization.Spatial-Visualization-Benchmark
Spatial Visualization Benchmark
This repository contains the Spatial Visualization Benchmark. The evaluation code is released on: wangst0181/Spatial-Visualization-Benchmark.
Dataset Description
The SpatialViz-Bench aims to evaluate the spatial visualization capabilities of multimodal large language models, which is a key component of spatial abilities. Targeting 4 sub-abilities of Spatial Visualization, including mental rotation, mental folding, visual penetration, and… See the full description on the dataset page: https://huggingface.co/datasets/PLM-Team/Spatial-Visualization-Benchmark.Celldega_Visualization_BenchmarkingNeural_3D_Reconstruction_and_Immersive_VR_Visualization_of_Row_Crops
🌱 BBCH-NeRF Plant Dataset
This dataset accompanies the paper:
Neural 3D reconstruction and immersive VR visualization of row crops across phenological growth stages
Overview
This dataset contains data for 3D plant reconstruction using NeRF and Gaussian Splatting (Gsplat) across multiple crops and BBCH growth stages.
It includes:
Raw video data
NeRF reconstructions (point clouds)
Gaussian Splatting outputs
Dataset Structure
Dataset/
├──… See the full description on the dataset page: https://huggingface.co/datasets/ShambhaviJoshi/Neural_3D_Reconstruction_and_Immersive_VR_Visualization_of_Row_Crops.figure4-visualization-dataset
Ordinal-Reference Visualization Dataset
Created: 2026-07-24 UTCStatus: final frozen 50-case package
This package is a hand-curated qualitative dataset for horizontal, same-case image-editing comparisons. Each image contains repeated instances of one subject class arranged in one or two visually legible horizontal rows. The cases deliberately span food, plants, household and laboratory objects, toys, vehicles, wildlife, and people so that ordinal grounding is tested across more… See the full description on the dataset page: https://huggingface.co/datasets/XIAOHAOYANG/figure4-visualization-dataset.STEP-visualization-subsetSR-Interpolation-Visualization
SR Interpolation Visualization Assets
Curated image pairs used by the continuous LR-to-HR interpolation page on nitec427.github.io.
realsr/: 12 matched Canon/Nikon samples across LR, HR, and three experiment-output folders.
div2k/: a fixed random subset of 30 matched LR/HR patches from the local DIV2K visualization set.
These files are published as visualization assets rather than a complete redistribution of either dataset.
visualization-chartslrmzero_visualizationVisualizationstereoadapter-2_visualizationvisualizationsdroid_pipeline_visualization
DROID 3D Bounding-Box Pipeline — Visualization Data
Per-frame fused 3-camera pointclouds backing the web viewer at
https://silicon23.github.io/droid_pipeline_visualization/.
Each episode comes from the DROID v1.0.1 dataset and completed the full
detection pipeline (FoundationStereo depth -> optimized extrinsics -> SAM3
masks -> SAM3D box -> FoundationPose video tracking). Point positions are in
the Franka base (world) frame, metres, Z up.
Layout… See the full description on the dataset page: https://huggingface.co/datasets/Silicon23/droid_pipeline_visualization.factorio-blueprint-visualizations-sdxl-lora-examples
Factorio Blueprint Visualizations SDXL Lora Examples
Examples of the usage of https://huggingface.co/piebro/factorio-blueprint-visualizations-sdxl-lora. The images are generated using 25 inference steps and a guidance_scale of 7. The filenames are composed like this: {counter}_{seed}_{prompt}.png.
SOC-Object-Segments-Visualization
Paper Link
SOS: Synthetic Object Segments Improve Detection, Segmentation, and Grounding
Code repo
Code for Generation
Citation
@misc{huang2025sossyntheticobjectsegments,
title={SOS: Synthetic Object Segments Improve Detection, Segmentation, and Grounding},
author={Weikai Huang and Jieyu Zhang and Taoyang Jia and Chenhao Zheng and Ziqi Gao and Jae Sung Park and Ranjay Krishna},
year={2025},
eprint={2510.09110},
archivePrefix={arXiv}… See the full description on the dataset page: https://huggingface.co/datasets/weikaih/SOC-Object-Segments-Visualization.WildChat-Legal-Visualization-V2Results-For-Music-Visualization-Generation-Pipelinefolding_250_high_loss_visualizationBEHAVIOR-1K-VISUALIZATION-DEMOholohub_bci_visualizationvisualization_test2your-output-visualizations-repodebug-visualizationThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 150,
"total_tasks": 1,
"total_videos": 1,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ryukin777/debug-visualization.debug-visualization-movingThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 120,
"total_tasks": 1,
"total_videos": 1,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ryukin777/debug-visualization-moving.Visualizationvisualization
