datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
semantic-segmentation-test-sampleThis dataset contains 10 examples of the segments/sidewalk-semantic dataset (i.e. 10 images with corresponding ground-truth segmentation maps).
Semantic-SVG-Benchmark
Semantic SVG Benchmark
A benchmark of 203 SVG files annotated with human-written semantic object-decomposition trees:
every rendered shape (<path>, <rect>, <circle>, …) in each SVG is assigned to a named semantic
object (e.g. judge, gavel), and objects may be further decomposed into parts
(e.g. Bamboo planter → pot, bamboo). It is the evaluation benchmark of
Compositional SVG Generation via VLM-Driven Hierarchical Semantic Parsing
(EMNLP 2026). The annotations are ours; the SVGs… See the full description on the dataset page: https://huggingface.co/datasets/KU-MIIL/Semantic-SVG-Benchmark.SemanticSTF
📌 SemanticSTF Dataset
SemanticSTF is a real multimodal LiDAR dataset collected under adverse weather conditions including rain, snow, and fog, for autonomous driving research.
It provides synchronized LiDAR point clouds, RGB images, and per-point semantic labels of 20 classes, designed for 3D semantic segmentation and sensor fusion tasks.
The dataset contains train/val/test splits, camera intrinsics/extrinsics, and high-quality annotations aligned at the frame level.… See the full description on the dataset page: https://huggingface.co/datasets/AR-X/SemanticSTF.semanticSearchVLM_semantics_SLO_benchmark
VLM Semantics SLO Benchmark
VLM Semantics SLO is a Slovenian multimodal benchmark for studying cultural and semiotic reasoning in vision-language models. It goes beyond object recognition by asking models to interpret visual hierarchy, spatial relations, colour and mood, composition, cultural symbols, metaphor, denotation and connotation, intertextuality, communicative intent, and relevance to Slovenia.
The released JSON contains 4,950 image-level records. Every record has ten… See the full description on the dataset page: https://huggingface.co/datasets/maticmatusek/VLM_semantics_SLO_benchmark.Semantic_Segmantation_Datasetssemantic_seg_ATL
Dataset Card for "semantic_seg_ATL"
More Information needed
PhenoBench_images_semanticsSemantic-Segmentation-Aerial-Imagery-DatasetThe dataset comprises aerial imagery of Dubai acquired by MBRSC satellites and annotated with pixel-level semantic segmentation across 6 distinct classes. The dataset comprises a total of 72 images, which are organised into 6 larger tiles. The categories are as follows:
Credit: Humans in the Loop is releasing an openly accessible dataset that has been annotated for a collaborative project with the Mohammed Bin Rashid Space Centre in Dubai, United Arab Emirates.
Deep Learning Projects for Final… See the full description on the dataset page: https://huggingface.co/datasets/gymprathap/Semantic-Segmentation-Aerial-Imagery-Dataset.semantic_seg_atl_resized
Dataset Card for "semantic_seg_atl_resized"
More Information needed
semanticsegmentationandposeestimationfromrgbdRGB-D dataset for instance segmentation (from RGB or depth) and pose estimation of individual objects. Data has been generated by randomizing bin contents in Webots.
Each instance contains a mask image as well meta data containing labels, position, and size of each object.
You can create your own data by opening webots_grasp.wbt in the world directory using Webots.
Semantic_Segmentation_CE
Dataset Card for "Semantic_Segmentation_CE"
More Information needed
semantic_segmentation_footpath_grass_road_water_v1
My Segmentation Dataset
Description
This dataset contains outdoor scene images and segmentation annotations.
Labels
0: background
1: footpath
2: grass
3: road
4: water
Structure
train/
validation/
annotations/
Annotation format
COCO segmentation annotations / semantic mask PNGs
Intended use
Training and evaluation for semantic segmentation models such as SegFormer.
pidray-semanticsPhenoBench_images_semantics02semantic-segmentation-dataset
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/dylangirrens/semantic-segmentation-dataset.semantic-segmentationSemanticSegmentationSemantic_Segmentation_Graptolodiea
