semantic
Datasets
All datasets matching “semantic”semantic-vad-eot
Semantic-VAD EOT
End-of-turn (semantic VAD) turns, schema-compatible with
livekit/eot-bench-data.
Each row is one user turn: an audio clip (16 kHz mp3), its words, and ordered silence_spans. Per the
eot-bench convention the last silence span is the true end-of-turn (eot); earlier spans are mid-turn hold
pauses (labels positional, not stored). Timestamps are on the decoded-audio axis, and the decoded length equals
duration.
from datasets import load_dataset, Audio
ds =… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/semantic-vad-eot.SemanticKITTIsemantic-image-search-assetsFUSU-Fine_grained_Urban_Semantic_Understanding
About:
FUSU dataset covers 5 whole urban areas, 847 km^2 located in the north and south of China, with 17 land use and land cover (LULC) classes and over 170K images and 30 billion pixels of annotations, supporting segmentation, change detection and domain adaptation tasks. This data comprises 2 parts:
Bi-temporal high-resolution satellite RGB images with fine-grained annotations.
Monthly revisited Sentinel-2 and Sentinel-1 images.
Details:
1.… See the full description on the dataset page: https://huggingface.co/datasets/sp-juni/FUSU-Fine_grained_Urban_Semantic_Understanding.semantic_patch_cache
HeatTok Semantic Patch Cache
Precomputed .pt caches for HeatTok. Use with HEATTOK_SEMANTIC_CACHE_DIR=/path/to/cache.
Filename pattern: {image_hash}_g1_s28.pt or {image_hash}_g1_gd1_s28.pt
semantic_patch_cache_vrsbench
Dataset: VRSBench (512×512)
Caches do not store precomputed global tokens or patch orientations.
Gaussian parameters and patch metadata are stored.
No need to regenerate .pt files — HeatTok computes global tokens and orientations online at load… See the full description on the dataset page: https://huggingface.co/datasets/Yingying11/semantic_patch_cache.BDD100K-Semantic-SegmentationBDD100K: Semantic Segmentation Dataset
Unofficial redistribution of BDD100K's semantic segmentation task (10K-image subset), under BDD100K's own data license (non-commercial/educational/research redistribution, with notice, explicitly permitted).
Disclaimer
This repository is not an official release of BDD100K.
BDD100K was created by Fisher Yu, Haofeng Chen, Xin Wang, Wenqi Xian, Yingying Chen, Fangchen Liu, Vashisht Madhavan, and Trevor Darrell at UC Berkeley… See the full description on the dataset page: https://huggingface.co/datasets/dronefreak/BDD100K-Semantic-Segmentation.
