Team Ai
Datasetpublic

rsoohyun/SpatialBlock-15k

SpatialBlock-15k This dataset accompanies the paper SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem. It contains 15,000 synthetic block-stacking problems for training large vision-language models (LVLMs) to improve spatial reasoning. The dataset includes three types of multiple-choice questions: Q1: 3D-to-2D projection Q2: viewpoint transformation Q3: structural combination The dataset is organized into a train split of 15,000… See the full description on the dataset page: https://huggingface.co/datasets/rsoohyun/SpatialBlock-15k.

sourceHugging Faceapache-2.0updated 1mo agoView on Hugging Face
7likes859downloads
Dataset Card

SpatialBlock-15k

This dataset accompanies the paper SpatialBlock: Enhancing Spatial Intelligence in LVLMs via Synthetic Block-Stacking Problem. It contains 15,000 synthetic block-stacking problems for training large vision-language models (LVLMs) to improve spatial reasoning. The dataset includes three types of multiple-choice questions:

  • —Q1: 3D-to-2D projection
  • —Q2: viewpoint transformation
  • —Q3: structural combination

The dataset is organized into a train split of 15,000 examples and a test split of 600 examples. Code, training scripts, and model checkpoints are available in the GitHub repository. The dataset is released under the Apache License 2.0.