Team Ai
Datasetpublic

e1879/showui-web-processed

ShowUI-Web Processed Flattened, normalized, and scenario-split version of showlab/ShowUI-web. Each row is a single (instruction, UI element) pair with normalized bounding-box coordinates. Schema Column Type Description sample_id string Unique row identifier ({row}_{element}) screenshot_id string Groups elements from the same screenshot image_relpath string Relative path to the screenshot image scenario string Website/domain inferred from the image… See the full description on the dataset page: https://huggingface.co/datasets/e1879/showui-web-processed.

sourceHugging Facecc-by-4.0updated 7mo agoView on Hugging Face
0likes44downloads
Dataset Card

ShowUI-Web Processed

Flattened, normalized, and scenario-split version of showlab/ShowUI-web.

Each row is a single (instruction, UI element) pair with normalized bounding-box coordinates.

Schema

ColumnTypeDescription
sample_idstringUnique row identifier ({row}_{element})
screenshot_idstringGroups elements from the same screenshot
image_relpathstringRelative path to the screenshot image
scenariostringWebsite/domain inferred from the image path
instructionstringNatural-language grounding instruction
bbox_xyxylist[float]Normalized bounding box [x1, y1, x2, y2] in [0, 1]
point_xylist[float] or nullNormalized click point [x, y]
element_typestring or nullUI element type label

Splits

SplitRowsStrategy
trainmajorityScenario-based holdout
validation~10% scenariosDomain holdout
test~15% scenariosDomain holdout

Repository Layout

The dataset repo contains both row-level parquet artifacts and image files:

  • —flat.parquet — full flattened table (all rows)
  • —splits/train.parquet — train split
  • —splits/val.parquet — validation split
  • —splits/test.parquet — test split
  • —splits/splits.json — split metadata
  • —images/... — screenshot and UI metadata files

Images

Screenshot images are hosted in the images/ directory of this repository. Use image_relpath to construct the path or fetch individual images on demand:

python
from huggingface_hub import hf_hub_download
from PIL import Image

path = hf_hub_download(
    repo_id="e1879/showui-web-processed",
    repo_type="dataset",
    filename=f"images/{row['image_relpath']}",
)
img = Image.open(path)

Usage

Load train split via datasets:

python
from datasets import load_dataset

ds = load_dataset("e1879/showui-web-processed")
print(ds["train"][0]["instruction"])

Or load parquet artifacts directly from the dataset repo:

python
import pandas as pd
from huggingface_hub import hf_hub_download

train_path = hf_hub_download(
    repo_id="e1879/showui-web-processed",
    repo_type="dataset",
    filename="splits/train.parquet",
)
train_df = pd.read_parquet(train_path)
print(train_df.shape)

Credit

Source dataset: showlab/ShowUI-web

e1879/showui-web-processed · Team Ai