datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
GenPoster100K
Dataset Card for GenPoster100K
Dataset Summary
GenPoster-100K is a large-scale dataset for content-aware graphic layout generation introduced in the SEGA paper.
The paper describes it as a high-quality poster dataset with layer-parseable source materials and rich metadata.
This repository provides a Hugging Face datasets loader implementation that reads the source release (BruceW91/GenPoster-100K) and exposes normalized examples with:
poster background image… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/GenPoster100K.PubLayNet
Dataset Card for PubLayNet
Dataset Summary
PubLayNet is a large document layout analysis dataset built by automatically matching XML representations and PDF content from more than one million PubMed Central Open Access articles. It contains more than 360,000 document images with COCO-style annotations for common layout elements such as text, title, list, table, and figure regions.
Supported Tasks and Leaderboards
The dataset supports document… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/PubLayNet.GQA-Scene-Graph
Dataset Card for GQA-35k
The GQA (Visual Reasoning in the Real World) dataset is a large-scale visual question answering dataset that includes scene graph annotations for each image.
This is a FiftyOne dataset with 35000 samples.
Note: This is a 35,000 sample subset which does not contain questions, only the scene graph annotations as detection-level attributes.
You can find the recipe notebook for creating the dataset here
Installation
If you haven't already… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/GQA-Scene-Graph.GVC-Graphsgraphical-bootstrap-correlator-dataset
Graphical Bootstrap Correlator Dataset
This dataset contains large-scale graph-structured data arising from high-order perturbative computations of four-point correlators in planar $\mathcal{N}=4$ super Yang--Mills theory.
The data consists of denominator graphs (d-graphs) appearing in the graphical bootstrap formulation of correlators. Each graph is associated with a binary label indicating whether it contributes to the correlator at a given perturbative order.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/Gabriele-dian/graphical-bootstrap-correlator-dataset.PKU-PosterLayout
Dataset Card for PKU-PosterLayout
Dataset Summary
PKU-PosterLayout is a content-aware visual-textual poster layout benchmark released with PosterLayout: A New Benchmark and Approach for Content-aware Visual-Textual Presentation Layout. The paper defines the task as arranging predefined text, logo, and underlay elements on a non-empty poster canvas while considering both inter-element and inter-layer relationships. The original benchmark contains 9,974… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/PKU-PosterLayout.CreativePSD
Dataset Card for CreativePSD
Dataset Summary
CreativePSD is the PSD-derived graphic design dataset released with PSDesigner. Each example is a poster archive containing PSD tree text, structured layer metadata, tool-call trajectories, source image resources, and stepwise rendered images.
This loader keeps the contents of each poster_*.zip archive: all metadata text/JSON files, all raw_resource images, all rendering_imgs images, and a manifest of every member in… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/CreativePSD.graphmemix-benchmarks
GraphMemix Benchmarks
Unified multimodal memory benchmark bundles used by
GraphMemix
(arXiv:2608.26983) — four long-term
personalized memory benchmarks with their raw media assets, packaged together
for reproducible evaluation.
Benchmark
Questions
Memories
Track
Upstream license
ATM-Bench (default + hard)
1,044
11,034
memory QA over one multimodal archive
MIT
Mem-Gallery
1,711
7,944
multimodal gallery memory QA
MIT
MemEye
1,855
3,392
comics-derived memory QA… See the full description on the dataset page: https://huggingface.co/datasets/oking0197/graphmemix-benchmarks.graphene-design-universe-256k
Graphene Design Universe 256K
256,000 unrelaxed atomistic graphene designs, with images, full periodic cells, standard extended XYZ coordinates, reproducible recipes, geometry-quality flags and a geometry-similarity explorer.
Interactive 256K explorer · Preserved 64K release · 4K movie
This expansion preserves all 64,000 prior designs, IDs and coordinate-file bytes and adds 192,000 new designs. It broadens the original 16 groups and adds eight hybrid motif groups. The… See the full description on the dataset page: https://huggingface.co/datasets/lamm-mit/graphene-design-universe-256k.Rico
Dataset Card for Rico
Dataset Summary
Rico is a mobile app UI dataset for building data-driven design applications. The original dataset mines Android apps at runtime and exposes visual, textual, structural, and interactive design properties from more than 9.3k apps across 27 categories and more than 66k unique UI screens. This packaging provides metadata, screenshots, view hierarchies, and semantic annotations as separate configs.
Supported Tasks… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/Rico.Graph-AlgorithmsGraphRL_2room_50-99_0310GBC10M
Graph-based captioning (GBC) is a new image annotation paradigm that combines the strengths of long captions, region captions, and scene graphs
GBC interconnects region captions to create a unified description akin to a long caption, while also providing structural information similar to scene graphs.
** The associated data point can be found at demo/water_tower.json
Description and data format
The GBC10M dataset, derived from the original images in CC12M, is… See the full description on the dataset page: https://huggingface.co/datasets/graph-based-captions/GBC10M.genvsr-video-benchmarksPrismLayersPro
PrismLayers: Open Data for High-Quality Multi-Layer Transparent Image Generative Models
We introduce PrismLayersPro, a 20K high-quality multi-layer transparent image dataset with rewritten style captions and human filtering.
PrismLayersPro is curated from our 200K dataset, PrismLayers, generated via MultiLayerFLUX.
Dataset Structure
📑 Dataset Splits (by Style)
The PrismLayersPro dataset is divided into 21 splits based on visual style categories.Each… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/PrismLayersPro.mm-graph
Multimodal Graph Benchmark
Paper: https://huggingface.co/papers/2406.16321
Project Page: https://mm-graph-benchmark.github.io/
Code: https://github.com/mm-graph-benchmark/mm-graph-benchmark
This repo contains all the datasets used in "Multimodal Graph Benchmark".
THE DATASET IS PROVIDED "AS IS", WITHOUT WARRANTY OF ANY KIND, EXPRESS OR
IMPLIED, INCLUDING BUT NOT LIMITED TO THE WARRANTIES OF MERCHANTABILITY,
FITNESS FOR A PARTICULAR PURPOSE AND NONINFRINGEMENT. IN NO EVENT SHALL THE… See the full description on the dataset page: https://huggingface.co/datasets/mm-graph-org/mm-graph.GraphicReasoning400CGL-Dataset-v2
Dataset Card for CGL-Dataset v2
Dataset Summary
CGL-Dataset v2 is an advertising-poster layout dataset released with Relation-Aware Diffusion Model for Controllable Poster Layout Generation. The paper argues that poster layouts should account for both visual-textual relationships and geometry relationships between elements. This version extends CGL-Dataset with richer element annotations, text annotations, and text features for controllable poster layout… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/CGL-Dataset-v2.CGL-Dataset
Dataset Card for CGL-Dataset
Dataset Summary
CGL-Dataset is a poster layout dataset released with Composition-aware Graphic Layout GAN for Visual-Textual Presentation Designs. The paper studies layout generation for a given image, emphasizing that both global semantics and spatial image composition affect where graphic elements should be placed. The original dataset contains 60,548 advertising posters with annotated layout information.
Supported… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/CGL-Dataset.Graph200K
VisualCloze: A Universal Image Generation Framework via Visual In-Context Learning
[Paper] [Project Page] [Github]
[🤗 Online Demo]
[🤗 Full Model Card (Diffusers)] [🤗 LoRA Model Card (Diffusers)]
Graph200k is a large-scale dataset containing a wide range of distinct tasks of image generation. If you find Graph200k is helpful, please consider to star ⭐ the Github Repo. Thanks!
📰 News
[2025-5-15] 🤗🤗🤗 VisualCloze has been merged into the… See the full description on the dataset page: https://huggingface.co/datasets/VisualCloze/Graph200K.Magazine
Dataset Card for Magazine
Dataset Summary
Magazine is a magazine layout dataset released with Content-aware Generative Modeling of Graphic Design Layouts. The paper studies graphic layout generation conditioned on visual and textual content and introduces a large-scale magazine layout dataset with fine-grained layout annotations and keyword labels.
Supported Tasks and Leaderboards
The dataset supports content-aware layout generation, graphic… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/Magazine.graphical-bootstrap-correlator-dataset
Graphical Bootstrap Correlator Dataset
This dataset contains large-scale graph-structured data arising from high-order perturbative computations of four-point correlators in planar $\mathcal{N}=4$ super Yang--Mills theory.
The data consists of denominator graphs (d-graphs) appearing in the graphical bootstrap formulation of correlators. Each graph is associated with a binary label indicating whether it contributes to the correlator at a given perturbative order.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/anonymous314/graphical-bootstrap-correlator-dataset.edge-ML-metabolomics-graph
edge_ML metabolomics co-response graph
An undirected graph of 18,494 nodes and 2,709,209 edges built from pairwise
metabolite co-response statistics across 83 MetaboLights
studies, together with the node properties, the PyTorch Geometric graph object, and the
full pipeline that produces them.
A node is one differential comparison within one study assay (MTBLS1405_0002_00003332
= study MTBLS1405, assay 002, feature 00003332). An edge carries the association
between two… See the full description on the dataset page: https://huggingface.co/datasets/kozo2/edge-ML-metabolomics-graph.graph-nucls
Graph-NuCLS: A Cell-Graph Dataset for Nucleus Classification from NuCLS
Graph-NuCLS is a node-level classification dataset derived from the NuCLS dataset with "main" labels. Each tissue patch is converted into a cell-graph where nodes represent detected cell nuclei and edges encode spatial proximity. The task is predicting the cell type of each nucleus across 7 classes. Note that node features describe cell morphology, texture, and color intensity whereas edge features are… See the full description on the dataset page: https://huggingface.co/datasets/ogutsevda/graph-nucls.Desigen
Dataset Card for Desigen
Dataset Summary
Desigen contains web advertisement design data with background images, content prompts, layout element annotations, and design canvas sizes. This loader reads the parquet shards hosted at creative-graphic-design/Desigen.
Dataset Structure
Data Fields
image: Background or rendered advertisement image.
prompt: Text prompt associated with the background image.
region: Image region boxes.… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/Desigen.GraphVQA-newtop-papers-graph-experts-datagraphene-design-universe-64k
Graphene Design Universe 64K
64,000 atomistic graphene geometries, with images, periodic cells, generator parameters, standard extended XYZ coordinates and a geometry-similarity embedding.
Explore the interactive map.
This is a geometry-only design library derived from the graphene metamaterial generators developed in an AI-built atomistic modeling workflow. It preserves the preceding 16,384-design visualization atlas and adds 47,616 newly generated designs. These 64,000… See the full description on the dataset page: https://huggingface.co/datasets/lamm-mit/graphene-design-universe-64k.graphene-agent-data
Models building models for discovery of graphene metamaterial design principles in the context of failure
Companion data for the GitHub repository lamm-mit/graphene-agent and the manuscript Models building models for discovery of graphene
metamaterial design principles in the context of failure. The repository holds the code, the experiment database (all 132 records
with structures and stress–strain curves), the report, the figures and the app; this dataset holds the files that… See the full description on the dataset page: https://huggingface.co/datasets/lamm-mit/graphene-agent-data.GBC1M
Graph-based captioning (GBC) is a new image annotation paradigm that combines the strengths of long captions, region captions, and scene graphs
GBC interconnects region captions to create a unified description akin to a long caption, while also providing structural information similar to scene graphs.
** The associated data point can be found at demo/water_tower.json
Description and data format
The GBC1M dataset, derived from the original images in CC12M, is… See the full description on the dataset page: https://huggingface.co/datasets/graph-based-captions/GBC1M.
