datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
posterPKU-PosterLayout
Dataset Card for PKU-PosterLayout
Dataset Summary
PKU-PosterLayout is a content-aware visual-textual poster layout benchmark released with PosterLayout: A New Benchmark and Approach for Content-aware Visual-Textual Presentation Layout. The paper defines the task as arranging predefined text, logo, and underlay elements on a non-empty poster canvas while considering both inter-element and inter-layer relationships. The original benchmark contains 9,974… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/PKU-PosterLayout.Text-Render-2M
Text Render 2M Dataset
A large-scale dataset containing 2 million text rendering image-text pairs for training generative models to improve text rendering performance.
Dataset Structure
image: Rendered text image in PNG format
text: Corresponding text content
file_name: Original filename
folder_id: Folder identifier
Usage
This dataset is designed for fine-tuning generative models to improve text rendering capabilities.
from datasets import load_dataset… See the full description on the dataset page: https://huggingface.co/datasets/PosterCraft/Text-Render-2M.imdb-posters-and-description-512Poster_Music_festivalPoster100K
Poster100K Dataset
A comprehensive dataset containing 93K+ movie and TV show posters with detailed captions and text region annotations for multimodal learning and poster generation tasks.
Dataset Structure
image: Poster image in JPG/JPEG/PNG format
caption: Detailed textual description generated by Gemini-2.5-flash-preview-04-17
mask_regions: Text region coordinates (bounding boxes) in JSON format
file_name: Original filename
folder_path: Normalized relative folder path… See the full description on the dataset page: https://huggingface.co/datasets/PosterCraft/Poster100K.codex-poster-layer-baseline-48
Codex poster layer baseline
Public source marketing posters, GPT-planned layer inventories, first image-tool
outputs, GPT-generated opacity masks, and derived RGBA layers. The paired Space
provides an English interactive inspector. These are model predictions, not
ground-truth segmentation or original design assets.
Planning used GPT-6 Astra through codex exec. Image generation also ran through
Codex's built-in image tool, which does not expose its exact backend model,
snapshot… See the full description on the dataset page: https://huggingface.co/datasets/Elfsong/codex-poster-layer-baseline-48.movie_posters-100k
Dataset Card for "movie_posters-100k"
More Information needed
movie_posters-genres-80k-transformed
Dataset Card for "movie_posters-genres-80k-transformed"
More Information needed
PosterIQ
PosterIQ-Benchmark
[CVPR2026] PosterIQ: A Design Perspective Benchmark for Poster Understanding and Generation[arXiv] https://arxiv.org/abs/2603.24078[HuggingFace] https://huggingface.co/datasets/ArtmeScienceLab/PosterIQ[Github] https://github.com/ArtmeScienceLab/PosterIQ-Benchmark
Citation
@inproceedings{cvpr2026posteriq,
title={PosterIQ: A Design Perspective Benchmark for Poster Understanding and Generation},
author={Feng, Yuheng and Zhang, Wen and Duan… See the full description on the dataset page: https://huggingface.co/datasets/ArtmeScienceLab/PosterIQ.movie_posters-genres-80k-torchvision-transforms
Dataset Card for "movie_posters-genres-80k-torchvision-transforms"
More Information needed
speechocean762-librispeech-posteriors
SpeechOcean762 LibriSpeech Acoustic Posterior Exports
This repository stores frame-level acoustic-model output exports for the
SpeechOcean762 train and test sets, computed with LibriSpeech-trained acoustic
models and Charsiu models.
Files
test/<model_id>/posteriors.h5
contains one model's padded tensor for the test split.
train/<model_id>/posteriors.h5
contains one model's padded tensor for the train split.
manifests/train.json and manifests/test.json
are the… See the full description on the dataset page: https://huggingface.co/datasets/Haopeng/speechocean762-librispeech-posteriors.movie-postersmovie-posters
Dataset Card for "movie_posters"
More Information needed
jerusalem-poster-detection
Jerusalem street poster sightings — detection dataset
Geotagged street photographs of a messianic postering campaign in central
Jerusalem, with bounding boxes around every campaign poster in frame. Each
image carries the GPS coordinates it was shot at, so the dataset supports both
detection and spatial analysis.
Each class is a specific poster design, not a generic "poster" category.
The intent is a recogniser that answers which known artwork is on the wall
and where its bounds… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/jerusalem-poster-detection.PosteriorBenchPoster100K
Poster100K Dataset
A comprehensive dataset containing 93K+ movie and TV show posters with detailed captions and text region annotations for multimodal learning and poster generation tasks.
Dataset Structure
image: Poster image in JPG/JPEG/PNG format
caption: Detailed textual description generated by Gemini-2.5-flash-preview-04-17
mask_regions: Text region coordinates (bounding boxes) in JSON format
file_name: Original filename
folder_path: Normalized relative folder path… See the full description on the dataset page: https://huggingface.co/datasets/momina884/Poster100K.themoviedb_tv_postersflux-poster-outpainting-trainMovie-Poster-WebURL-Dataset-1874-2025
Movie Poster WebURL Dataset 1874–2025
A TMDB-derived metadata index of movie poster WebURLs covering 1874–2025.
The dataset contains metadata and external TMDB poster URLs. Poster image binaries are not redistributed in this repository.
Data
Split: train
Rows: 804,304
Format: Parquet
Columns: 15
The publication artifact was produced from a larger local TMDB harvest and passed a conservative metadata-based content filtering and post-filter verification process… See the full description on the dataset page: https://huggingface.co/datasets/ROSCOSMOS/Movie-Poster-WebURL-Dataset-1874-2025.PosterErase
Dataset Card for PosterErase
Dataset Summary
PosterErase is a poster text-erasing dataset released with Self-supervised Text Erasing with Controllable Image Synthesis. It contains high-resolution poster images with text regions and structured annotations for text-erasing research.
This Hugging Face version exposes the original train, validation, and test splits as parquet files. The validation and test splits include ground-truth erased poster images; the… See the full description on the dataset page: https://huggingface.co/datasets/creative-graphic-design/PosterErase.movie-posters-audio-video
Movie Posters Audio Video Data Notes
Dataset summary
This repository contains a preparation pipeline and a small metadata sample for Movie Posters work with Audio Video inputs. It does not claim to be a complete benchmark release; the loader documents how source data is normalized and validated.
Included material
dataloader.py — loading, cleaning, and split preparation code.
dataset_infos.json — schema and split metadata.
metadata_sample.jsonl —… See the full description on the dataset page: https://huggingface.co/datasets/vitaliymtzm/movie-posters-audio-video.poster_data
Paper2Poster: Poster Data
Training and test poster data for the Paper2Poster layout training.
The dataset is sourced from the PosterLayout dataset.
Citation
If you use this dataset, please cite:
@inproceedings{qiang2016learning,
title={Learning to Generate Posters of Scientific Papers},
author={Yuting Qiang and Yanwei Fu and Yanwen Guo and Zhi-Hua Zhou and Leonid Sigal},
booktitle={Proceedings of the 30th AAAI Conference on Artificial Intelligence}… See the full description on the dataset page: https://huggingface.co/datasets/Paper2Poster/poster_data.OmniPSD_Layered_PosterPosterSum
POSTERSUM Dataset
Dataset Summary
The POSTERSUM dataset is a multimodal benchmark designed for the summarization of scientific posters into research paper abstracts. The dataset consists of 16,305 research posters collected from major machine learning conferences, including ICLR, ICML, and NeurIPS, spanning the years 2022-2024. Each poster is provided in image format along with its corresponding abstract as a summary. This dataset is intended for research in… See the full description on the dataset page: https://huggingface.co/datasets/rohitsaxena/PosterSum.Wang_Leehom_PosterCraft_F1_CaptionedAutoDeisgn-PosterBenchmovie_posters_100k_controlnetDataset Name: 10k Movie Poster Images with Layouts and Captions
Description:
This dataset contains 10,000 movie poster images, along with their extracted layout information and captions. The captions are generated by concatenating the movie title and genre(s). The layout annotations were extracted using PaddleOCR, providing precise structural details of the posters.
Source:
The dataset is a curated set of the movie-posters-100k dataset.
Key Features:
Images: 10,000 high-resolution movie… See the full description on the dataset page: https://huggingface.co/datasets/stzhao/movie_posters_100k_controlnet.movie-posters-genres-80k
Dataset Card for "movie-posters-genres-80k"
More Information needed
tmdb-poster
TMDB Poster
A collection of movie and TV poster images obtained from The Movie Database (TMDB).
Configurations
movie: 21,999 rows
tv: 3,979 rows
