datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Synth-Text-Eng-512x128
Synthetic Text Images (English)
A synthetic dataset of rendered text images with rich per-sample
annotations: the text itself, its rendering attributes, background
description, applied post-processing, and a natural-language caption.
Each image is generated by compositing English text over a procedurally
generated background with random font, color, position, rotation, blur,
brightness and noise. All samples are accompanied by a structured
metadata.csv and a ready-to-use… See the full description on the dataset page: https://huggingface.co/datasets/Nininkkka/Synth-Text-Eng-512x128.text-auto-illustrate
Text Auto Illustrate — passage-to-image relevance judgements
Relevance judgements for illustrating prose: given a paragraph of Wikipedia
text, which images from a 5.4-million-image collection actually suit it?
Two judgement sets over the same 25 passages — one annotated by hand, one
generated and far broader — plus the metadata for every image either set names,
so the benchmark can be used without downloading the underlying corpus.
Built for a University of Glasgow final-year… See the full description on the dataset page: https://huggingface.co/datasets/domeist/text-auto-illustrate.Curated-Fox-News-Headlines-and-Full-Text
Curated Fox News Headlines and Full Text
This dataset contains a clean, curated collection of Fox News articles, including both headlines and full article text. It is designed for use in natural language processing (NLP) tasks such as sentiment analysis, summarization, topic classification, and media analysis.
📁 Dataset Format
Format: CSV
Encoding: UTF-8
Fields:
headline: The article title or headline
publish_date: Date the article was published (YYYY-MM-DD)
content:… See the full description on the dataset page: https://huggingface.co/datasets/crawlfeeds/Curated-Fox-News-Headlines-and-Full-Text.laion_text_debiased_60MFilter zxbsmk/laion_text_debiased_60M by image size and get 512 subset(12,009,641 pairs), 768 subset(4,915,850 pairs), 1024 subset(1,985,026 pairs).
text2food-mmc4This dataset is filtered version of MMC4 Multimodal-C4 core fewer-faces dataset . It contains 144 474 pair of food image url and image caption.
All the code and model in the repository.
movies
