Team Ai
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Rapidata /text-2-image-Rich-Human-Feedback Building upon Google's research Rich Human Feedback for Text-to-Image Generation we have collected over 1.5 million responses from 152'684 individual humans using Rapidata via the Python API. Collection took roughly 5 days. If you get value from this dataset and would like to see more in the future, please consider liking it. Overview We asked humans to evaluate AI-generated images in style, coherence and prompt alignment. For images that contained flaws, participants were… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/text-2-image-Rich-Human-Feedback.imagetext-to-image10K<n<100K37 likes622 downloads15d agoHugging Face02Rapidata /text-2-image-Rich-Human-Feedback-32k Building upon Google's research Rich Human Feedback for Text-to-Image Generation, and the smaller, previous version of this dataset, we have collected over 3.7 million responses from 307'415 individual humans for the open-image-preference-v1 dataset using Rapidata via the Python API. Collection took less than 2 weeks. If you get value from this dataset and would like to see more in the future, please consider liking it ♥️ Overview We asked humans to evaluate AI-generated… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/text-2-image-Rich-Human-Feedback-32k.image10K<n<100K26 likes452 downloads15d agoHugging Face03pszemraj /text2image-multi-prompt text2image multi-prompt(s): a dataset collection collection of several text2image prompt datasets data was cleaned/normalized with the goal of removing "model specific APIs" like the "--ar" for Midjourney and so on data de-duplicated on a basic level: exactly duplicate prompts were dropped (after cleaning and normalization) updates Oct 2023: the default config has been updated with better deduplication. It was deduplicated with minhash (params: n-gram size set to 3… See the full description on the dataset page: https://huggingface.co/datasets/pszemraj/text2image-multi-prompt.texttext-generation1M<n<10M10 likes104 downloads10mo agoHugging Face04datapointai /text-2-image-human-preferences-2mgated Text-to-image human preferences: 2M votes across 30 models This dataset contains the complete voting record behind the Datapoint Image Bench leaderboard: 2,161,160 validated pairwise votes — exactly 10 for each of 216,116 image pairs. The votes compare 30 text-to-image models in a complete round-robin on 500 prompts, judged by annotators from over 200 countries. Every vote includes the annotator's trust score at the time the vote was cast. Built on the Datapoint annotation… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-human-preferences-2m.imagetext-to-image1M<n<10M21 likes102 downloads2mo agoHugging Face05HankYe /Sampled_AIGCBench_text2image_ar_0.625 Description This dataset is intended for the implementation of image-to-video generation evaluations in the paper of AdaptiveDiffusion, which is composed of the original text-image pairs collected from AIGCBench v1.0 and a text file listing the randomly selected samples. Data Organization The dataset is organized into the following files: AIGCBench_t2i_aspect_ratio_625.zip: 2002 images named by the index and the text description, adjusted to an aspect ratio of 0.625.… See the full description on the dataset page: https://huggingface.co/datasets/HankYe/Sampled_AIGCBench_text2image_ar_0.625.imageimage-to-video1K<n<10K0 likes60 downloads2y agoHugging Face06TREC-AToMiC /atomic2023-small_text2imageimage10K<n<100K1 likes50 downloads2y agoHugging Face07LLAAMM /text2image100kimage100K<n<1M0 likes48 downloads2y agoHugging Face08JackyZhuo /ShareGPT-4o-Text2Imageimage10K<n<100K0 likes47 downloads1y agoHugging Face09haoxianc /antonkozyriev_ai-text2image-tweets AI Text-2-Image Tweets (+ Sentiment labels) Mirror of the Kaggle dataset antonkozyriev/ai-text2image-tweets by Anton Kozyriev, released under CC0: Public Domain. All credit goes to the original author; please cite and link the Kaggle page when using this data. Tweets about DALLE-2, GLIDE, Imagen, and Stable Diffusion AI models Original description (from Kaggle) Attribution Dataset thumbnail by OpenAI. Context These tweets are about… See the full description on the dataset page: https://huggingface.co/datasets/haoxianc/antonkozyriev_ai-text2image-tweets.tabular10K<n<100K0 likes32 downloads5d agoHugging Face10datapointai /text-2-image-dpo-human-preferences-fullgated Text-2-Image DPO Human Preferences (Full) The complete human preference dataset for text-to-image generation. 416,360 pairwise judgments from ~20,000 annotators comparing AI-generated images across two evaluation dimensions: prompt alignment and overall preference. This is the full, unfiltered version with uniform vote weights. For quality-filtered subsets with calibrated annotator weighting, see: datapointai/text-2-image-dpo-human-preferences (5,000 pairs, trust-weighted)… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-dpo-human-preferences-full.imageimage-classification10K<n<100K1 likes30 downloads6mo agoHugging Face11zqman /Text2image-ChinesePainting Text2image-ChinesePainting Dataset This repository provides 2192 pairs of Traditional Chinese Landscape Paintings and their corresponding descriptive texts. The paintings are sourced from the Chinese Landscape Painting Dataset, and the descriptive texts are generated using GPT-4o's API. The dataset is intended for use in training and fine-tuning text-to-image models, focusing on generating Chinese landscape art from textual descriptions. Dataset Overview Number of… See the full description on the dataset page: https://huggingface.co/datasets/zqman/Text2image-ChinesePainting.image1K<n<10K5 likes22 downloads2y agoHugging Face12shirsho12 /text2image-10k-with-spectacles-pairs text2image-10k-with-spectacles-pairs A 10k text-only dataset combining: 9,000 prompts sampled (streaming) from jackyhate/text-to-image-2M 1,000 rows drawn from user-provided spectacles pairs (both base and with "wearing spectacles" versions are included as separate rows) Schema text: the prompt is_from_pair: whether this row came from the spectacles pairs has_spectacles_phrase: whether the text explicitly includes "wearing spectacles" source:… See the full description on the dataset page: https://huggingface.co/datasets/shirsho12/text2image-10k-with-spectacles-pairs.text10K<n<100K0 likes22 downloads11mo agoHugging Face13datapointai /text-2-image-dpo-human-preferences-smallgated Text-2-Image DPO Human Preferences (Small) A quality-controlled human preference dataset for text-to-image generation. 40,000 trust-weighted pairwise judgments from calibrated annotators comparing AI-generated images across two evaluation dimensions: prompt alignment and overall preference. This is the highest-annotator-quality subset. For the full 5,000-pair dataset, see datapointai/text-2-image-dpo-human-preferences. Built on the Datapoint annotation platform — purpose-built… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-dpo-human-preferences-small.imageimage-classification1K<n<10K3 likes16 downloads6mo agoHugging Face14wusize /puffin_text2imagetext1K<n<10K0 likes15 downloads1y agoHugging Face15LLAAMM /text2image10kimage10K<n<100K1 likes14 downloads2y agoHugging Face16LLAAMM /text2image1mimagetext-to-image1M<n<10M0 likes11 downloads2y agoHugging Face17datapointai /text-2-image-dpo-human-preferencesgated Text-2-Image DPO Human Preferences A large-scale, quality-controlled human preference dataset for text-to-image generation. 80,000 trust-weighted pairwise judgments from calibrated annotators comparing AI-generated images across two evaluation dimensions: prompt alignment and overall preference. Built on the Datapoint annotation platform — purpose-built infrastructure for collecting high-quality human preference data at scale. Overview Metric Value Total… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-dpo-human-preferences.imageimage-classification1K<n<10K1 likes11 downloads6mo agoHugging Face18xchuan /text2image-fupoimagen<1K1 likes10 downloads2y agoHugging Face19nirmalendu01 /text2image-10k-with-spectacles-pairs text2image-10k-with-spectacles-pairs A 10k text-only dataset combining: 9,000 prompts sampled (streaming) from jackyhate/text-to-image-2M 1,000 rows drawn from user-provided spectacles pairs (both base and with "wearing spectacles" versions are included as separate rows) Schema text: the prompt is_from_pair: whether this row came from the spectacles pairs has_spectacles_phrase: whether the text explicitly includes "wearing spectacles" source:… See the full description on the dataset page: https://huggingface.co/datasets/nirmalendu01/text2image-10k-with-spectacles-pairs.text10K<n<100K0 likes7 downloads1y agoHugging Face20JohnTeddy3 /text2image-multi-prompt###转载 pszemraj/text2image-multi-prompt text2image multi-prompt(s): a dataset collection collection of several text2image prompt datasets data was cleaned/normalized with the goal of removing "model specific APIs" like the "--ar" for Midjourney and so on data de-duplicated on a basic level: exactly duplicate prompts were dropped (after cleaning and normalization) contents DatasetDict({ train: Dataset({ features: ['text', 'src_dataset'], num_rows:… See the full description on the dataset page: https://huggingface.co/datasets/JohnTeddy3/text2image-multi-prompt.text1M<n<10M0 likes6 downloads3y agoHugging Face21nthuy652 /text2image_en_vi_captionsimage10K<n<100K0 likes5 downloads2y agoHugging Face22xchuan /text2image-manimagen<1K0 likes3 downloads2y agoHugging Face23aa-nadim /text2images0 likes3 downloads2y agoHugging Face24WindSun /text2image0 likes2 downloads3y agoHugging Face25Takeru /Text2image0 likes2 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.