Team Ai
Datasetpublic

whosouravsharma/text-to-image-diffusiondb-2M

DiffusionDB text-to-image subset A cleaned, safety-filtered image-prompt dataset for training a text-to-image model, built from DiffusionDB. Built on Hugging Face Jobs directly from poloclub/diffusiondb. It covers part_id 1-20 (20,000 source images) before filtering. The same content is also kept on the 20k-subset branch. Load it with: load_dataset("whosouravsharma/text-to-image-diffusiondb-2M") Note on the repo name: despite "2M" in the name, this is a small slice of… See the full description on the dataset page: https://huggingface.co/datasets/whosouravsharma/text-to-image-diffusiondb-2M.

sourceHugging Facecc0-1.0updated 2mo agoView on Hugging Face
2likes205downloads
settings

This repository belongs to whosouravsharma on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nametext-to-image-diffusiondb-2M
visibilitypublic
licencecc0-1.0
gatedno
ownerwhosouravsharma
Account settings
whosouravsharma/text-to-image-diffusiondb-2M · Team Ai