Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01poloclub /diffusiondbDiffusionDB is the first large-scale text-to-image prompt dataset. It contains 2 million images generated by Stable Diffusion using prompts and hyperparameters specified by real users. The unprecedented scale and diversity of this human-actuated dataset provide exciting research opportunities in understanding the interplay between prompts and generative models, detecting deepfakes, and designing human-AI interaction tools to help users more easily use these models.text-to-imagen>1T666 likes13k downloads3y agoHugging Face02yjernite /prof_report__wavymulder-Analog-Diffusion__multi__24 Dataset Card for "prof_report__wavymulder-Analog-Diffusion__multi__24" More Information needed tabular1K<n<10K0 likes5.2k downloads3y agoHugging Face03Gustavosta /Stable-Diffusion-Prompts Stable Diffusion Dataset This is a set of about 80,000 prompts filtered and extracted from the image finder for Stable Diffusion: "Lexica.art". It was a little difficult to extract the data, since the search engine still doesn't have a public API without being protected by cloudflare. If you want to test the model with a demo, you can go to: "spaces/Gustavosta/MagicPrompt-Stable-Diffusion". If you want to see the model, go to: "Gustavosta/MagicPrompt-Stable-Diffusion". text10K<n<100K528 likes4.4k downloads4y agoHugging Face04robotics-diffusion-transformer /rdt-ft-data Dataset Card This is the fine-tuning dataset used in the paper RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation. Source Project Page: https://rdt-robotics.github.io/rdt-robotics/ Paper: https://arxiv.org/pdf/2410.07864 Code: https://github.com/thu-ml/RoboticsDiffusionTransformer Model: https://huggingface.co/robotics-diffusion-transformer/rdt-1b Uses Download all archive files and use the following command to extract: cat rdt_data.tar.gz.* |… See the full description on the dataset page: https://huggingface.co/datasets/robotics-diffusion-transformer/rdt-ft-data.25 likes4.2k downloads2y agoHugging Face05hw-liang /Diffusion4D Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models [Project Page] | [Arxiv] | [Code] News 2024.6.28: Released rendered data from curated objaverse-xl. 2024.6.4: Released rendered data from curated objaverse-1.0, including orbital videos of dynamic 3D, orbital videos of static 3D, and monocular videos from front view. 2024.5.27: Released metadata for objects! Overview We collect a large-scale, high-quality dynamic… See the full description on the dataset page: https://huggingface.co/datasets/hw-liang/Diffusion4D.3dtext-to-3d1M<n<10M33 likes4.1k downloads2y agoHugging Face06tyDiffusion /Diffusion0 likes3.7k downloads2y agoHugging Face07hanamizuki-ai /stable-diffusion-v1-5-glazed Dataset Card for Stable Diffusion v1.5 Glazed Samples Dataset Description Dataset Summary This dataset contains image samples originally generated by runwayml/stable-diffusion-v1-5 and subsequently processed by Glaze tool. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information… See the full description on the dataset page: https://huggingface.co/datasets/hanamizuki-ai/stable-diffusion-v1-5-glazed.imageimage-classification100K<n<1M4 likes3.6k downloads3y agoHugging Face08LoveAronaPlana /Stable-diffusion2 likes3.3k downloads1y agoHugging Face09bayes-group-diffusion /OAS95-aligned-cleanedtext100M<n<1B3 likes3.1k downloads10mo agoHugging Face10sasha /prof_images_blip__22h-vintedois-diffusion-v0-1 Dataset Card for "prof_images_blip__22h-vintedois-diffusion-v0-1" More Information needed image10K<n<100K0 likes2.5k downloads3y agoHugging Face11yjernite /prof_report__22h-vintedois-diffusion-v0-1__multi__24 Dataset Card for "prof_report__22h-vintedois-diffusion-v0-1__multi__24" More Information needed tabularn<1K0 likes2.5k downloads3y agoHugging Face12CaiYuanhao /DiffusionGS [ICCV 2025] DiffusionGS: Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction Data Description These are the demo results of our ICCV 2025 paper. HuggingFace Model Link We also release our models in HuggingFace: https://huggingface.co/CaiYuanhao/DiffusionGS Here are some video generation results demo: · (a) Object-level Generation… See the full description on the dataset page: https://huggingface.co/datasets/CaiYuanhao/DiffusionGS.3dimage-to-3dn<1K11 likes2.4k downloads9mo agoHugging Face13DiffusionWave /models0 likes2.3k downloads1mo agoHugging Face14TIGER-Lab /RationalRewards_DiffusionNFT_TrainDataTLDR: this is the diffusion RL training dataset for text-to-image generation and image editing, from the following paper. RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time Haozhe Wang1   Cong Wei2   Weiming Ren2   Jiaming Liu3   Fangzhen Lin1   Wenhu Chen2 1 HKUST   2 University of Waterloo   3 Alibaba RationalRewards is a… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/RationalRewards_DiffusionNFT_TrainData.tabulartext-to-image10K<n<100K1 likes2.2k downloads6mo agoHugging Face15AbstractPhil /diffusion-pretrain-set-ft1 diffusion-pretrain-set-ft1 A multi-source image-caption pretraining dataset assembled from ten upstream sources via a uniform ingest pipeline. Designed for a full pretrain or finetune pipeline meant to curate for any major diffusion model preliminary, with the sole intent to create a more powerful baseline preliminary train and a baseline for synthesizing images to train the next generation of the VLM model. This is a lot like the snake eating it's own tail, so it must be… See the full description on the dataset page: https://huggingface.co/datasets/AbstractPhil/diffusion-pretrain-set-ft1.image1M<n<10M3 likes2.1k downloads4mo agoHugging Face16DiffusionLight /text2dataset Text2Dataset the dataset generate from ChatGPT's output and text2light for training the LoRA using in DiffusionLight Code for training the LoRA can be found at DiffusionLight-LoRA-Trainer imagen<1K0 likes2.1k downloads2y agoHugging Face17sasha /prof_images_blip__stabilityai-stable-diffusion-2 Dataset Card for "prof_images_blip__stabilityai-stable-diffusion-2" More Information needed image10K<n<100K1 likes2k downloads3y agoHugging Face18WhiteAiZ /stable-diffusion-webui-reForge Stable Diffusion WebUI Forge/reForge Stable Diffusion WebUI Forge/reForge is a platform on top of Stable Diffusion WebUI (based on Gradio) to make development easier, optimize resource management, speed up inference, and study experimental features. The name "Forge" is inspired from "Minecraft Forge". This project is aimed at becoming SD WebUI's Forge. Forge2/reForge2 You can read more on… See the full description on the dataset page: https://huggingface.co/datasets/WhiteAiZ/stable-diffusion-webui-reForge.0 likes1.9k downloads10mo agoHugging Face19Yzl-code /RS-Diffusion0 likes1.9k downloads2y agoHugging Face20DenisKochetov /Diffusion4D-Animated-Mocap-7440 Diffusion4D Animated Mocap 7440 Rolling export of 7,440 technically and semantically selected animated assets. Each asset has 12 yaw views (0..330 degrees, step 30), 24 sampled frames, BVH motion, DINOv2 image embeddings, and one PNG preview. Batch archives are uploaded only after local stage validation. imagefeature-extraction1K<n<10K3 likes1.5k downloads2mo agoHugging Face21gmongaras /Stable_Diffusion_3_RecaptionThis dataset is the one specified in the stable diffusion 3 paper which is composed of the ImageNet dataset and the CC12M dataset. I used the ImageNet 2012 train/val data and captioned it as specified in the paper: "a photo of a 〈class name〉" (note all ids are 999,999,999) CC12M is a dataset with 12 million images created in 2021. Unfortunately the downloader provided by Google has many broken links and the download takes forever. However, some people in the community publicized the dataset.… See the full description on the dataset page: https://huggingface.co/datasets/gmongaras/Stable_Diffusion_3_Recaption.image10M<n<100M7 likes1.5k downloads2y agoHugging Face22fffffchopin /DiffusionDream_DatasetThis is the dataset of the diffusion dream dataset. The dataset contains the following columns: info: A string describing the action taken in the frame keyword: A string describing the keyword of the action action: A string describing the action taken in the frame current_frame: The current frame of the video previous_frame_1: The frame before the current frame previous_frame_2: The frame before the previous frame previous_frame_3: The frame before the previous frame previous_frame_4: The… See the full description on the dataset page: https://huggingface.co/datasets/fffffchopin/DiffusionDream_Dataset.image100K<n<1M1 likes1.4k downloads2y agoHugging Face23DenisKochetov /Diffusion4D-Animated-Raw Diffusion4D Animated Assets This dataset provides animated 3D assets referenced by Diffusion4D and Objaverse-XL in a directly browsable format. The default split contains 67,988 rows. Each row includes metadata, a preview image, and a short preview video so that assets can be inspected in the Hugging Face Data Studio without first downloading the original 3D file. The repository also mirrors available raw assets and keeps their original source links and hashes. The current… See the full description on the dataset page: https://huggingface.co/datasets/DenisKochetov/Diffusion4D-Animated-Raw.3d10K<n<100K5 likes1.4k downloads3mo agoHugging Face24HuggingFace-CN-community /Diffusion-book-cn 《从零开始学扩散模型》 术语表 词汇 翻译 Corruption Process 退化过程 Pipeline 管线 Timestep 时间步 Scheduler 调度器 Gradient Accumulation 梯度累加 Fine-Tuning 微调 Guidance 引导 目录 第一部分 基础知识 第一章 扩散模型的原理、发展和应用 1.1 扩散模型的原理 1.2 扩散模型的发展 1.3 扩散模型的应用 第二章 HuggingFace介绍与环境准备 2.1 HuggingFace Space 2.2 Transformer 与 diffusers 库 2.3 环境准备… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFace-CN-community/Diffusion-book-cn.imagen<1K82 likes1.3k downloads3y agoHugging Face25brandonyang /diffusion_policy_robocasa_activations_latest_chkpt Diffusion Policy — RoboCasa Activations (latest checkpoint) Per-step, per-episode activation traces collected from a DiffusionTransformerHybridImagePolicy (the diffusion_policy library) rolled out on RoboCasa benchmark tasks. Captured with collect_activations_robocasa.py at the latest training checkpoint. These traces are the input expected by the conceptor / SAE steering pipelines under diffusion_policy/experiments/robocasa_steering/ and diffusion_policy/experiments/sae/ — see… See the full description on the dataset page: https://huggingface.co/datasets/brandonyang/diffusion_policy_robocasa_activations_latest_chkpt.robotics100B<n<1T0 likes1.3k downloads5mo agoHugging Face26dnlpy /Stable-Diffusion1 likes1.3k downloads2y agoHugging Face27ANWERFATEHY /encoded_images_captions_for_t2i_diffusion_traininggatedhundreds of thousands of images of mainly beautiful females, nature/landskape, birds/cats, architecture/buildings, etc some are handpicked from free images sites and the rest from several images/captions datasets in huggingface no images of violence or harm or ugly things are included all encoded with their captions that are made by several models, mainly qwen and lfm, the script i used to encode them are included in this dataset there are about 5-10% images of consensually sexually-nude… See the full description on the dataset page: https://huggingface.co/datasets/ANWERFATEHY/encoded_images_captions_for_t2i_diffusion_training.0 likes1.2k downloads4d agoHugging Face28diffusion-bench /blip3o-256image1K<n<10K1 likes1.2k downloads7mo agoHugging Face29PennyJX /stable-diffusion-webui Stable Diffusion web UI A browser interface based on Gradio library for Stable Diffusion. Features Detailed feature showcase with images: Original txt2img and img2img modes One click install and run script (but you still must install python and git) Outpainting Inpainting Color Sketch Prompt Matrix Stable Diffusion Upscale Attention, specify parts of text that the model should pay more attention to a man in a ((tuxedo)) - will pay more attention to tuxedo a man in a… See the full description on the dataset page: https://huggingface.co/datasets/PennyJX/stable-diffusion-webui.imagen<1K1 likes1.2k downloads3y agoHugging Face30tauhuang /diffusion_rewardimage10K<n<100K1 likes1.2k downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.