Team Ai
Modelpublic

JuyeopDang/simple-latent-diffusion-models

sourceHugging Facemitupdated 6mo agoView on Hugging Face
0likes
Model Card

Simple Latent Diffusion Model (LDM)

This repository contains the pre-trained weights and configuration files for the Simple Latent Diffusion Model project.

For the full source code, detailed explanations, and implementation logic, please visit the original GitHub repository.

๐Ÿš€ Model Description

This project implements a Latent Diffusion Model (LDM) from scratch. The repository includes:

  • โ€”Custom-trained VAE: For compressing images into a latent space.
  • โ€”Diffusion Model: A U-Net based architecture for the reverse diffusion process.
  • โ€”CLIP Weights: Integrated for text-guided image generation.

๐Ÿ“‚ Available Models & Checkpoints

The repository provides weights for three different datasets, covering both unconditional and conditional generation tasks:

DatasetTypeDescription
CIFAR-10Unconditional32x32 image generation based on CIFAR-10 classes.
CelebAUnconditionalHuman face generation trained on the CelebA dataset.
Asian CompositeText-to-Image (T2I)CLIP-based conditional generation using the Asian Composite Dataset.

๐Ÿ›  How to Use

If you want to experiment with these models and generate your own images, we provide a hands-on example notebook.

  1. 1.Open the `cifar10_example.ipynb` file provided in this repository.
  2. 2.Follow the step-by-step instructions to load the configurations and model weights.
  3. 3.Run the cells to start the sampling process and generate images.

๐Ÿ”— References


Note: Ensure you have the necessary dependencies installed as specified in the GitHub repository's requirements.