huggingface.co

AI & ML interests

Generative approaches for visual synthesis, Invertible deep models for explainable AI, Deep metric and representation learning, self-supervised learning paradigms

Recent Activity

Robin Rombach's profile picture Apolinário from multimodal AI art's profile picture Anton Lozhkov's profile picture Patrick Esser's profile picture Suraj Patil's profile picture Emad Mostaque's profile picture Katherine Crowson's profile picture Pedro Cuenca's profile picture Chigozie's profile picture Stella Biderman's profile picture Leandro von Werra's profile picture Takyon236's profile picture David Marx's profile picture Ahmed Jawed's profile picture Yufan Zhou's profile picture avaer's profile picture John Doe's profile picture Corentin coKervadec's profile picture Sven Cattell's profile picture BM's profile picture Adam Chen's profile picture Stefan Baumann's profile picture Kolja Bauer's profile picture Nick Stracke's profile picture Owen Vincent's profile picture Vincent Tao Hu's profile picture Olga Grebenkova's profile picture Dima Kotovenko's profile picture Pingchuan Ma's profile picture Thomas Ressler-Antal's profile picture Enrico Shippole's profile picture Jannik Wiese's profile picture Ulrich Prestel's profile picture Tommaso Martorella's profile picture Reiner Birkl's profile picture

Welcome to CompVis!

We host public weights for Latent Diffusion and Stable Diffusion models. There are several options to choose from, please check the details below.

Stable Diffusion Models

Stable Diffusion is a latent text-to-image diffusion model capable of generating photo-realistic images given any text input. For more information about how Stable Diffusion works, please have a look at 🤗's Stable Diffusion with 🧨 Diffusers blog.

We recommend you use Stable Diffusion with 🤗 Diffusers library. You can also use the original CompVis code. There are variants of the weights depending on:

  • The library they are intended for.
  • The training regime. There are 4 training versions: v1-1 through v1-4 . Each one was created from the checkpoint of the previous version, and was trained for additional steps in specific variants of the dataset.

    Please, refer to the details in the following table to choose the weights appropriate for your use.

    Model Library Details
    stable-diffusion-v1-1 🤗 Diffusers 237k steps at resolution 256x256 on laion2B-en.
    194k steps at resolution 512x512 on laion-high-resolution.
    stable-diffusion-v1-2 🤗 Diffusers v1-1 plus:
    515k steps at 512x512 on "laion-improved-aesthetics".
    stable-diffusion-v1-3 🤗 Diffusers v1-2 plus:
    195k steps at 512x512 on "laion-improved-aesthetics",
    with 10% dropping of text-conditioning.
    stable-diffusion-v1-4 🤗 Diffusers v1-2 plus:
    225k steps at 512x512 on "laion-aesthetics v2 5+",
    with 10% dropping of text conditioning.
    stable-diffusion-v-1-1-original CompVis 237k steps at resolution 256x256 on laion2B-en.
    194k steps at resolution 512x512 on laion-high-resolution.
    stable-diffusion-v-1-2-original CompVis v1-1 plus:
    515k steps at 512x512 on "laion-improved-aesthetics".
    stable-diffusion-v-1-3-original CompVis v1-2 plus:
    195k steps at 512x512 on "laion-improved-aesthetics",
    with 10% dropping of text-conditioning.
    stable-diffusion-v-1-4-original CompVis v1-2 plus:
    225k steps at 512x512 on "laion-aesthetics v2 5+",
    with 10% dropping of text conditioning.

Read the original on huggingface.co ↗