Stable Diffusion v2 1 base Model Card This model card focuses on the model associated with the Stable Diffusion v2 1 base model. This stable diffusion 2 1 base model fine tunes stable diffusion 2 base ( 512 base ema.ckpt ) with 220k extra steps taken, with punsafe=0.98 on the same dataset. Use it with the stablediffusion repository: download the v2 1 512 ema pruned.ckpt here. Use it with 🧨 diffusers Model Details Developed by: Robin Rombach, Patrick Esser Model type: Diffusion based text to image generation model Language(s): English License: CreativeML Open RAIL++ M License Model Description: This is a model that can be used to generate and modify images based on text prompts. It is a Latent Diffusion Model that uses a fixed, pretrained text encoder (OpenCLIP ViT/H). Resources for more information: GitHub Repository. Cite as: @InProceedings{Rombach 2022 CVPR, author = {Rombach, Robin and Blattmann, Andreas and Lorenz, Dominik and Esser, Patrick and Ommer, Bj\"orn}, title = {High Resolution Image Synthesis With Latent Diffusion Models}, booktitle = {Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)}, month = {June}, year = {2022}, pages = {10…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy