Playground v2.5 – 1024px Aesthetic Model This repository contains a model that generates highly aesthetic images of resolution 1024x1024, as well as portrait and landscape aspect ratios. You can use the model with Hugging Face 🧨 Diffusers. Playground v2.5 is a diffusion based text to image generative model, and a successor to Playground v2. Playground v2.5 is the state of the art open source model in aesthetic quality. Our user studies demonstrate that our model outperforms SDXL, Playground v2, PixArt α, DALL E 3, and Midjourney 5.2. For details on the development and training of our model, please refer to our blog post and technical report. Model Description Developed by: Playground Model type: Diffusion based text to image generative model License: Playground v2.5 Community License Summary: This model generates images based on text prompts. It is a Latent Diffusion Model that uses two fixed, pre trained text encoders (OpenCLIP ViT/G and CLIP ViT/L). It follows the same architecture as Stable Diffusion XL. Using the model with 🧨 Diffusers Install diffusers = 0.27.0 and the relevant dependencies. Notes: The pipeline uses the EDMDPMSolverMultistepScheduler scheduler by default, fo…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy