Stable Diffusion 3 Medium Model Stable Diffusion 3 Medium is a Multimodal Diffusion Transformer (MMDiT) text to image model that features greatly improved performance in image quality, typography, complex prompt understanding, and resource efficiency. For more technical details, please refer to the Research paper. Please note: this model is released under the Stability Community License. For Enterprise License visit Stability.ai or contact us for commercial licensing details. Model Description Developed by: Stability AI Model type: MMDiT text to image generative model Model Description: This is a model that can be used to generate images based on text prompts. It is a Multimodal Diffusion Transformer (https://arxiv.org/abs/2403.03206) that uses three fixed, pretrained text encoders (OpenCLIP ViT/G, CLIP ViT/L and T5 xxl) License Community License: Free for research, non commercial, and commercial use for organisations or individuals with less than $1M annual revenue. You only need a paid Enterprise license if your yearly revenues exceed USD$1M and you use Stability AI models in commercial products or services. Read more: https://stability.ai/license For companies above this revenue…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy