Stable Diffusion 3 Medium Model Stable Diffusion 3 Medium is a Multimodal Diffusion Transformer (MMDiT) text to image model that features greatly improved performance in image quality, typography, complex prompt understanding, and resource efficiency. For more technical details, please refer to the Research paper. Please note: this model is released under the Stability Non Commercial Research Community License. For a Creator License or an Enterprise License visit Stability.ai or contact us for commercial licensing details. Model Description Developed by: Stability AI Model type: MMDiT text to image generative model Model Description: This is a model that can be used to generate images based on text prompts. It is a Multimodal Diffusion Transformer (https://arxiv.org/abs/2403.03206) that uses three fixed, pretrained text encoders (OpenCLIP ViT/G, CLIP ViT/L and T5 xxl) License Non commercial Use: Stable Diffusion 3 Medium is released under the Stability AI Non Commercial Research Community License. The model is free to use for non commercial purposes such as academic research. Commercial Use : This model is not available for commercial use without a separate commercial license from…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy