Qwen 360 Diffusion General Qwen 360 Diffusion is a rank 128 LoRA built on top of a 20B parameter MMDiT (Multimodal Diffusion Transformer) model, designed to generate 360 degree equirectangular projection images from text descriptions. The model was trained from the Qwen Image model on an extremely diverse dataset composed of tens of thousands of equirectangular images, depicting landscapes, interiors, humans, animals, and objects. All images were resized to 2048x1024 before training. The model was also trained with a diverse dataset of normal photos for regularization, making the model a realism finetune when prompted correctly. Based on extensive testing, the model's capabilities vastly exceed all other currently available T2I 360 image generation models. Thus when given the right prompt, the model should be capable of producing almost anything you want. The model is designed to be capable of producing equirectangular images that can be used for non VR purposes such as general imagery, photography, artwork, architecture, portraiture, and many other concepts. Training Details The training dataset consists of 32k unique 360 degree equirectangular images. Each image was randomly rota…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy