Model Overview [ Github ] [ CVPR 2025 ] [ CVPR 2024 ] Description This model performs visual feature extraction. For instance, RADIO generates image embeddings that can be used by a downstream model to classify images. C RADIOv3 models are available in multiple sizes: Base (90M parameters). Large (320M parameters). Huge (653M parameters). (In training) Gigantic (1.1B parameters). C RADIOv3 was trained for 1M steps (400k more steps than v1), using inverse frequency sampling for data balancing, and PHI Standardization for teacher distribution balancing. As well as new techniques for summary distribution matching, and domain generalization. This model is ready for commercial/non commercial use. License/Terms of Use GOVERNING TERMS: Use of this model is governed by the NVIDIA Open Model License Agreement. Deployment Geography Global. Use Case The embeddings generated by this model are expected to be used by a downstream application. For example: Image level understanding (image classification, curation, etc.). Dense processing (semantic segmentation, depth estimation, etc.). Integration into a Vision Language Model. Release Date Huggingface: 03/26/2025 via RADIO Collection of Models. R…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy