Model Overview Description This model performs visual feature extraction. For instance, RADIO generates image embeddings that can be used by a downstream model to classify images. C RADIOv4 models are available in multiple sizes: Shape Optimized (431M parameters). Huge (653M parameters). C RADIOv4 was trained using an updated set of teach models: SigLIP2 g DINOv3 7B SAM3 This model is ready for commercial/non commercial use. License/Terms of Use GOVERNING TERMS: Use of this model is governed by the NVIDIA Open Model License Agreement. Deployment Geography Global Use Case The embeddings generated by this model are expected to be used by a downstream application. For example: Image level understanding (image classification, curation, etc.). Dense processing (semantic segmentation, depth estimation, etc.). Integration into a Vision Language Model. Release Date Hugging Face: 01/27/2026 via RADIO Collection of Models. References AM RADIO: Agglomerative Vision Foundation Model Reduce All Domains Into One PHI S: Distribution Balancing for Label Free Multi Teacher Distillation RADIOv2.5: Improved Baselines for Agglomerative Vision Foundation Models FeatSharp: Your Vision Model Features, Sh…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy