[FEATURE] Add ViT weights: RADIO
Author: seefunCreated May 14, 2024Updated Jun 27, 2026
Labelsenhancement
https://github.com/NVlabs/RADIO
The code and model weights of paper [CVPR 2024] AM-RADIO: Agglomerative Vision Foundation Model - Reduce All Domains Into One has been released by Nvidia
RADIO , a new vision foundation model (actually a new vit pretrained weight), excels across visual domains, serving as a superior replacement for vision backbones. Integrating CLIP variants, DINOv2, and SAM through distillation, it preserves unique features like text grounding and segmentation correspondence.
Source: huggingface/pytorch-image-models