ModelRadar
/

Mage-ViT

Microsoft🇺🇸 USMajor lab· Jul 26, 2026

Embeddingopen weights

Mage-ViT is the visual encoder at the core of Mage-VL. It is a Codec-ViT built primarily for video, where a single image is simply the degenerate one-frame case.

From the model card

License
mit
Modalities
image → embedding

1 variant

StandardHF

Sources

HuggingFace

Provenance

Detected
Sep 27, 2026
First source
huggingface
Description
From the model card
Editorially reviewed
—