ModelRadar
/

VISTA

Ant Group (inclusionAI)๐Ÿ‡จ๐Ÿ‡ณ CNMajor labยท Jun 12, 2026

TextVisionopen weights

VISTA-9B are GUI-grounding vision-language models trained from Qwen3.5 9B backbones with VISTA: View-Consistent Self-Verified Training for GUI Grounding.

From the model card

License
apache-2.0
Modalities
text, image โ†’ text

2 variants

9B9BHF
4B4BHF

Sources

HuggingFacePaper

Provenance

Detected
Sep 27, 2026
First source
huggingface
Description
From the model card
Editorially reviewed
โ€”