VISTA
Ant Group (inclusionAI)๐จ๐ณ CNMajor labยท Jun 12, 2026
TextVisionopen weights
VISTA-9B are GUI-grounding vision-language models trained from Qwen3.5 9B backbones with VISTA: View-Consistent Self-Verified Training for GUI Grounding.
From the model card
- License
- apache-2.0
- Modalities
- text, image โ text
2 variants
Sources
HuggingFacePaperProvenance
- Detected
- Sep 27, 2026
- First source
- huggingface
- Description
- From the model card
- Editorially reviewed
- โ