Qwen3-VL-235B-A22B-Instruct
Alibaba · 2025-09-22 · 235.7B parameters
Qwen3-VL-235B-A22B-Instruct is the instruction-tuned variant of Alibaba's flagship open-weight Qwen3-VL vision-language MoE model. It supports text, image, and video understanding, multilingual OCR, spatial grounding, and visual-agent workflows in a non-thinking instruction-following mode. Its native context window is 262,144 tokens and can be extended to 1,000,000 tokens with YaRN.