Qwen3-VL-4B-Instruct is now available on Ollama
Qwen3-VL-4B-Instruct is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.
View sourceQwen
Qwen3-VL-4B-Instruct is a open-weight Qwen multimodal model with a 262,144 token context window.
Running this yourself: consumer gpu should be enough.
38.9
Quality Score
---
Arena ELO
4B
Parameters
262K
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
14 of 22 public signals
Sign in to join the discussion
4.0M
Downloads
450
Likes
Oct 2025
Released
4/5 signals
1/4 signals
4/5 signals
3/4 signals
2/4 signals
Gaps we are still tracking
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Qwen3-VL-4B-Instruct is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.
View sourceQwen3-VL-4B-Instruct is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.