Qwen3-VL-4B-Thinking is now available on Ollama
Qwen3-VL-4B-Thinking is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.
View sourceQwen
Qwen3-VL-4B-Thinking is a open-weight Qwen multimodal model with a 262,144 token context window.
Running this yourself: consumer gpu should be enough.
34.6
Quality Score
---
Arena ELO
4B
Parameters
262K
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
14 of 22 public signals
Sign in to join the discussion
32.2K
Downloads
120
Likes
Oct 2025
Released
4/5 signals
1/4 signals
4/5 signals
3/4 signals
2/4 signals
Gaps we are still tracking
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Qwen3-VL-4B-Thinking is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.
View sourceQwen3-VL-4B-Thinking is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.