Qwen3 VL 8B Instruct Benchmark Update
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
View sourceQwen
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
Running this yourself: desktop gpu should be enough.
33.1
Quality Score
---
Arena ELO
8B
Parameters
262K
Context
Sign in to join the discussion
0
Downloads
0
Likes
Oct 2025
Released
Benchmarks
19
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
View sourceQuality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
View sourceQuality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
View sourceQuality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
View sourceQuality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.2/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.4/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.4/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.4/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.4/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Quality: 8.4/100 | Price: $0.31/M tokens | Output: 0 tok/s | MMLU: 0.686% | HumanEval: 0.332%
Qwen3 VL 8B Instruct is now available through local Ollama runtime and Ollama Cloud. 256K context window listed. The most powerful vision-language model in the Qwen model family to date.