Qwen3 Next 80B A3B Instruct Benchmark Update
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 171.206 tok/s | MMLU: 0.819% | HumanEval: 0.684%
View sourceQwen
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Running this yourself: likely needs a high-memory cloud gpu.
OpenRouter
Price record: 2026-10-03. Source: openrouter.
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
46.3
Quality Score
---
Arena ELO
81B
Parameters
262K
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
18 of 22 public signals
Sign in to join the discussion
292.3K
Downloads
1.1K
Likes
Sep 2025
Released
5/5 signals
3/4 signals
5/5 signals
3/4 signals
2/4 signals
Gaps we are still tracking
Benchmarks
19
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 171.206 tok/s | MMLU: 0.819% | HumanEval: 0.684%
View sourceQuality: 9.6/100 | Price: $0.412/M tokens | Output: 171.206 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 171.206 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 171.439 tok/s | MMLU: 0.819% | HumanEval: 0.684%
View sourceQuality: 9.6/100 | Price: $0.412/M tokens | Output: 160.767 tok/s | MMLU: 0.819% | HumanEval: 0.684%
View sourceQuality: 9.6/100 | Price: $0.412/M tokens | Output: 177.993 tok/s | MMLU: 0.819% | HumanEval: 0.684%
View sourceQuality: 9.6/100 | Price: $0.412/M tokens | Output: 171.206 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 171.439 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 160.767 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 177.993 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 183.498 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 182.87 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 172.636 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 166.183 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 158.252 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 166.844 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 166.844 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 184.291 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 173.554 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 177.194 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 179.891 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 191.447 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 196.439 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Quality: 9.6/100 | Price: $0.412/M tokens | Output: 191.761 tok/s | MMLU: 0.819% | HumanEval: 0.684%
Qwen3 Next 80B A3B Instruct is now available through local Ollama runtime. 256K context window listed. The first installment in the Qwen3-Next series with strong performance in terms of both parameter efficiency and inference speed.