Qwen/Qwen2.5-14B-Instruct — Open LLM Leaderboard #173
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
View sourceQwen
In the past three months since Qwen2’s release, numerous developers have built new models on the Qwen2 language models, providing us with valuable feedback. During this period, we have focused on creating smarter and more knowledgeable language models.
Running this yourself: desktop gpu should be enough.
Today, we are excited to introduce the latest addition to the Qwen family: Qwen2.5.
Live access is not confirmed for this model. No current purchase price is advertised.
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
36.8
Quality Score
---
Arena ELO
15B
Parameters
131K
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
18 of 22 public signals
Sign in to join the discussion
82.0K
Downloads
157
Likes
Sep 2024
Released
4/5 signals
3/4 signals
4/5 signals
3/4 signals
4/4 signals
Parameters
15B
Training compute
1.6e24 FLOP
Dataset scale
Not reported
Base model
Not reported
Source-reported access: Open weights (unrestricted) · Confident confidence
Gaps we are still tracking
Benchmarks
19
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
View sourceAvg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
View sourceAvg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
View sourceAvg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
View sourceAvg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
View sourceAvg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Avg: 41.6 | IFEval: 81.4 | BBH: 48.6 | MATH: 55.3 | GPQA: 10.5 | MMLU-PRO: 43.2
Avg: 41.3 | IFEval: 81.6 | BBH: 48.4 | MATH: 54.8 | GPQA: 9.6 | MMLU-PRO: 43.4
Qwen2.5-14B is now available through local Ollama runtime. 32K context window listed. Qwen2.5 models are pretrained on Alibaba's latest large-scale dataset, encompassing up to 18 trillion tokens. The model supports up to 128K tokens and has multilingual support.