Qwen2-57B-A14B — BigCodeBench #36
Complete: 46.8 | Instruct: 36.1 | 57B params
View sourceQwen
This report introduces the Qwen2 series, the latest addition to our large language models and large multimodal models. We release a comprehensive suite of foundational and instruction-tuned language models, encompassing a parameter range from 0.
Running this yourself: likely needs a rented cloud gpu.
This report introduces the Qwen2 series, the latest addition to our large language models and large multimodal models.
Live access is not confirmed for this model. No current purchase price is advertised.
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
39.5
Quality Score
---
Arena ELO
57B
Parameters
131K
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
19 of 22 public signals
Sign in to join the discussion
10.9K
Downloads
59
Likes
May 2024
Released
5/5 signals
3/4 signals
4/5 signals
3/4 signals
4/4 signals
Parameters
57B
Training compute
3.8e23 FLOP
Dataset scale
Not reported
Base model
Not reported
Source-reported access: Open weights (unrestricted) · Confident confidence
Gaps we are still tracking
Benchmarks
19
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Complete: 46.8 | Instruct: 36.1 | 57B params
View sourceComplete: 46.8 | Instruct: 36.1 | 57B params
View sourceComplete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
View sourceComplete: 46.8 | Instruct: 36.1 | 57B params
View sourceQwen2-57B-A14B is now available through local Ollama runtime. 32K context window listed. Qwen2.5 models are pretrained on Alibaba's latest large-scale dataset, encompassing up to 18 trillion tokens. The model supports up to 128K tokens and has multilingual support.
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Complete: 46.8 | Instruct: 36.1 | 57B params
Qwen published benchmark or leaderboard evidence for Qwen2-57B-A14B.