OpenThinker2-32B - Arena-Hard-Auto
Arena-Hard-Auto official Gemini-2.5 judged score 3.2 with CI -0.3/0.3
View sourceemilied137799
hi is a open-weight emilied137799 image generation model.
Running this yourself: can likely run on your own machine.
Live access is not confirmed for this model. No current purchase price is advertised.
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
This model is still tracked for research and discovery, but it is excluded from default public rankings until it returns to active status.
---
Quality Score
908
Arena ELO
Unknown
Parameters
---
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
10 of 22 public signals
Sign in to join the discussion
0
Downloads
0
Likes
---
Released
2/5 signals
3/4 signals
2/5 signals
1/4 signals
2/4 signals
Gaps we are still tracking
Benchmarks
5
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Arena-Hard-Auto official Gemini-2.5 judged score 3.2 with CI -0.3/0.3
View sourceSWE-Bench Verified resolved rate 66.6
View sourceLiveCodeBench pass@1 59.4 across 1055 tasks
LiveCodeBench pass@1 78.1 across 1055 tasks
View sourceSWE-Bench Verified resolved rate 66.6
View sourcehi is now available through local Ollama runtime. 4K context window listed. An experimental 1.1B parameter model trained on the new Dolphin 2.8 dataset by Eric Hartford and based on TinyLlama.
Arena-Hard-Auto official Gemini-2.5 judged score 3.2 with CI -0.3/0.3
SWE-Bench Verified resolved rate 66.6
LiveCodeBench pass@1 59.4 across 1055 tasks
LiveCodeBench pass@1 78.1 across 1055 tasks
SWE-Bench Verified resolved rate 66.6