Llama-3.1-Nemotron-70B-Instruct-HF - Arena-Hard-Auto
Arena-Hard-Auto official Gemini-2.5 judged score 10.3 with CI -0.8/1
View sourcedesert-ant-labs
emo is a proprietary desert-ant-labs specialized model.
Live access is not confirmed for this model. No current purchase price is advertised.
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
---
Quality Score
1224
Arena ELO
Undisclosed
Parameters
---
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
13 of 22 public signals
Sign in to join the discussion
26.9K
Downloads
5
Likes
Jun 2026
Released
2/5 signals
3/4 signals
3/5 signals
3/4 signals
2/4 signals
Gaps we are still tracking
Benchmarks
3
Open Source
1
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Arena-Hard-Auto official Gemini-2.5 judged score 10.3 with CI -0.8/1
View sourceSWE-Bench Verified resolved rate 68.2
View sourceLiveCodeBench pass@1 81.0 across 1055 tasks
emo is now available through local Ollama runtime. 128K context window listed. NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
View sourceemo is now available through local Ollama runtime. 128K context window listed. NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
Arena-Hard-Auto official Gemini-2.5 judged score 10.3 with CI -0.8/1
SWE-Bench Verified resolved rate 68.2
LiveCodeBench pass@1 81.0 across 1055 tasks