OpenAI
OpenAI's most powerful reasoning model. Achieves state-of-the-art results on complex math, science, and code tasks through extended chain-of-thought.
Achieves state-of-the-art results on complex math, science, and code tasks through extended chain-of-thought.
Current rates are not verified. Missing pricing does not mean free usage.
No current subscription pricing is tracked for this model.
Confirm this specific model, usage limits, and billing terms with the provider. A subscription does not automatically include API credits.
No login needed to compare. Prices are in USD; provider charges are separate from AI Market Cap plans. Context length, caching, tools, taxes, and regional terms can change the final cost. Open weights do not mean free hosting.
48.1
Quality Score
1428
Arena ELO
Undisclosed
Parameters
200K
Context
This measures the amount of verifiable public evidence we have, not how capable the model is. A missing field means it has not been verified yet, not that its value is zero.
17 of 22 public signals
Sign in to join the discussion
0
Downloads
0
Likes
Apr 2025
Released
4/5 signals
4/4 signals
5/5 signals
1/4 signals
3/4 signals
Parameters
—
Training compute
Not reported
Dataset scale
Not reported
Base model
Not reported
Source-reported access: API access · Unknown confidence
Gaps we are still tracking
Launches
3
Benchmarks
10
General
6
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Arena-Hard-Auto official Gemini-2.5 judged score 50.0 with CI 0/0
Arena-Hard-Auto official Gemini-2.5 judged score 50.0 with CI 0/0
View sourceSWE-Bench Verified resolved rate 58.4
View sourceWe're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases may
SWE-Bench Verified resolved rate 58.4
LiveCodeBench pass@1 70.6 across 1055 tasks
Introducing GPT‑5 for developers | OpenAI Skip to main content Research Products Business Developers Company Foundation (opens in a new window) Log in Try ChatGPT (opens in a new window) Research Products Business Developers Company Foundation (opens in a new window) Try ChatGPT (opens in a new window) Login OpenAI August 7, 2025 Product Introducing GPT‑5 for developers The best model for coding and agentic tasks. Loading… Share Introduction Introduction Coding Frontend engin
SWE-Bench Verified resolved rate 58.4
Introducing OpenAI o3 and o4-mini | OpenAI Skip to main content Research Products Business Developers Company Foundation (opens in a new window) Log in Try ChatGPT (opens in a new window) Research Products Business Developers Company Foundation (opens in a new window) Try ChatGPT (opens in a new window) Login OpenAI April 16, 2025 Release Product Introducing OpenAI o3 and o4-mini Try on ChatGPT (opens in a new window) Loading… Share What’s changed What’s changed Continuing to
Arena-Hard-Auto official Gemini-2.5 judged score 50.0 with CI 0/0
Arena-Hard-Auto official Gemini-2.5 judged score 85.9 with CI -0.8/0.9
GAIA score 30.2 from Test_agent_part3
ChatGPT Start searching API Dashboard Try ChatGPT Home API Codex Docs Guides, concepts, and product docs for Codex Use cases Example workflows and tasks teams can take on with ChatGPT or Codex Docs Use cases Training Resources ChatGPT Plugins Extend ChatGPT and Codex Workspace Agents Trigger published ChatGPT workspace agents Commerce Build commerce flows in ChatGPT Ads Publish and measure ads in ChatGPT Resources Showcase Demo apps to get inspired Blog Learnings and experien