OpenAI
Cost-efficient reasoning model with adjustable reasoning effort. Optimised for STEM and coding tasks at a fraction of o3's inference cost.
44.4
Quality Score
1318
Arena ELO
Undisclosed
Parameters
200K
Context
Sign in to join the discussion
0
Downloads
0
Likes
Jan 2025
Released
Launches
3
Benchmarks
5
Research
1
General
5
Recent launch, pricing, benchmark, and API signals linked to this model or its provider.
Arena-Hard-Auto official Gemini-2.5 judged score 50.0 with CI 0/0
Now in preview: The ChatGPT desktop app for Linux. Use ChatGPT, ChatGPT Work, and Codex where you already work and build, with your projects and browser workflows on supported Linux systems. https://t.co/OtsPt5N5QC
View sourceArena-Hard-Auto official Gemini-2.5 judged score 50.0 with CI 0/0
View sourceSWE-Bench Verified resolved rate 42.4
View sourceAs models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research
ChatGPT can now remember your activity across the apps and websites on your computer. With Computer History in the desktop app, future interactions feel more personalized and require less explanation. https://t.co/WHZPxPp31R

Now in preview: The ChatGPT desktop app for Linux. Use ChatGPT, ChatGPT Work, and Codex where you already work and build, with your projects and browser workflows on supported Linux systems. https://t.co/OtsPt5N5QC
SWE-Bench Verified resolved rate 42.4
LiveCodeBench pass@1 70.6 across 1055 tasks
Introducing GPT-5.2 | OpenAI Skip to main content Research Products Business Developers Company Foundation (opens in a new window) Log in Try ChatGPT (opens in a new window) Research Products Business Developers Company Foundation (opens in a new window) Try ChatGPT (opens in a new window) Login OpenAI December 11, 2025 Product Release Introducing GPT‑5.2 The most advanced frontier model for professional work and long-running agents. Loading… Share Model performance Model per
Arena-Hard-Auto official Gemini-2.5 judged score 50.0 with CI 0/0