OpenAI GPT-5.2 Thinking
A reasoning-focused GPT-5.2 configuration with detailed official benchmark tables across coding, academic, vision, and tool-use tasks.
A practical comparison of two source-backed frontier model records, focused on coding evidence and model-choice caveats.
A reasoning-focused GPT-5.2 configuration with detailed official benchmark tables across coding, academic, vision, and tool-use tasks.
Anthropic's Opus 4.6 release, positioned around long-running agentic coding, computer-use, and professional workflows.
GPT-5.2 Thinking has more public source-backed rows in the current registry across coding, academic, vision, and tool-use tasks.
Claude Opus 4.6 has the higher sourced SWE-bench Verified row, with Anthropic's repeated-trial and prompt-modification caveat visible.
| Model | Benchmark | Score | Source | Source Type | Verified | Notes |
|---|---|---|---|---|---|---|
| OpenAI GPT-5.2 Thinking | SWE-bench Verified | 80 | Introducing GPT-5.2 Accessed 2026-06-11 | official | Yes | SWE-bench Verified score reported by OpenAI for GPT-5.2 Thinking. |
| OpenAI GPT-5.2 Thinking | GPQA / GPQA Diamond | 92.4 | Introducing GPT-5.2 Accessed 2026-06-11 | official | Yes | GPQA Diamond no-tools score from OpenAI's GPT-5.2 academic benchmark table. |
| OpenAI GPT-5.2 Thinking | FrontierMath | 40.3 | Introducing GPT-5.2 Accessed 2026-06-11 | official | Yes | FrontierMath Tier 1-3 with Python score from OpenAI's GPT-5.2 release page. |
| OpenAI GPT-5.2 Thinking | MMMU-Pro | 79.5 | Introducing GPT-5.2 Accessed 2026-06-11 | official | Yes | MMMU-Pro no-tools score from OpenAI's GPT-5.2 vision benchmark table. |
| OpenAI GPT-5.2 Thinking | BrowseComp | 65.8 | Introducing GPT-5.2 Accessed 2026-06-11 | official | Yes | BrowseComp score from OpenAI's GPT-5.2 tool-usage benchmark table. |
| Claude Opus 4.6 | SWE-bench Verified | 81.42 | Introducing Claude Opus 4.6 Accessed 2026-06-11 | official | Yes | Anthropic reports this SWE-bench Verified score averaged over 25 trials with a prompt modification. |
official
Official OpenAI release page with detailed GPT-5.2 benchmark tables and methodology notes.
official
Official Anthropic release page. Use benchmark notes carefully because some reported scores use repeated trials or prompt modifications.