OpenAI GPT-5.4
A frontier OpenAI model release with strong reported performance across coding, tool use, multimodal, and academic reasoning benchmarks.
Structured model records with provider, status, capability tags, and update dates.
A frontier OpenAI model release with strong reported performance across coding, tool use, multimodal, and academic reasoning benchmarks.
A reasoning-focused GPT-5.2 configuration with detailed official benchmark tables across coding, academic, vision, and tool-use tasks.
Anthropic's Opus 4.6 release, positioned around long-running agentic coding, computer-use, and professional workflows.
Google DeepMind's Gemini 3.1 Pro preview, with official benchmark coverage across reasoning, coding, multimodal, tool-use, and long-context tasks.
A Gemini 3 Pro baseline included in Google DeepMind's Gemini 3.1 Pro comparison table.
A reasoning model from DeepSeek with official benchmark results published on the DeepSeek-R1 model card.
A widely used model family for general assistance, coding, multimodal analysis, and tool-using workflows.
A model family commonly evaluated for coding, long-form writing, analysis, and document-heavy workflows.
A multimodal model family with emphasis on broad input types and long-context workflows.
A model family frequently discussed for reasoning, math, and coding-oriented performance.
A broad model family spanning general, coding, multilingual, and multimodal use cases.