Twelve dimensions across seven providers. Re-graded as announcements land, with the market's reaction shown alongside.
Raw capability · updated 2026-06-01
Frontier reasoning and the assistant most users measure everyone else against.
How smart is the assistant at the core reasoning, knowledge, and generation tasks people actually throw at it? Frontier benchmark standing, but weighted toward real assistant quality rather than leaderboard maxing.
A models/apps/harnesses framework: the harness that lets a model take real actions now matters as much as IQ.
Frames the rapid GPT-5.2 release as a competitive response to Gemini 3, with internal anxiety about losing the frontier.
OpenAI's launch of GPT-5 as a unified system that decides when to answer fast vs reason longer, priced to maximize reach.