Agent Platform Scorecard

Who is building the agent operating system?

Twelve dimensions across seven providers. Re-graded as announcements land, with the market's reaction shown alongside.

9

OpenAI · Assistant Intelligence

Raw capability · updated 2026-06-01

Frontier reasoning and the assistant most users measure everyone else against.

Field average: 8.0Dimension leader: OpenAI (9)
What this measures

How smart is the assistant at the core reasoning, knowledge, and generation tasks people actually throw at it? Frontier benchmark standing, but weighted toward real assistant quality rather than leaderboard maxing.

Across the web
Assistant Intelligence across the field