Agent Platform Scorecard

Who is building the agent operating system?

Twelve dimensions across seven providers. Re-graded as announcements land, with the market's reaction shown alongside.

5+1

Apple · Agent Capability

Can it actually do work? · updated 2026-06-08

WWDC 2026 Siri understands what apps can do and 'gets more done', but true multi-step autonomy is still to be demonstrated.

Field average: 6.6Dimension leader: Anthropic (9)
What this measures

Can the system take a goal and complete multi-step work autonomously (browse, write code, operate tools, recover from errors), not just answer questions? Measures shipped agentic behavior, not demos.

What moved this grade
  1. Apple

    Siri, rebuilt: conversational and “profoundly more capable”

    Siri is now a profoundly more capable assistant, more conversational, so you can go back and forth like never before and get detailed, engaging answers.

    The personal-context Siri Apple promised back in 2024 finally ships in iOS 27, with a new, customizable expressive voice and a major jump in system-wide dictation accuracy. It is the most consequential Siri change in a decade, though the real-world quality only proves out as iOS 27 reaches devices this fall, so we move the grades with a verification caveat rather than to the top of the field.

    Assistant Intelligence46+2Conversational UX46+2Agent Capability45+1
    AAPL▲ +0.4%intradayA shrug, not a pop. The market had largely priced a Siri overhaul in.
    source: Engadget liveblog
More on Apple

No outside coverage tagged to this exact cell yet. Here is recent Apple analysis.

Agent Capability across the field