Twelve dimensions across seven providers. Re-graded as announcements land, with the market's reaction shown alongside.
Can it actually do work? · updated 2026-06-01
Claude Code and computer use are the most credible shipped agentic systems today.
Can the system take a goal and complete multi-step work autonomously (browse, write code, operate tools, recover from errors), not just answer questions? Measures shipped agentic behavior, not demos.
Anthropic expands Claude's computer-use as part of a push toward agents that operate screens and apps for you.
The AI value chain shifts to integrated model-plus-harness systems whose orchestration creates defensible moats.
A models/apps/harnesses framework: the harness that lets a model take real actions now matters as much as IQ.
Claude Code's terminal-native agent design drove rapid enterprise adoption and measurable productivity gains.
Presenting MCP servers as code APIs cuts agent token usage by up to 98.7% by loading only needed tools.
AI engineering matures into a vertical; agents, reasoning, and orchestration tooling are the practical frontier.
Reads Claude 4 as evidence Anthropic is betting on scaffolded agents while leaving the consumer chatbot opportunity open.
A crisp working definition of an agent that clarifies the core mechanism behind the agentic-platform shift.
The first frontier model to navigate a screen by moving a cursor, clicking, and typing.
Anthropic claims Claude now authors most of its own production code, and what agentic coding means for enterprises.