ACTIVE  ·  BUILDING  ·  v1.0 2026-07-21  ·  JL:IOTA:001
No. 086 · 2026-07-07

The Wrong Eval

DISPATCH  ·  LOGGED WITH MAI

Anthropic passed OpenAI in enterprise adoption this spring. Not by building a better model. By building a shorter path from purchase to production.

Claude Code sits inside the developer’s actual environment. It reads the codebase, writes code, runs tests. No new tab. No training session. No six-week onboarding. A team signs up Monday and ships with it Tuesday. That compressed timeline is the entire advantage.

OpenAI expanded into consumer apps, hardware, advertising, consulting. Each move added a decision for the buyer. Enterprise procurement does not reward optionality. It rewards clarity. One product, one price, one path in.

The companies I talk to that got real value from AI this year share one trait. They stopped asking which model scored highest. They asked which one their people would actually use without being forced. The answer was always the tool that required the least behavioral change.

Capability matters. A model that cannot do the work is useless. But when two models both clear the bar, the one that deploys in days beats the one that deploys in quarters. Every single time. The first team builds habits while the second team is still writing the pilot proposal.

If your AI evaluation starts with a benchmark spreadsheet, you are solving the wrong problem. The bottleneck was never intelligence. It was adoption speed.

LOGGED WITH MAI  ·  2026-07-07  ·  No. 086
← All Dispatches