06-reference

every gpt6 sol vs opus 5 5 vibe check

2026-09-22·reference·source: Every·by Dan Shipper
openaianthropicgpt-6-solopus-5-5model-comparisonevery-vibe-checkcodexclaude

Why this is in the vault

Every's same-day head-to-head of OpenAI's GPT-6 Sol and Anthropic's Claude Opus 5.5 — Dan Shipper's verdict is a split by task type, not a single winner, and it's the latest data point in Every's running Vibe Check series tracking which model wins which job.

The core argument

The newsletter itself was a paywalled preview (only the intro rendered — "Two big models dropped today," teasing the comparison without the verdict). The actual assessment came through cleanly in Shipper's companion X thread, which carries the same content: Sol is now his "new daily driver in Codex... faster, and 50% cheaper than 5.6 Sol," while Opus 5.5 was "the bigger surprise" of the two releases.

Shipper's per-category breakdown:

Overall framing: Sol is "an S-class iPhone release" — most of Astra's power at roughly a fifth of the price, best for people who live in Codex reading, writing, and shipping day to day. Opus 5.5 is the one worth handing a hard, open-ended coding or visual project to see how far it can run — strong enough that some of Every's Codex converts are wobbling back toward Claude.

Mapping against Ray Data Co

This directly validates RDCO's existing per-station model-effort pairing pattern (the skill-agent-brigade's station-code-author/sw-builder split, and the Delegation memory rule that Fable delegations run at high/xhigh effort for meaningfully large tasks only): the same segmentation Shipper describes — a fast/cheap model for high-volume daily read-write work, a higher-ceiling model reserved for long autonomous builds — is exactly the logic behind routing fanout-style extraction to a cheap low-effort model while reserving Fable-high/xhigh for real build tasks. It's also a live argument against picking one "house model" for everything: RDCO's own harness (Scribble Works' AI Gateway model routing, Claude Code station dispatch) is built around task-conditional model choice, not a single default, which this Vibe Check reinforces from the outside.

Related