06-reference

innermost loop opus55 life sciences swarm superintelligence

2026-09-24·reference·source: The Innermost Loop·by Alex Wissner-Gross
ai-capabilityagentic-aiai-governancedecision-infrastructurelife-sciences-ai

Why this is in the vault

Daily Innermost Loop digest without a single named framing device this issue — it runs paragraph-by-paragraph through Anthropic's Opus 5.5 release (beats Fable 5.1 on most work at 40% less cost than Opus 5, debuts #1 on Artificial Analysis), OpenAI's cheaper GPT-6 Sol/Luna and Astra's autonomous-driving debut (with a "nail in the coffin for specialized models" claim from OpenAI's Boris Power), a ~950-agent Anthropic life-sciences lab, and a governance paragraph where POTUS rebrands AI as "super intelligence" against a 22-country human-control declaration and a Sanders/Casar ban bill.

Mapping against Ray Data Co

The load-bearing item is Boris Power's "nail in the coffin for specialized models" claim, made after GPT-6 Astra alone finished DrivingBench's cone course in a real Corolla (5:22, $7.74) where Fable 5.1 only managed 45% — it's a direct counter-signal to the live due-diligence thread RDCO has open on TypeSafe's Jev, a specialized "System One" classification model already in RDCO's stack ([[2026-09-15-every-typesafe-jev-vibe-check]], [[2026-09-20-innermost-loop-claude-rd-share-jev-embedded-evaluators]], [[2026-09-23-every-jev-usage-guide]]). The claim's scope should be read carefully before treating it as a verdict on Jev specifically: DrivingBench is an embodied, physical-world task where a frontier generalist absorbing driving as one more skill is a different bet than Jev's narrow-and-cheap classification niche (4.2¢/million input tokens vs. $10 for Astra/Fable input) — but "generalist subsumes specialist" is exactly the industry-level thesis RDCO's Jev due-diligence beat needs to keep testing against, not just accumulate confirming evidence for. Separately, Anthropic's ~950-Claude-agent life-sciences lab (combing 200,000 reverse transcriptases for a new CRISPR-like enzyme system) is a concrete look at what "agent fleet in production" actually means at scale in a closed, well-specified domain — a useful outside data point for RDCO's own fleet-dispatch pattern (one-subagent-per-item in /deep-research and /family-research-round) as those patterns scale past a handful of parallel agents. Opus 5.5's headline (leads coding/knowledge-work/computer-use at 40% below Opus 5's cost) is also a fresh data point worth checking against RDCO's model/effort delegation heuristic ([[feedback_delegation_model_effort_pairing]]), which is still informal.

Curation section

Attempted one deep-fetch on the Boris Power "nail in the coffin" item since it anchors the mapping section — the substack.com/redirect link resolves to an x.com/BorisMPower status, but X returned HTTP 402 to WebFetch, so the note relies on the newsletter's own paraphrase rather than the primary tweet. Every other link in the issue is the same substack.com/redirect tracking wrapper; none of the remaining ~25 items had a specific enough hook to justify the second deep-fetch budget.

Related

[[2026-09-23-every-jev-usage-guide]] [[2026-09-20-innermost-loop-claude-rd-share-jev-embedded-evaluators]] [[2026-09-15-every-typesafe-jev-vibe-check]] [[2026-09-22-innermost-loop-precautionary-scarcity-osec-gpt6-sol]] [[feedback_delegation_model_effort_pairing]]