Why this is in the vault
Lead story is another confirmed data point in the recurring AlphaSignal thread on multi-agent swarms cracking formally-verified math/proof problems — this time OpenAI running 10,000 agents for 88 hours to resolve an open question about Navier-Stokes blowup, Lean-checked for correctness.
Mapping against Ray Data Co
Directly reinforces the fan-out-and-verify pattern RDCO already runs operationally (deep-research's one-sub-agent-per-question isolation, family-research-round's Workflow fleet, process-newsletter's own per-message sub-agent dispatch): OpenAI's approach — thousands of agents working sub-problems in parallel, with a Lean formal-verification pass as the trust gate before the result is accepted — is the same "fan out cheap workers, gate the output through an independent, mechanical verifier" shape RDCO uses (station-critic, audit-newsletter-outputs.py, graph-ingest). The scale difference (10,000 agents vs RDCO's single-digit sub-agent fan-outs) is the interesting delta: it's a proof point that the pattern keeps working as N grows, as long as the verification step stays formal/mechanical rather than another LLM judging its own swarm's work — the same principle behind RDCO's insistence on a deterministic audit script rather than an LLM self-grading its vault notes.
Secondary: the Mistral €3B raise (French sovereign/self-hosted-model pitch, no vendor lock-in) is a data point for L5 positioning conversations about model-provider concentration risk, but doesn't cross a threshold to change anything here — noted for completeness, not actioned.
⚠️ Sponsorship
Four distinct paid placements in this issue, none overlapping with the previously-confirmed 12-sponsor rotating pool (OpenRouter, Teleport, Granola, Unblocked, Datadog, Vanta, Finest, Google Cloud, Tiger Data, Redis, Attio, Flint AI):
- QA.tech — masthead "In Partnership with" slot (webinar promo: Vilhelm von Ehrenheim, Co-founder & Chief AI Officer, "Your AI Agent Wrote the Code. Who Verifies It?", Sept 16 webinar on agentic PR verification). This is the first issue where the previously-unresolved masthead "In Partnership with" slot (flagged unresolved across 2026-09-03, 09-04, 09-07, 09-08) actually resolves to a named, legible sponsor in plaintext — it rendered as a full webinar-promo block this time rather than an unlabeled image logo. Worth confirming on the next issue whether this is now a stable plaintext-legible format or a one-off.
- Ory — standalone "Presented by" block, "Deterministic Controls for AI Agents" (Ory Agent Security — policy checks inside the agent runtime).
- Launch Darkly — standalone "Presented by" block, "AgentControl" (runtime behavior control/rollback for production agents).
- Voices — native ad embedded inside the Signals numbered list (#2), DNSMOS speech-data calibration framework.
Rotating pool now 16+ confirmed distinct sponsors across 9 tracked days. Bias implication: none of the four sponsors this issue touch the lead story's actual claims (multi-agent math-proof swarms) — the sponsor content is adjacent-market (agent security/observability/verification tooling) rather than a direct promotional angle on the news being reported, consistent with the pattern established across prior issues.
Curation section
Top News
- OpenAI agents crack a 90-year-old math problem in 88 hours — 10,000 agents on an unreleased (more-capable-than-GPT-6-Astra) model resolved an open question on Navier-Stokes blowup; result formally checked in Lean. Newsletter flags "controversy around credit" as a secondary story without detail. This is the lead/mapped item above.
- OpenAI ships ChatGPT Images 2.5 — two new API models (Flare, Sunburst), 50% lower latency than GPT-Image-2 at comparable/better quality, sketch-input feature, priced 2x GPT-Image-2 rate. No RDCO mapping — image-gen product update, not agent-orchestration relevant.
- Mistral raises €3B Series D at €21B valuation, largest European tech equity round ever; self-hosted/no-lock-in model pitch, 1GW EU compute buildout by 2030, 125+ enterprise customers. Noted above, not actioned.
Signals
- Google DeepMind mapped predicted impact of all 9 billion possible DNA mutations.
- Voices' DNSMOS speech-data calibration framework (sponsored native placement, see above).
- Sony AI's open-source tool predicted two real aging-gene discoveries from 1.5M hypotheses.
- Reasoning models produce fractal patterns when solving hard problems — offered as a partial explanation for "overthinking."
- OpenBMB released MiniCPM5-2B, a 2B-parameter on-device model.
- An open-source tool lets 20,000 developers run Claude Code free via DeepSeek — worth a skeptical read (free-tier routing through a third-party model backend) but not deep-fetched this issue; no RDCO action implied without verifying the mechanism.
No deep-fetches this issue: the plaintext body strips outbound hyperlinks (only "READ MORE" anchor text survives, no resolvable URL), so none of the curated items met the "third-party domain, clear hook" bar for a fetch — flagging per the curation-mode rule rather than silently passing.
Related
- [[2026-08-11-alphasignal-riemann-hypothesis-subagent-swarm]]
- [[2026-07-13-alphasignal-subagents-math-proof-cycle-cover]]
- [[2026-09-07-alphasignal-lean-proof-swarm-cheat-whistleblower]]