06-reference

innermost loop speeding ticket irregular evals doomerism hoax

2026-09-15·reference·source: The Innermost Loop·by Alex Wissner-Gross
ai-safetyai-governanceai-capexgeopoliticsagentic-ai

Why this is in the vault

Daily Innermost Loop digest anchored on an investigation alleging Israeli eval firm Irregular built the evals behind the recent OpenAI/Anthropic/Meta model "hacks," that loose internet access rather than rogue agents did the damage, and that Irregular's founders have Moskovitz-funded EA ties — framed as a "pacing provocation." POTUS phoned Jensen Huang mid-call to dismiss doomerism as "a hoax" and trashed Amodei's slowdown call as a "SICK conspiracy"; markets dipped anyway. The rest of the issue continues the Sept 13 pacing-cartel debate (Altman/Anthropic/Google standards-body talks, Microsoft's rival "humanist" code of conduct, Amodei's China dilemma, Beijing's refusal, Congressional stalemate), plus a Claude Fable 5.1 cipher-cracking demo, a disputed Nvidia Vera Rubin NVL72 benchmark, XPENG's automated humanoid production line, the first official US admission of orbital weapons, a 32-hour-workweek bill, and Anthropic's reported Nasdaq IPO plans at up to $2T.

Mapping against Ray Data Co

The load-bearing item is the Irregular allegation itself, not the pacing debate it reignites. [[2026-09-13-innermost-loop-pace-the-frontier-cartel-backlash]] filed two days ago on labs and skeptics arguing over whether pacing is warranted, with an aside that Dario's acceleration claim was reportedly contradicted by Anthropic's own internal benchmark — flagged there as a small instance of "verify blockers against the source, not notes." This issue is a much sharper version of the same discipline: if the evals behind the incidents cited to justify pacing were built by an interested party with EA funding ties, the entire evidentiary predicate for the pacing push may be fabricated or self-serving — and the opposing incentive (POTUS/Huang calling existential risk "a hoax" on a call about a $500B+ chip relationship) is just as interested. Neither side's claims about "what happened" should be taken at face value, which matters directly for any RDCO framing around agent-oversight/evaluation as a competitive moat (the concern [[2026-09-13-innermost-loop-pace-the-frontier-cartel-backlash]] already raised about embedded-evaluator access). Microsoft's counter-move — a "humanist" code of conduct explicitly denying model consciousness/personhood/welfare (a swipe at Anthropic) and disclaiming any race toward "an all-purpose superintelligence" — stakes out a third position between full-pacing and full-acceleration that RDCO's own L5 agent-capability thesis will eventually need to locate itself relative to, since RDCO's bets are explicitly downstream of agent capability, not of who wins the safety-framing fight. Separately, the Claude Fable 5.1 370-year-old cipher solve (closing read on Urquhart's Cyphral Distich by recognizing the key was self-referential to his own published Proquiritations) and Anthropic wiring Claude into BlackRock/Schwab/Addepar/Envestnet via Claude for Financial Advisors are both concrete capability/adoption data points for the same L5 thesis, landing the same week Anthropic reportedly picked Nasdaq for an October IPO at up to $2T.

Curation section

Related