06-reference

every compound engineering guide

2026-02-09·reference·source: Every·by Kieran Klaassen (plugin co-maintained with Trevin Chow)

Compound Engineering: The Definitive Guide

Re-read at full length 2026-10-03 (paid).

Why this is in the vault

This is Every's canonical reference for the compound engineering loop (Plan, Work, Review, Compound), its plugin, a five-stage adoption ladder, and a set of team norms for agent-written code. It is the most complete external spec of a practice RDCO already runs, so it is useful as a checklist and as citable language. The earlier version of this note was a roughly 200-word teaser built from the guide's intro; this rewrite covers the full guide.

The core argument

Klaassen's thesis: each unit of engineering work should make the next unit easier. Most codebases go the other way. Every feature adds complexity, and after a decade teams spend more time negotiating with old code than building new code. Compound engineering reverses this by turning bug fixes, patterns and review findings into codified knowledge that the agent reads next time. Every says it runs five products (Cora, Monologue, Sparkle, Spiral, Every.to) mostly with one-person engineering teams on this system.

The main loop

Plan → Work → Review → Compound → Repeat. The first three steps are ordinary engineering. The fourth step is the differentiator; skip it and "you've done traditional engineering with AI assistance."

Two time rules. Per feature, plan plus review should take about 80% of the time and work plus compound about 20%. Across a developer's whole job, split 50/50 between shipping features and improving the system (review agents, documented patterns, test generators). Klaassen's arithmetic: one hour building a review agent saves about 10 hours of review over a year.

The plugin (house product)

The captured guide lists 26 agents, 23 commands and 13 skills. Key commands: /workflows:plan (parallel researchers plus a spec-flow analyzer), /workflows:review (14+ parallel reviewers such as security-sentinel, performance-oracle, data-integrity-guardian), /triage (human approve/skip per finding), /workflows:compound (six subagents write a searchable solution doc into docs/solutions/), and /lfg (the whole pipeline, 50+ agents, pausing only for plan approval). Note: [[2026-05-29-every-compound-engineering-upgrade]] covers a later version that grew the loop to more stages.

Beliefs to drop, beliefs to adopt

Drop eight beliefs: code must be hand-written; every line must be manually reviewed; solutions must come from the engineer; code is the primary artifact; writing code is the job; first attempts should be good (Klaassen puts first attempts at 95% garbage, second at 50%); code is self-expression; more typing means more learning.

Adopt: extract your taste into CLAUDE.md, agents, skills and commands; build safety nets instead of manual review ("if you don't trust the results, fix the system"); make the environment agent-native (the agent can run tests, read logs, take screenshots, open PRs); parallelize, because the bottleneck is now compute, not attention; and treat the plan as the new primary artifact.

Five-stage adoption ladder

0 manual; 1 chat-based copy-paste; 2 agentic tools with line-by-line approval (where most developers plateau); 3 plan-first, PR-only review (where compounding starts); 4 idea to PR on one machine; 5 parallel cloud execution with proactive agents. Each transition has a "compounding move": keep a prompt log (0→1), start CLAUDE.md (1→2), document what each plan missed (2→3), build a library of outcome-style instructions (3→4), document which work parallelizes and which is inherently serial (4→5). Skipping stages fails because trust has not been built.

Practical extras

Mapping against Ray Data Co

The most concrete connection is the copilot agent factory: its 50+ stateless skills, document-tracked state, eval plans and human review gates are this loop, minus an explicit Compound step with a "would the system catch this next time?" check. Ray has the parts ([[2026-04-04-nightly-learn-and-ship-loop]], /improve, /self-review, memory feedback files) but the capture is ad hoc rather than a required close-out of every run.

Why it matters for RDCO / The Denominator

⚠️ Sponsorship

The guide is house promotion for Every's free open-source Compound Engineering plugin, and it promotes Cora and other Every products. Every also sells AI consulting and runs Compound Engineering camps. Bias implications: claimed results (one-person teams, five to 10 times faster, 95% first-attempt garbage) are self-reported and unaudited, and the guide describes the plugin's own command set as the way to do the practice. The method stands without the plugin.

Related