Spend is easy to total. Hard to locate.

The Money Map decomposes spend by step (operation, task class, model), so a rising bill has an address: a retry loop on one task class, rather than a total that went up. It shows cost per outcome first, then each step’s waste split between retries and cache, so the fix is aimed at the line that moved. A metric is real or it does not render.

Locating spend is not the same as ranking it. The map ranks by dollars, because dollars are the part we can measure; two steps burning the same amount can be worth very different things to your business, and nothing on this page knows which is which. Only you do.

See the map, at a buyer’s scale ↓
Our own books, located: the biggest line is 84% of the bill. Opened, the 11% line spent 63.7% of its full-budget arm on a knob, not the work.Figure 1

Every dollar our committed evidence cost us, ranked by step. Panel A is the estate’s tree from the previous part, read as a bill. Panel B opens one line, docpipeline-reasoning: the same 115 tasks re-run at thinking budget 128 landed more outcomes for $0.0352, so the rest of the full-budget arm’s bill bought nothing the outcomes needed.

A. The bill, ranked by step
margin-oss-seed (aider-edit, crewai-solve, crewai-pipeline-solve) · fit-scoring · gpt-4o$1.04
docpipeline-reasoning · reasoning-answer · gemini-2.5-flash$0.1322
highlightmagic-tape · sidekick-tagging · claude-sonnet-4-6 vs claude-haiku-4-5$0.0699
B. One line, opened
docpipeline-reasoningfull-budget arm · $0.0970
what the outcomes needed · $0.0352 default thinking budget · $0.0618

An address, not a total. The biggest line on our own bill is not where the fix was. The fix sat on the 11% line: a budget knob left at its default, and quality rose when it was cut. On your estate the waste hides elsewhere — retries, cold caches, a route nobody re-priced — but the move is the same: locate first, then act.

n = 794 metered calls, 3 workflowstotal $1.24quality 89% → 94% with the budget cutis_simulated=false in all threecaptured 2026-08-19 → 2026-09-19

Source: the committed artifacts provenance/{oss_cost_per_outcome_usd, autoroute_defended_savings_frac, reasoning_effort_savings_frac}.json. The ranking, both shares and the split derive from them at build time; none is typed.

The map itself, at a buyer’s scale.

Our own books are three workflows; yours will not be. This is the Money Map drawn on a stand-in company, simulated and labelled so, every system priced by the same code that priced Fig. 01. Each box is a system with its share of the bill drawn as a bar, and each opens to the operations inside it. Press the preview under the map and it re-prices off the first fix Margin would make. The numbers come from the server; nothing is computed on this page.

Next: The Parity Gate →

quality slips → it reverts → the loop re-runsThe Metermeasures every callThe Estatewhat you actually runThe Money Mapwhere the spend goesThe Parity Gateprove it held qualityThe Governoract, with revert armed
  1. The Meter: measures every call, its cost, tokens, substrate, and whether it worked.
  2. The Estate: the org chart of your AI workforce.
  3. The Money Map: where the spend goes, decomposed by step.
  4. The Parity Gate: proves the cheaper route held quality before it ships.
  5. The Governor: acts with a human ratifying, and a revert armed; if quality slips it reverts and the loop re-runs.
A cheaper route, once a person approves it, is held only while quality holds. Today that bar is one we set on your behalf.

Watch the loop run on real spend.

The console shows the loop on open-source agents we metered ourselves: measured spend, recorded face-offs and the gate’s verdicts, every number tagged real or sample. No live route has slipped yet, so the revert has not fired.