Panout review packOverviewFinal memoMVPLive outputsJudge roundsAppendicesDecisions

Panout: the commit-boundary record for agent labor

Review pack, 2026-09-04. Five judge rounds, eight appendices of measurement, a shipped MVP, and its first live outputs.

One-liner. Panout records and enforces the decision to stop reading agent work. It runs a team's contracts at the commit boundary, records every override with its outcome, grants autonomy per contract from measured fault-injection catch rates, and is paid per evaluation, the same whether it says yes or no.

Where the score landed. Five independent judge rounds using the gstack office-hours rubric scored 62/72, 66, 67, 64, 59. The judge's stated ceiling without external evidence is about 65. Reaching 90 needs ten operators in writing, three external maintainers, a paying team, and a powered predictive measurement that cannot exist before roughly 2026-10-20. The loop was stopped there rather than manufacturing further rounds.

Final memo

The round-5 strategy memo: moat, $1B path, desperate human, wedge, pricing, and what the founder owes.

MVP

Goal prompt, what shipped in prototype/panout.py, acceptance status, and what remains user-only.

Live outputs

Real status, audit, inject, test run, and trailers from the founder's repos as of build time.

Judge rounds

All five memos and the fresh-judge verdicts that moved the idea from five candidates to one.

Appendices A to H

Every measurement, including the ones that disconfirmed earlier moats.

Decisions

The decision log, including the MVP departures from the goal prompt.

Read this first

Current dogfood status

Panout status
Contracts installed: 4 (tests-touch, no-secrets, size-guard, agent-attribution)
Evaluations recorded: 5 total, 5 in the last 7 days, 2 for commit(s) no longer on HEAD (amended or reset; kept as recorded)
Exposure: 5/5 commits with a non-skip result (100%; target 30%)
Overrides per contract:
- tests-touch: 2 override(s) over 4 evaluation(s)
- no-secrets: 0 override(s) over 5 evaluation(s)
- size-guard: 3 override(s) over 5 evaluation(s)
- agent-attribution: 0 override(s) over 5 evaluation(s)

Full outputs, including the fault-injection table, are on the live outputs page.