Panout review packOverviewFinal memoMVPLive outputsJudge roundsAppendicesDecisions

Panout: the commit-boundary record for agent labor

Review pack, updated 2026-09-04 evening. Six judge rounds, eleven appendices, a shipped MVP, and its first live outputs.

One-liner. Panout records and enforces the decision to stop reading agent work. It runs a team's contracts at the commit boundary, records every override with its outcome, grants autonomy per contract from measured fault-injection catch rates, and is paid per evaluation, the same whether it says yes or no.

Where the score landed. Six independent judge rounds using the gstack office-hours rubric scored 62/72, 66, 67, 64, 59, 61. The judge's stated ceiling without external evidence is about 65. Reaching 90 needs ten operators in writing, three external maintainers, a paying team, and a powered predictive measurement that cannot exist before roughly 2026-10-20. The loop was stopped there rather than manufacturing further rounds.

Final memo

The round-5 strategy memo: moat, $1B path, desperate human, wedge, pricing, and what the founder owes.

MVP

Goal prompt, what shipped in prototype/panout.py, acceptance status, and what remains user-only.

Live outputs

Real status, audit, inject, test run, and trailers from the founder's repos as of build time.

Judge rounds

All five memos and the fresh-judge verdicts that moved the idea from five candidates to one.

Appendices A to H

Every measurement, including the ones that disconfirmed earlier moats.

Decisions

The decision log, including the MVP departures from the goal prompt.

New on 2026-09-04: one more day of measurement

Read this first

Current dogfood status

Panout status
Contracts installed: 4 (tests-touch, no-secrets, size-guard, agent-attribution)
Evaluations recorded: 8 total, 8 in the last 7 days, 2 for commit(s) no longer on HEAD (amended or reset; kept as recorded)
Exposure: 8/8 commits with a non-skip result (100%; target 30%)
Overrides per contract:
- tests-touch: 2 override(s) over 4 evaluation(s)
- no-secrets: 0 override(s) over 8 evaluation(s)
- size-guard: 4 override(s) over 8 evaluation(s)
- agent-attribution: 0 override(s) over 8 evaluation(s)

Full outputs, including the fault-injection table, are on the live outputs page.