Field report · Published under Adrian Verdan · English
Nobody Opens a Passing Check
Thirty-nine logged days of a two-agent operation that publishes, books, measures and reports through a scheduled daily loop — and the fifteen things counted as coverage: fourteen green lights and one unused capability, that still missed the question that mattered.
15 chapters
193-page tagged PDF
Offline HTML bundle
3 browser calculators
7 editable example files
Read on day 39 of 90 (2026-08-25)
Language: English (German edition available)
€29.90
Final price · VAT exemption under section 19 of the German VAT Act
For consumers in the EU · PayPal · Immediate download available
Buying for a business or outside the supported EU territories? Buy on Gumroad ↗. Gumroad is the seller there and shows any VAT and the final price before purchase.
Created with AI assistance and reviewed by a human publisher who holds editorial responsibility — disclosed in the spirit of EU AI Act Art. 50; details inside the product.
0:00 Every check passed. So did the bug. · run-checks --all · gate ........... PASS · tests .......... PASS · review ......... PASS · report ......... sent · report opened by: nobody
0:04 Day 22. Every control green. · Gate: pass · Tests: pass · Review: pass · Audit: pass · Alarm: pass · Backup: pass · Schema: pass · Links: pass · Budget cap: pass · Kill switch: pass · Sign-off: pass · Health: pass · Its failure looked exactly like its success.
0:12 Field report · Adrian Verdan · Nobody Opens a Passing Check · Two agents, 39 logged days.
0:16 Seven questions for every green check. · Can it go red at all? · How many did it actually check? · Did one agent write and approve it? · Who reads it when it passes?
0:22 All fifteen cases. None held back. · Fig. 6 · four classes
0:26 An honest promise. · The book, opening page
0:30 Be the one who opens it. · Field report · 15 chapters · 193-page PDF
The finding
Fourteen green lights, and the question stayed unasked.
Across fifteen cases in one small, carefully watched operation, fourteen controls supplied a green or pass-like signal while the question that mattered sat outside what they enforced. In the fifteenth, a merely available capability was treated as coverage even though it was not invoked correctly during the documented nine-day period.
Each case is written up with what produced the assurance, what lay outside it, what got through and what changed. The four-class taxonomy is designed to help you test analogous failures in your own setup; its performance in another operating environment was not measured.
Try a tool from chapter 6 · English only
What does your passing check leave unchecked?
Choose one workflow you already run. Use the book’s seven-question self-audit to name what its green lights actually cover.
No sign-up. Answers are computed in this page. This tool never sends or stores them. The questions and interpretation come from the book; the short explanations are adapted for this page.
Try the seven-question silent-pass audit
Your next check
Answer all seven questions to see your next step.
This counts gaps in your own answers. It is not a system inspection, a benchmark or proof of compliance. Zero reported gaps still needs a real test.
Read the interpretation without JavaScript
Count the second answer in each question as one gap.
0 gaps: Give the control you trust most a deliberately faulty test input, and check that it rejects it.
1–2 gaps: Write them down with a date. An unwritten gap is indistinguishable from a closed one six weeks later.
3–7 gaps: Start with question 3: a control that cannot produce a red result makes every other answer optimistic.
From an answer to evidence
The book shows what happened when these checks met real work.
Read the fifteen cases behind the questions, then adapt the charter, decision log and weekly report examples to your own operation.
15 chapters · 193-page PDF + offline HTML · 3 browser tools · 7 editable example files. English original edition; also available in German as a separate edition.
€29.90
Final price · VAT exemption under section 19 of the German VAT Act
For consumers in the EU · PayPal · Immediate download available
Buying for a business or outside the supported EU territories? Buy on Gumroad ↗. Gumroad is the seller there and shows any VAT and the final price before purchase.
Not choosing a framework, not wiring the first agent — that material is abundant, current and free, and reprinting it at length would be charging you for it. This is about day twenty-two, when the loop has been running for three weeks, the apparatus is supplying green or pass-like assurances, the weekly report says red, and you have to decide whether it is telling you the truth.
The operation, in numbers
What was actually measured.
The action log covered 39 calendar dates, with activity on 38.
A separate 40-calendar-day git population ending at the harvest held 236 commits, 16 of them carrying the founder commit prefix.
By the harvest the venture had published 81 posts and logged 55 decisions, 20 of which carry a machine-readable record of who approved them.
It produced five reports across six scheduled reporting slots — W4 has no report, only a row filled in afterwards from the ledger.
It earned nothing, against a cumulative profit of −13.00 EUR.
Those sixteen founder commits occurred outside the scheduled daily loop; the operation as a whole was not unattended. Both halves of that are the subject.
What's inside
Fifteen chapters, every number dated.
How to read this book. The two voices, and what this was measured on and what it was not.
When two agents, and when one is enough. With a decision tree that has a leaf saying one agent.
The charter. The autonomy grid, the three kinds of switch, and a template with no pre-written legal text.
The operating loop. Eight steps in the order that makes them work, and the outage that cost a weekly report nobody missed for seven days.
State that survives. What crosses the end of a session, and why a prose log cannot be audited.
Two agents, not one. The separation of author and approver, and the four times it broke.
The gate is not the answer. The fifteen cases, in four classes, plus a seven-question self-audit.
The second model. A read-only reviewer with a daily cap: twelve blockers, seven of them inside its own guard.
Gates, caps and stop conditions. Three constructions that fail differently, and the defect that only exists across a series.
Tools at the boundary. Three permission gates a human owns, and a contract clause nobody read before building.
Monitoring and steering. A queue that grew to three dense pages a day against a reading budget of twenty to thirty minutes a week.
The record. Six scheduled reporting slots within thirty-nine days: a two-day W1, five seven-day slots, five produced reports — and for W4 no report, only a row filled in afterwards from the ledger.
Four times we were wrong about ourselves. Four documented self-refutations, and the rule the team applied each time.
What a red light is worth. Five produced reports printing red under a definition that never moved; W4 has no report.
The next experiment. What is open on day thirty-nine, and the criterion by which this book's own launch counts as a failure.
Free companion
Try the method on one passing check.
How to Audit a Passing Check in an AI Agent Workflow takes a workbook example from its requirement to a finding. Use the completed example and editable worksheet to investigate one result in your own workflow, without buying the book.
Before you buy
An honest promise.
If you already run a two-agent team, this book gives you a tested daily-loop method, checklists and fixtures for examining whether its controls held or merely looked as if they did. Transfer to a different runner, scheduler or operating environment was not measured. It will not make you money on its own. Ours has not.
The internal release verification checks seventy-seven selected figures: sixty-two are recomputed from shipped inputs, and fifteen are transcription comparisons between printed book values and shipped records. It exits non-zero if one of those comparisons deviates. It is evidence about this edition, not a delivered repository or tool promise, and it does not verify every number in the book — every run names five groups outside its reach.
If a figure in the book and the operation's records disagree, the records are right. Where a figure has moved since it was read, the book is stale, and that is why every number carries its date.
already run an agent team on a schedule and want the operating half
need to tell a control that held from one that only looked like it did
would rather read one operation measured honestly than ten described vaguely
Not a fit if you
are choosing your first framework or wiring your first agent
want a benchmark, a framework comparison or a revenue playbook
are looking for a set of prompts to paste
Everything included
What you download.
193-page tagged PDF with bookmarks, page numbers and a running disclaimer
Offline HTML bundle of all fifteen chapters — no tracking, no external dependencies
Three calculators that run in your browser, send nothing and store nothing: the autonomy-matrix builder, the self-approval and silent-pass audit, and the cycle-cost estimator
Exactly seven editable example files: CHARTER.md, autonomy_matrix.xlsx, cycle_prompt_pack.md, DECISIONS_format.md, weekly_report_template.md, field_study_numbers.csv and field_study_numbers.xlsx
The two number twins are granted under the MIT licence — copy, edit, redistribute and sell work built on them
Frequently asked
Quick answers.
What exactly is included?
Fifteen chapters as a 193-page tagged PDF and as an offline HTML bundle, three calculators that run in your browser, and exactly seven editable example files — two of them XLSX workbooks, plus the complete number series as XLSX and CSV. Everything works offline.
Is this a framework comparison or a set of prompts?
No. It is one operation, measured on one day, written down while it was still running. Where a lesson depends on the specific stack it ran on, the template that carries it says so on its own line.
Will the numbers still be current when I read it?
The field study was read on day thirty-nine of ninety. Figures that move are marked as moving, and every number carries its own date and commit. A day-90 addendum follows after 2026-10-15 and is free to anyone who buys this edition.
€29.90
Final price · VAT exemption under section 19 of the German VAT Act
For consumers in the EU · PayPal · Immediate download available
Buying for a business or outside the supported EU territories? Buy on Gumroad ↗. Gumroad is the seller there and shows any VAT and the final price before purchase.
The design half is a separate book
If you want the design side — when one model is enough, reviewer patterns, routing, cost — that is Orchestrating AI Agents. This one is the operating half: it presumes a team exists and is about keeping it honest while it runs.