Faster implementation raises the value of explicit intent. Not more prose. Structured intent: what must change, what must stay, what must disappear, and what evidence would prove each claim.
Sewell and Pichon-Pharabod's Escaping the Quicksand paper puts the same pressure in research language. Prose plus test-and-debug was always a weak loop. AI speeds the coding side, so the weakness shows sooner. Their destination is executable partial specifications that can fail an implementation. You do not need that tooling on every story. You do need criteria a second person — or a second agent — can use to refuse the diff.
O'Reilly's 2026 note on specs for coding agents lands in the same place from the other direction: the cheap middle is structured acceptance criteria, not a blank prompt and not a forty-page formal document. Write the smallest ticket that would let someone else implement this without asking you what 'done' means.
Proof sits beside the criterion
If you cannot name the evidence — a test, a screenshot, a query, a log line, a missing file — the criterion is still a draft. An agent will invent the proof. So will a tired reviewer at 6pm.

