- 18 September 2026
The unit economics of the human-in-the-loop exception tail
In a production automation, cost is set by the small fraction of cases that route to a person, not by the volume the system handles cleanly. Here is how to quantify that tail from the same benchmarks that describe each process. - 28 August 2026
Idempotency and exactly-once guarantees when an agent writes into a system of record
When an agent posts an invoice or issues a payment, retries make duplicate writes a near-certainty unless the writes are designed to be safe to repeat. Here is the discipline that makes them correct. - 16 June 2026
What the agent-reliability curve says about which workflows are automatable now
A mid-2026 reading of the METR task-length curve, turned into a workflow-selection rule for operators deciding what to automate this quarter and what to wait on. - 4 June 2026
How production reliability gets engineered, and why a demo is not evidence of it
A pilot at eighty percent on a clean slice tells you almost nothing about whether the workflow runs unattended on Monday. The machinery that closes the gap, with named benchmark numbers.