
The Seven Sins of AI Agents
Agents don't fail at random: they fail with seven systematic vices that all survive the 'looks correct' test. The instinct is to add process — a PEP, a KEP, a committee — but every graduation stage rests on a human who approves it, and the agent's speed outruns the human you put at the gate. This is the honest comparison: what ontoref's ADRs inherit from PEP and KEP, and where they surpass both with witnessed, decidable, bounded-slice graduation criteria.

Trust Is an Output, Not an Input
An agent skipped the one invariant that would have caught the bug in thirty seconds. The honest diagnosis was not 'the agent forgot' — it was that the rule was prose, and prose never binds. This is the story of turning that failure into a falsifiable mechanism: a Statement of Work (the terms you own) and a Work Order (the execution it can't edit), where 'done' carries the validator's output instead of the agent's word.
2 items