Every release makes the harness harder to fool: LLMKube 0.9.19
We found nine fully-tested rails on main with zero production callers, reviews approved by an agent running on an empty system prompt, and safety rails that silently skipped when their inputs went missing. 0.9.19 is what shipped after we pointed the harness's own skepticism at itself: a CI guard that makes tested-but-dead code impossible, verdict records that name what was not checked, a config floor for agents, and an execution-first reviewer rubric. Much of it was written by the agent fleet through the pipeline it improves.