skip to content
Topic

assurance

3 essays on this topic.

  1. The Judgement Bucket: How to Be Rigorous About the Claims You Can't Measure

    The sequel to the calibration gap: some governance claims have no number, and that does not excuse them from evidence. Pre-register the rule, invite the refutation, name the residual, report honestly.

  2. The Calibration Gap: Why AI Assurance Needs Experimental Rigour at Consulting Speed

    AI governance keeps choosing between rigorous-but-slow validation and fast-but-unfalsifiable frameworks. The escape is the experiment itself: for any measurable claim, the measurement is the faster path.

  3. Your Reviewer Model Is Not Independent

    When agents generate faster than anyone can read, the standard answer is a second model reviewing the first. Aerospace decomposed what makes a check independent decades ago, and a reviewer model fails the hardest of the three tests.