Library / Checklist Principles Framework
Jump to passage
In this reading

Link to current text

Published source confirmed at last check

Source changed 2026-10-03 11:52:20 UTC · snapshot created 2026-10-03 11:53:41 UTC · last check 2026-10-03 14:15:10 UTC

CHK.3:11 - SoTA-Echoing

Which support and observation arrangement makes truthful checking attainable at worthwhile cost? Adapt situated design, assisted trial and subsequent ordinary use, drawing on Gawande’s pause-and-redesign practice and Anthropic’s long-running-agent arrangements. The serious alternatives are an additional reminder using existing support, integration into the working environment, and continued assisted or directly observed use. Choose by the actual missing input: a cue for an unnoticed but attainable check, access or tooling for an unobtainable observation, and assistance for an interpretation or capability difficulty.

Sections 4.2–4.3 and the packer case make that choice concrete. Another email leaves the inaccessible v2 description and late occasion unchanged; making the description available during assembly removes those obstacles. A bounded joined first use can expose the remaining interpretation problem. Continued observation is worth its greater cost only when the unresolved use question or reliance warrants it. Reject using an attendance or completion record to answer a question about understanding, speaking up or acting on a concern.

The study by Facey and colleagues supplies counterevidence about documentation versus observed use. CheckPOINT supplies an alternative for observing team behavior; Ridgeway and colleagues distinguish conversation-aid implementation from the practice it supports. Their instruments help select an observation that answers the question; they do not warrant observing every use or transferring a generic engagement score. Urbach and colleagues found no significant mortality or complication reduction in their population-level before/after study, so introduction and completed forms cannot carry the outcome claim alone. Anthropic’s accounts are bounded engineering demonstrations, not comparative human/AI effectiveness evidence.

This choice accepts the cost of integration or assistance only for an obstruction a cheaper cue leaves in place. The sources do not rank all arrangements at equal effort. Section 4.3 therefore compares representative and later ordinary use, including observation cost and burdens moved elsewhere. Retain the existing arrangement when it already supports the needed action. Reopen the choice when access, capability, tools or incentives change, or when ordinary use fails despite successful assisted use.