SYSE.50:4 - Solution
SYSE.50:4.1 - Select the decision and what the policy can observe
Name one decision unit: a particular claim, tool argument, action proposal or task continuation. Fix the performer or performing arrangement, relevant task family and information available at that point; identify the configured model when one is used. Distinguish an uncertain interpretation from a missing current fact, missing access, unresolved effect and absent authority. Further thinking can resolve some interpretation questions; it cannot manufacture an unavailable fact or authority.
List feasible responses and what each can supply. Retrieval may supply a source premise; clarification may settle the user’s intended target; an outcome lookup may resolve a previous effect. A generic request for “more confidence” supplies none of these by itself.
Choose signals obtainable before the decision. For a person these may include observed error patterns, the presence of needed information, current working demands and an elicited judgement of confidence. For a model they may include exposed response scores, sampled disagreement, evidence coverage or a trained estimator. Record the signal’s meaning and access conditions. A person’s confidence report and a model score have different qualification bases; text alone is not an internal probability. A hidden-state estimator is unavailable to a text-only API. Keep instruction and evidence access comparable.
SYSE.50:4.2 - Obtain evidence of useful and necessary assistance
Assemble representative decision inputs with independently qualified target answers or allowed continuations. Preserve the evidence available to the performer. Distinguish a contribution indispensable to the result from an aid that improves reliability or reduces burden despite adequate unaided ability. Record what the complete assisted way changes, including obtaining, checking and using its return. Another person’s or teacher model’s tool use does not establish this performer’s need.
Use paired conditions where they discriminate need: present versus withheld premise, ambiguous versus settled target, applicable versus revised rule. Obtain actual assistance returns where the policy relies on them. An oracle answer unavailable in deployment cannot qualify the deployed help branch.
Separate construction examples, calibration examples and final policy comparisons. Select enough variation for the consequence and reliance; do not infer population calibration from a few illustrative examples. SYSE.49 can construct the missing distinction and challenge the labels. A disputed label remains unresolved until its result meaning has grounds. When the target is human acquisition, use HCD’s practice purpose and permitted help: a hint and a complete answer can have different effects on the action being learned.
SYSE.50:4.3 - Construct the rule and its fallback
Relate the available signal to qualified outcomes, then select a decision rule under the receiving loss and resource constraints. For a threshold rule, compare supported continuation and assistance outcomes on calibration cases across candidate thresholds. Inspect unsupported continuation and unnecessary help separately, including the delay or cost of obtaining that help. A favorable average is insufficient when a protected error exceeds the declared allowance.
A set-valued construction is another option: score the allowed candidate actions and calibrate which remain plausible under the source method’s assumptions. One supported candidate can permit selection; several consequentially different candidates can trigger clarification. An empty or inapplicable set needs an explicit fallback. A probability or coverage claim requires the corresponding sampling, labeling and calibration basis; a hand-chosen cutoff is an engineering rule until such evidence exists.
Make the output implementable: response, target question or source, decisive grounds, applicability and stop condition. Keep required facts, permissions and effect recovery as independent conditions. No confidence value authorizes replay of an attempt whose effect remains unknown.
Define behavior outside qualification. A changed performer/model condition, inaccessible signal, unrecognized input or defeated calibration premise can return to an adequate direct rule, request the specific missing contribution, or stop with the exact gap. It must not silently become permission to guess.
SYSE.50:4.4 - Test the acted policy and return the right defect
Put the rule into the person’s practicable procedure or the technical controller through SYSE.47, and use SYSE.46 to compare whole receiving tasks with the incumbent or direct rule. Test both unsupported continuation and over-asking, plus confident common-source error, unavailable assistance and relevant drift. Calibrated scores alone do not establish that the selected help is timely, accurate or consumed.
Inspect the actual next input and transition when a supported premise is requested again. Return missing input to SYSE.52 and ignored return to SYSE.47. Repair the assistance rule here only when the decision uses the relevant input yet selects the wrong help behavior. Reopen calibration when a relied-on model, input, signal, source or task population changes. Stop at the bounded rule and evidence, or its exact qualification gap.