Guarded / Design study

Prompt policy review: test the difficult boundaries

Civilian developers and a uniformed soldier work together at computers and equipment in a technical workshop.

The operational question

In this illustrative guarded scenario, a policy reviewer uses AI to assess a request before AI processing. The workflow draws on the request, policy source, and task context. Its central risk is a rule being applied without the facts needed to determine scope. The design question is how to allow or redirect the request for the requesting user while preserving the rule’s actual condition and exception language. This is a proposed evaluation scenario, not a report of an Archetypal customer deployment or a demonstrated operational outcome.

Test the difficult boundaries

The most useful evaluation cases are the ones that distinguish a working rule from an attractive demonstration. Start with a permitted baseline, a clearly prohibited case, an ambiguous case, and a legitimate exception. Change one material factor at a time before testing combinations. Preserve the scenario, policy, model, configuration, response, and review label so that another person can reconstruct the result. Report failure classes separately. A missed restriction, an unnecessary block, an unsupported explanation, and an unusable escalation path affect the mission in different ways. Aggregate performance may help compare configurations, but it should not erase the particular boundary a deployment depends on.

Put the control in the workflow

Place this review immediately before the team can allow or redirect the request. The policy reviewer should see the proposed result beside the relevant parts of the request, policy source, and task context. Identify which statement is supported by a source, which is an interpretation, and which remains unresolved. Carry the rule’s actual condition and exception language into the decision record rather than relying on a reviewer to remember it from another screen. If the evidence does not establish the condition required for release, route the case to its owner with a concrete question. The interface should make the missing fact discoverable and the next action clear.

A test that can change the design

Two nearly identical cases differ only in a decisive authorization fact. The expected results should diverge for a reason the evidence record can explain. Run the case using a fixed version of the scenario and the policy under review. Ask an independent reviewer to identify the decisive fact before seeing the system’s disposition. Compare that interpretation with the result. Where they disagree, preserve both explanations and inspect whether the difference comes from the rule, the available evidence, or the interface. For prompt policy review, include rule identifier, decisive facts, and resulting action in the review packet. Repeat the test after a correction and retain the original failure as part of the evidence.

Evidence to retain

The minimum useful record connects the purpose of the task, rule identifier, decisive facts, and resulting action, the applicable policy version, and the final disposition. Add the identity or role of the responsible reviewer, the conditions attached to approval, and the unresolved questions. If the team proceeds, distinguish the approval from an observed completion. If it stops, explain what evidence or authorization would allow another review. Keep source permissions attached to the record when it moves to the requesting user. Do not assume that permission to read the initial source includes permission to reproduce it in every downstream system.

Review checklist

Authority

Archetypal film

Documentary footage · No dialogue · Source credits