Block criteria in practical terms
Block when key claims are unsupported and users are likely to take action from them.
Block when errors can cause compliance, legal, financial, or safety harm.
Block also when the answer overstates certainty in areas where even one bad instruction could trigger a serious downstream failure.
Review versus block
Use review for ambiguous but recoverable cases where a human can quickly validate and fix.
Use block for high-impact uncertainty, fabricated references, or strong confidence with weak support.
The dividing line is whether the response can be salvaged safely with targeted review or whether it should be stopped before anyone sees it.
Why this improves user trust
Blocking high-risk output prevents visible failures that reduce confidence in your product.
It also keeps your review workload focused on cases that can be safely salvaged.
That makes the whole system easier to operate, because teams are not wasting time debating clearly unsafe output.