When Human Review Is Required
CoreDetermine when human review is required · Difficulty 2/5
Explanation
Not Every Output Needs a Human Gate -- But Some Clearly Do
Escalate an output to human review (and, when the task is technical, to an Architect or Developer) when any of the following apply:
- The output is high-stakes: legal, financial, medical, compliance, safety, or reputational.
- It will be published or sent externally without further checks.
- It makes factual claims that can't be easily verified by the user.
- It involves regulated or sensitive data, or a policy decision.
- The task exceeds the Associate's scope (complex system design, API/agent work) and should be escalated to a more technical role.
Human-in-the-Loop Is a Feature, Not a Failure
Requiring a human check before a high-stakes output goes out is not an admission that Claude or the process is broken. It is the expected, responsible design for any workflow where the cost of an undetected error is high. Treating human review as evidence of tool inadequacy misreads its purpose.
The Skill Is Calibration, Not a Universal Rule
The judgment call is matching the review level to the actual stakes of the specific output -- not applying one rule everywhere:
- Escalating everything wastes the tool's value and slows down low-stakes work that didn't need a human gate.
- Escalating nothing ignores risk and lets an unverified, high-stakes claim go out unchecked.
A team-lunch brainstorm and a customer-facing legal notice do not warrant the same review intensity, even though both are Claude outputs.
Common exam traps
- Thinking automation means removing humans everywhere; high-stakes outputs specifically warrant a human check regardless of how automated the rest of the workflow is.
- Escalating everything (wastes the tool's value) or nothing (ignores risk) -- the tested skill is matching the review level to the stakes, not picking one extreme.
Key Takeaways
- Escalate to human review when output is high-stakes, external-facing, contains unverifiable claims, involves sensitive/regulated data, or exceeds Associate scope
- Human-in-the-loop is a feature of responsible use, not a sign the tool failed
- The tested skill is calibrating review intensity to actual stakes, not a blanket policy
- Escalating everything wastes the tool's value; escalating nothing ignores risk -- both are wrong answers
Glossary Terms
Related Concepts