← Research
Research · Definition

What is appropriate reliance?

The behaviour every oversight policy assumes, and nobody has worked out how to produce.

Last reviewed: 26 September 2026 · Next review due: 26 September 2027

Trust is the attitude, reliance is the behaviour, calibration is the match between them and the system. What has been tried, what it produced, and what that means for a policy that says a human will review the output.

Question this page partly answersAll 996 questions this research covers

Appropriate reliance means taking the machine's answer when it is right and overriding it when it is wrong, and after two decades of research nobody has an intervention that reliably produces it. That absence is the most important thing a leadership team can know about human oversight, because every oversight policy in circulation assumes the opposite.

The answer, in one line

Accepting automated advice when it is correct and rejecting it when it is not. Lee and See drew the distinction in 2004: trust is the attitude, reliance is the behaviour, and calibration is the match between them and the system's actual reliability.

Share as a card

Definition#

Appropriate reliance: the behaviour of accepting automated advice when it is correct and rejecting it when it is not. Trust is the attitude, reliance is the behaviour, and calibration is the match between the two and the system's actual reliability. The distinction is Lee and See's, from 2004.

Share this definition as a card

Why it is hard#

To rely appropriately, a person has to judge whether this particular output is right, which usually requires the capability the tool was brought in to supply. That is the circularity at the centre of human oversight, and it does not dissolve with training or with better interfaces.

The reliability of the system works against the human too. A tool that is right nine times in ten teaches the operator to accept the tenth, and the acceptance is rational given the base rate. This is the mechanism behind automation bias and complacency, reviewed by Parasuraman and Manzey in 2010, and it appears in novices and experts alike.

What has been tried, and what it produced#

Explanations. The intuitive fix, and the one regulation tends to assume. Bansal and colleagues, in a controlled study at CHI in 2021, found that explanations increased the rate at which people accepted the model's answer without improving the accuracy of the team, which is the definition of the wrong outcome.

Confidence scores. They help where the confidence is calibrated, and models are frequently confident in the wrong places, which moves the problem rather than solving it. The comparison this estate keeps is at calibration.

Cognitive forcing. Making the person commit to an answer before seeing the machine's. The most promising family, and the results remain mixed; it also costs the time saving that justified the tool.

Training. Warnings and instruction reduce automation bias somewhat and do not remove it, and the effect decays as the system proves reliable again.

What follows for oversight policy#

A policy that says a human will review the output is not a control until three things are true: the reviewer could produce or check the answer unaided, the review has time and authority to change the outcome, and somebody counts how often it does. The third is the one nobody instruments, and a review that never disagrees is not a review.

That is why this research treats override rate as the number worth watching. It is cheap to collect, hard to fake and it distinguishes an oversight arrangement that operates from one that exists on paper. The regulatory version of the same requirement is discussed at human in the loop is not a safeguard.

Key sources

Explainer · SS-2026-315 · Graded against the published rubric

Cite this page

Hirji, R. (2026). What is appropriate reliance?. The SuperSkills evidence base, SS-2026-315. https://thesuperskills.com/research/what-is-appropriate-reliance. Last reviewed 26 September 2026.

An evidence review by Rahim Hirji, not peer-reviewed research. For a material claim, cite the underlying study as well; every study here carries its own permanent link.

How citations and IDs work
Questions answered on this page

What is appropriate reliance?

Accepting automated advice when it is correct and rejecting it when it is not. Lee and See drew the distinction in 2004: trust is the attitude, reliance is the behaviour, and calibration is the match between them and the system's actual reliability.

Why is appropriate reliance difficult?

Because judging whether a particular output is right usually requires the capability the tool was brought in to supply. Reliability works against the person too: a tool that is right nine times in ten teaches the operator to accept the tenth, and that acceptance is rational given the base rate.

Do explanations help people rely appropriately?

The controlled evidence says no. Bansal and colleagues found at CHI in 2021 that explanations increased the rate at which people accepted the model's answer without improving the accuracy of the team, which is the wrong outcome rather than a partial one.

What does work?

Nothing reliably. Cognitive forcing, where the person commits to an answer before seeing the machine's, is the most promising family and the results are mixed; it also costs the time saving that justified the tool. Training reduces automation bias somewhat and does not remove it.

What should an oversight policy require instead?

Three things, all checkable: that the reviewer could produce or check the answer unaided, that the review has the time and the authority to change the outcome, and that somebody counts how often it does. The third is the one nobody instruments, and a review that never disagrees is not a review.

In this hub

Definitions

The terms this field uses, defined against their primary sources.

Ask the evidence
What does the evidence actually show?What should our board be asking about this?Where does Rahim disagree with the consensus?
Bring this into your organisation

If this describes something happening in your teams, say so.

Keynotes, board sessions and advisory work, drawing on research across more than 200 organisations in 30 countries. Tell me the room, the date and the shift you need. A reply within 24 hours.

Start a conversation

Topics and audiences  ·  All research

Box of Amazing

Rahim’s free weekly letter on AI and human capability

If this was useful, the weekly letter is where the thinking happens first. Most of what ends up on this site starts there. Weekly essays on AI, capability and the future of work. Read by 25,000 people, every week since 2017. Free, and one click to stop.

Opens Substack to confirm. No pitch in it, unsubscribe in one click, and nobody follows up because you read something.

Running an event, or responsible for how AI arrives in your organisation? Keynotes  ·  Advisory  ·  Boards  ·  Enquire