← Research
Research

Should I let an AI agent act on my behalf?

An agent without a stated boundary is not delegation. It is abdication with a progress bar.

Last reviewed: 26 August 2026

The question that replaces do I trust it, four things to fix before granting autonomy, why the evidence points at the front rather than the back, and what is genuinely unknown.

Questions this page answersAll 780 questions this research covers

Sometimes, and trust has little to do with it. The question is whether you have specified the boundary, because an agent without a stated boundary is abdication with a progress bar.

The distinction that matters is simple and almost nobody makes it. A system that answers gives you something to accept or reject. A system that acts has already done it. Every oversight model in common use assumes a pause for review, and agents remove the pause.

The question that replaces "do I trust it?"#

What is the worst thing this can do before a human sees it, and can I live with that? Answer that and the trust question resolves itself. Leave it unanswered and no amount of confidence in the model helps you, because you have not bounded the downside.

Four things to fix before granting autonomy#

1 · Reversibility. Sort the actions the agent can take into reversible, expensive to reverse, and irreversible. Grant autonomy freely in the first category, reluctantly in the second, and not at all in the third without a stop. This does more work than any accuracy estimate, because it bounds the loss rather than the probability.

2 · Blast radius. Not what the agent does, but how far the consequence travels. Sending one email is small. Sending one email to a client list is not, and the action is identical.

3 · The stop. Can you halt it mid-sequence, and does halting leave things in a safe state or a broken one? Article 14 of the EU AI Act requires exactly this for high-risk systems: intervention or interruption bringing the system to a halt in a safe state. Most consumer and internal agent deployments have not been designed for it.

4 · The reconstruction test. If this goes wrong, can you reconstruct why the agent did what it did? If the reasoning exists only inside a sequence of model calls, it has already gone, and you will be explaining an outcome you cannot account for.

Why the evidence points at the front rather than the back#

The Vaccaro meta-analysis of 106 experiments found human-AI combinations underperforming the better party alone, with losses concentrated in decision tasks, where a human judges whether a system was right, and gains in creation tasks, where the pair produces something together.

Agents make the decision task worse in the one way that matters: they perform it at machine speed, in volume, without the human present. If review after the fact was already the weakest available position, reviewing after the fact and after execution is weaker still.

Which means the human contribution has to move to where it counts: the problem, the intent, the constraints and the rejection criteria, all before anything runs. That is Human at the Start, and with agents it stops being a preference and becomes the only place the human can meaningfully be.

What is genuinely uncertain#

Most of it. There is no equivalent of the Vaccaro analysis for agentic systems, no field evidence on agent oversight failures at scale, and no established practice for what adequate autonomy limits look like. Agent capability is also moving faster than any other part of this field, so anything written now dates quickly.

Anyone offering confident guidance on agent delegation, including this page, is reasoning from adjacent evidence. The difference is whether they say so.

Agents make the argument unavoidable#

Agents are the point at which the argument this research has been making becomes unavoidable rather than advisory. When a system answers, you can compensate for a weak oversight design by being careful at the end. When a system acts, there is no end to be careful at. The decision was made when you set the boundary, or it was not made at all.

There is a second consequence, quieter and worse. Agents absorb exactly the sequences of small tasks that used to constitute learning a job: chasing the thing, checking the thing, noticing the anomaly, following it up. Those look like overhead and they are where judgement is formed. An organisation that hands them wholesale to agents has not just automated coordination, it has removed the last visible route by which anyone learned how the work actually holds together. See the missed reps.

A working rule#

On judgement with agents, AI agents and human judgement. On the stage-by-stage version, the Delegation Boundary Map. On why review fails, human in the loop is not a safeguard. On the legal duty, meaningful human oversight. On the override rule, when should I override AI.

Key sources

About this research#

Rahim Hirji is the author of SuperSkills (Kogan Page, 2026), keynote speaker on AI and human capability, and founder of The SuperSkills Intelligence Company. Agent-specific evidence is thin and this page reasons from adjacent findings, which it states above. Not legal advice. On a 90-day review cycle, because this is the fastest-moving question on the site.

How this research works  ·  Reviewed quarterly  ·  Found an error? Tell me and it is corrected on the page.

Cite this

Hirji, R. (2026). Should I let an AI agent act on my behalf? The SuperSkills Intelligence Company. Last reviewed 26 August 2026. thesuperskills.com/research/should-i-let-an-ai-agent-act-on-my-behalf

Questions answered on this page

Should I let an AI agent act on my behalf?

Sometimes, and the useful question is not whether you trust it. It is what the worst thing is that it can do before a human sees it, and whether you can live with that. Answer that and the trust question resolves itself. Leave it unanswered and no amount of confidence in the model helps, because you have not bounded the downside.

What should you decide before giving an agent autonomy?

Four things. Reversibility: sort actions into reversible, expensive to reverse, and irreversible, and grant autonomy accordingly. Blast radius: how far the consequence travels, since sending one email and sending one email to a client list are the same action. The stop: whether you can halt it mid-sequence and whether halting leaves things in a safe state. And the reconstruction test: whether you could explain afterwards why it did what it did.

Why are agents harder to oversee than assistants?

Because every oversight model in common use assumes a pause for review, and agents remove the pause. The Vaccaro meta-analysis found human-AI combinations underperforming the better party alone, with losses concentrated in decision tasks where a human judges whether a system was right. Agents perform that task at machine speed, in volume, without the human present. The human contribution has to move to the problem, intent, constraints and rejection criteria, before anything runs.

What is the quiet cost of agent delegation?

Agents absorb exactly the sequences of small tasks that used to constitute learning a job: chasing the thing, checking the thing, noticing the anomaly, following it up. Those look like overhead and they are where judgement is formed. An organisation that hands them wholesale to agents has removed the last visible route by which anyone learned how the work holds together.

In this hub

Judgement, oversight and accountability

Who decides, who checks, and who is answerable when the machine was involved.

Bring this into your organisation

If this describes something happening in your teams, say so.

Keynotes, board sessions and advisory work, drawing on research across more than 200 organisations in 30 countries. Tell me the room, the date and the shift you need. A reply within 24 hours.

Start a conversation

Topics and audiences  ·  All research

Agents are where this stops being a thought experiment. There is the AI agents and accountability version, and the full range of topics and audiences.

Box of Amazing

Rahim’s free weekly letter on AI and human capability

If this was useful, the weekly letter is where the thinking happens first. Most of what ends up on this site starts there. Weekly essays on AI, capability and the future of work. Read by 25,000 people, every week since 2017. Free, and one click to stop.

Opens Substack to confirm. No pitch in it, unsubscribe in one click, and nobody follows up because you read something.