← Research
Research

What should a board do about the AI safety warnings?

What was said in September 2026 and by whom, why boards cannot settle the odds, and the five decisions and five questions that hold regardless.

Last reviewed: 17 September 2026

Not adjudicate the extinction odds. A board should be able to show, in writing, which decisions machines make in the company name, who can stop each system, what people can still do unaided, how it would learn of an incident, and what management claims rest on. An evidence review by Rahim Hirji; every figure resolves to a graded entry in the evidence base that says what it does not show.

Questions this page answersAll 811 questions this research covers

A board does not need a view on whether AI will end the human race. It needs to be able to show, in writing, that the company is in control of the AI it has already deployed. September 2026 has produced resignations, extinction odds from 5 to 70 per cent, an essay from Anthropic’s chief executive calling for a slower frontier, bills in Washington and Westminster, and confident dismissals. None of that can be adjudicated from a boardroom. What can be decided there is which decisions machines may make in the company’s name, who can stop each system, what people must remain able to do unaided, how the board would find out if something went wrong, and what management’s assurances rest on. Those five hold whatever the odds turn out to be.

The answer, in one line

The board's job is not to have a view on superintelligence; it is to be able to show, in writing, that the company is in control of the AI it has already deployed.

Share as a card

What has been said, and by whom#

Jacob Coxon resigned from Anthropic on 8 September and posted that “Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives.” TIME reported over 90 million views in 24 hours, and that Evan Hubinger, Anthropic’s head of alignment stress testing, posted: “Jacob is correct here, we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade”, adding that Anthropic “do not yet have a plan to solve alignment for superintelligence”. Dario Amodei’s essay We Must Pace the Frontier said: “We must slow the pace at which we improve the capabilities of AI models.” Sam Altman told Axios: “I agree with Dario that we need to pace the frontier”.

Jensen Huang of Nvidia called extinction claims “complete nonsense”, the BBC reported. David Bellamy of the Institute of Foundation Models told TIME: “AI killing us all by creating dangerous viruses is total bogus”. In Washington, Senator Sanders and Representative Casar introduced a bill on 3 September to ban artificial superintelligence and pause frontier development, with criminal penalties. In the UK, AI minister Kanishka Narayan rejected a Lords emergency shutdown amendment on 11 September; binding rules remain “on the table” if voluntary testing proves insufficient, Crypto Briefing reported. First Secretary Louise Haigh told the TUC of “huge risks to our national security and society if the right guardrails are not put in place”; Business Secretary Jonathan Reynolds warned against complacency and hyperbole alike, according to The British Eye. The UK has no dedicated AI statute; Alex Sobel’s Private Member’s Bill on superintelligence has its second reading on 13 November, Lewis Silkin noted.

Why the board should not try to adjudicate the odds#

The figures range from Hubinger’s more than 10 per cent to Geoffrey Irving’s 50 per cent and Marcus Williams’s 70 per cent, as TIME reported on 15 September. Grace et al. 2024, a survey of 2,778 AI researchers cited on the AGI page, found a median of 5 per cent. Those giving the numbers are researchers rather than forecasters, and most say so. The International AI Safety Report 2026 calls this the evidence dilemma: capability moves fast and evidence about new risks arrives slowly, so acting early may entrench the wrong intervention and waiting may leave people exposed. A board that spends its September meeting deciding whether 5 or 50 is right is doing work it cannot do well and that changes none of its obligations. Whether the AI systems the company already runs are under its control today is answerable.

The five decisions that hold whatever the odds#

First, which decisions machines may make in the company’s name: the outcomes a system may determine without a person deciding: a credit limit, a refund, a hiring shortlist, a public reply. If the board cannot produce that list, the systems have made the decision by default. This is the ground covered by Rules Before Tools and by what board oversight of AI looks like.

Second, who can stop each system, and whether they would. Every debate about a national AI kill switch, including the Lieu-Moran House bill reported by the Wall Street Journal, runs into the same questions: which official, what threshold, and how to switch off something running across several jurisdictions’ cloud infrastructure. Inside a company, a stop often means a vendor change request and a week. Being able to stop a system is a technical fact. Being willing to stop it, mid-quarter, with revenue attached, is a leadership fact, and should be rehearsed before it is needed.

Third, what people must remain able to do unaided. Budzyn et al. 2025, in the Lancet Gastroenterology and Hepatology, found that 19 endoscopists averaging 27.6 years of experience saw unassisted adenoma detection fall from 28.4 to 22.4 per cent within months of routine AI use. That is capability debt: skill the organisation would need if the system stopped, lost without anyone deciding to lose it. The board should know which tasks the company could no longer do by hand, and whether that was chosen.

Fourth, how the board would find out if something went wrong. That requires an incident definition before there is an incident; what counts as a serious AI incident sets one out. Without it, the board hears first from a journalist or a regulator. The agent incidents of July and August were described by the companies involved as tests with guardrails disabled and monitoring not enabled. A company with no incident definition has made the same choice.

Fifth, what management’s claims rest on. When management says a system is safe, accurate or overseen, the board should ask what evidence supports that, and expect more than a vendor slide. A sampling regime, a reviewer disagreement rate, a log of overrides, a rehearsed stop: any of these is evidence. An assurance that oversight exists is a claim. The difference between governance and leadership is that governance files the assurance and leadership asks what it rests on.

Five questions to ask management this month#

One: list every decision a machine makes in this company’s name without a person deciding, and who approved each. Two: for each system on that list, who can stop it, how long a stop takes, and when it was last rehearsed. Three: which tasks could we no longer do by hand if the system were withdrawn, and did we decide that or discover it? Four: what is our definition of an AI incident, how many have we had this year, and how did the board hear? Five: for each assurance in the last board pack about AI accuracy or safety, what was the evidence and who produced it?

What “pace the frontier” changes for a company#

Less than the headlines suggest. The essay says “Progress will still seem fast, and we must make wise use of the time we gain.” Amodei’s three steps, embedded third-party evaluators, common safety standards with US government mediation, and international coordination including China, are addressed to frontier developers and governments. The Register called the plan regulatory capture; Speaker Mike Johnson said it could “smother innovation”. A slower frontier does not restore a decayed skill, define an incident nobody defined, or tell a board who can stop a system. Whether to pause AI development is a question for governments and laboratories. The handover decisions belong to the board, and were made, by choice or default, some time ago.

What this does not show#

This page does not show that the extinction warnings are wrong, or right. The odds quoted are individual estimates; the one systematic survey cited, Grace et al. 2024, found no consensus on pace. The Sanders-Casar, Lieu-Moran and Sobel bills had not passed at the time of review; UK ministers say existing powers suffice. Nothing here shows that a company taking the five decisions above would be protected from a frontier-scale failure. What it shows is that the questions a board can answer are the ones about its own systems, and those answers are available now.

Essay · SS-2026-262

Cite this page

Hirji, R. (2026). What should a board do about the AI safety warnings?. The SuperSkills evidence base, SS-2026-262. https://thesuperskills.com/research/what-should-a-board-do-about-the-ai-safety-warnings. Last reviewed 17 September 2026.

An evidence review by Rahim Hirji, not peer-reviewed research. For a material claim, cite the underlying study as well; every study here carries its own permanent link.

How citations and IDs work
Questions answered on this page

What should a board do about the AI safety warnings?

The board's job is not to have a view on superintelligence; it is to be able to show, in writing, that the company is in control of the AI it has already deployed. That means five things: which decisions machines may make in the company's name, who can stop each system and whether they would, what people must remain able to do unaided, how the board would find out if something went wrong, and what management's claims rest on.

Why should a board not try to decide whether the extinction warnings are right?

The figures given in September 2026 range from more than 10 per cent to 70 per cent, from people who say they are researchers rather than forecasters, while a 2024 survey of 2,778 AI researchers found a median of 5 per cent and no consensus. The International AI Safety Report 2026 calls this the evidence dilemma. A board cannot resolve it, and resolving it would change none of its obligations over the systems the company already runs.

What five questions should a board ask management this month?

Which decisions does a machine currently make in the company's name without a person deciding, and who approved each. Who can stop each system, how long it takes, and when it was last rehearsed. Which tasks the company could no longer do by hand, and whether that was a choice. What the definition of an AI incident is, how many there have been, and how the board heard. And for each AI assurance in the board pack, what evidence it rests on.

Does the call to pace the frontier change what a company should do?

Little. Dario Amodei's essay and Sam Altman's agreement with it are addressed to frontier developers and governments, and the essay itself says progress will still seem fast. A slower frontier does not restore a skill that has already decayed, define an incident nobody defined, or tell a board who can stop a system. The handover decisions a company has already made remain its own to review.

In this hub

Judgement, oversight and accountability

Who decides, who checks, and who is answerable when the machine was involved.

Ask the evidence
What does the evidence actually show?What should our board be asking about this?Where does Rahim disagree with the consensus?
Bring this into your organisation

If this describes something happening in your teams, say so.

Keynotes, board sessions and advisory work, drawing on research across more than 200 organisations in 30 countries. Tell me the room, the date and the shift you need. A reply within 24 hours.

Start a conversation

Topics and audiences  ·  All research

The five decisions that hold whatever the odds turn out to be, written down with names against them. Getting a board to that page in one session is the engagement. Board advisory.

This argument is one a board usually meets for the first time in the room. There is the boards and leadership version, and the full range of topics and audiences.

Box of Amazing

Rahim’s free weekly letter on AI and human capability

If this was useful, the weekly letter is where the thinking happens first. Most of what ends up on this site starts there. Weekly essays on AI, capability and the future of work. Read by 25,000 people, every week since 2017. Free, and one click to stop.

Opens Substack to confirm. No pitch in it, unsubscribe in one click, and nobody follows up because you read something.