← Research
Research

What is alarm fatigue?

The best measured failure of a safety control anywhere in the record, and the reason flag-it-for-a-human is a design problem rather than an answer.

Last reviewed: 16 September 2026

A definition page on alarm fatigue, the desensitisation that follows exposure to a high volume of warnings, built on three sources read at source: an intensive care measurement, a regulator's own event analysis with its own limits attached, and one accident investigation in which an engineer traded false alarms against missed events on the record.

Question this page answersAll 811 questions this research covers

Alarm fatigue is what happens to a person who is warned too often. The warnings stop registering, then they get silenced, and eventually the one that mattered goes the same way as the thousands that did not. It is the best measured failure of a safety control anywhere in the record. The standard answer to AI oversight, flag the risky cases and have a human check those, inherits that failure whole.

The answer, in one line

It is the desensitisation of a person to warning signals caused by exposure to a high volume of them, most of which turn out not to require action.

Share as a card

Definition#

Alarm fatigue: the desensitisation of a person to warning signals caused by exposure to a high volume of them, most of which turn out not to require action, leading to slower responses, silenced alarms and missed true events. An established term in patient safety and human factors, used here in its original sense and not claimed by this research.

Share this definition as a card

The number that makes the argument#

Barbara Drew and colleagues instrumented five adult intensive care units at the University of California, San Francisco and stored every monitor signal for the thirty-one days of March 2013: seven ECG leads, pressure, oxygen saturation and respiration waveforms, user settings and every alarm, for 461 consecutive patients. The count was 2,558,760 unique alarms in a single month, made up of 1,154,201 arrhythmia alarms, 612,927 parameter alarms and 791,632 technical alarms. Of those, 381,560 were audible, an audible burden of 187 per bed per day.

Nurse scientists then annotated 12,671 arrhythmia alarms against a defined protocol, with 95 per cent agreement between raters and a Cohen kappa of 0.86. 88.8 per cent of those were false positives. Among the true ventricular tachycardia alarms, 93 per cent were not sustained long enough to warrant treatment.

That figure is quoted carelessly and this page will not join in. The 88.8 per cent applies to the 12,671 annotated arrhythmia alarms and to nothing else. It does not describe the 2.56 million, and a page saying that 88.8 per cent of clinical alarms are false has misread the paper. The study is single-centre, one month, five units, and it was funded by GE Healthcare.

Ignoring them is the rational response#

A clinician hearing 187 audible alarms a bed a day, where nearly nine in ten of the annotated ones carried no signal, is receiving an instruction about what those sounds mean. Treating them as noise is an accurate inference from the evidence available, and the language of fatigue slightly misdescribes it: this is a base rate doing what base rates do. The person has learned the true positive rate of the instrument and adjusted, and the adjustment is correct on average and catastrophic on the exception.

The Joint Commission documented the exception. Its Sentinel Event Alert 50, issued on 8 April 2013, recorded 98 alarm-related events between January 2009 and June 2012, of which 80 resulted in death, 13 in permanent loss of function and five in unexpected additional care or extended stay, with 94 occurring in hospitals. The contributing factors it recorded are the shape of the behaviour: alarm signals inappropriately turned off (36), absent or inadequate alarm system (30), alarm signals not audible in all areas (25), improper alarm settings (21). Clinicians turn the volume down, turn the alarm off, or set it outside safe limits because of the volume of signals. The Alert led to National Patient Safety Goal NPSG.06.01.01, phased in from 1 July 2014, which made the authority to change or silence an alarm something an organisation has to assign to named people.

The Commission's own footnote limits what those 98 events can be used for: reporting is voluntary, represents only a small proportion of actual events, and supports no conclusion about relative frequency or trend. The widely repeated claim that 85 to 99 per cent of alarm signals require no clinical intervention is quoted by the Commission from AAMI Horizons in 2011 and is not a Joint Commission measurement; the estate has not read the AAMI source and does not use the range.

A designer making the trade explicitly, and the cost of it#

The clearest record of an engineer weighing false alarms against missed events sits in a road accident investigation. In its report on the fatal collision in Tempe, Arizona on 18 March 2018, the National Transportation Safety Board found that the developer had disengaged the vehicle's factory forward collision warning and automatic emergency braking during automated operation, and that on detecting an emergency the system entered a one-second period of action suppression, withholding braking while it verified the hazard or waited for the operator to take control. No alert was given to the operator when action suppression began. The system had detected the pedestrian 5.6 seconds before impact and tracked her all the way without ever classifying her correctly. The stated reason for suppressing the response was concern about false alarms.

The NTSB determined the probable cause to be the operator's failure to monitor the driving environment while visually distracted, and found she would likely have had time to react had she been attentive, so nobody should say the missing alert caused the crash. What the report does establish is a design in which a human being was named as the primary countermeasure in an emergency and simultaneously denied the signal that would have let them act as one, with the volume of false alarms given as the reason. That is alarm fatigue anticipated by the engineer and paid for in advance.

Why this is the first thing to say about AI oversight#

Lisanne Bainbridge saw in 1983 that a person cannot watch a quiet system indefinitely, and her proposal was to hand the watching to an automatic alarm system connected to sound signals. Medicine then ran that experiment for forty years at enormous scale. The result is the evidence above: the remedy relocates the problem to whoever must now decide which alarms deserve a response, and it relocates it into a form where the correct everyday behaviour and the correct exceptional behaviour are opposites.

Most proposals for AI oversight have the same shape. Route the uncertain outputs to a reviewer. Flag the high-risk decisions. Surface a confidence score. Each creates a queue whose base rate the reviewer will learn, and the fluency of model output makes the learning worse, because a false flag on a well-written answer looks like a false flag rather than like a near miss. An oversight design that generates more signals than a person can act on has not added a control. It has added a number that an auditor can count.

Four questions that tell you whether a flag is a control#

Key sources

The argument this completes is the ironies of automation, and the limit that made an alarm seem necessary is the vigilance decrement. On controls that exist and do not work, designing a stop button people will use and human in the loop is not a safeguard. On the reviewer's side of it, automation complacency and who owns verification when AI does the work. On what a genuine oversight standard requires, meaningful human oversight.

Explainer · SS-2026-255 · Graded against the published rubric

Cite this page

Hirji, R. (2026). What is alarm fatigue?. The SuperSkills evidence base, SS-2026-255. https://thesuperskills.com/research/what-is-alarm-fatigue. Last reviewed 16 September 2026.

An evidence review by Rahim Hirji, not peer-reviewed research. For a material claim, cite the underlying study as well; every study here carries its own permanent link.

How citations and IDs work
Questions answered on this page

What is alarm fatigue?

It is the desensitisation of a person to warning signals caused by exposure to a high volume of them, most of which turn out not to require action. The consequences are slower responses, alarms turned down or switched off, thresholds set outside safe limits, and true events missed. The term is established in patient safety and human factors.

How many false alarms are there in an intensive care unit?

Drew and colleagues recorded 2,558,760 unique alarms across five adult intensive care units at the University of California, San Francisco in the thirty-one days of March 2013, for 461 consecutive patients: 1,154,201 arrhythmia, 612,927 parameter and 791,632 technical. Of 12,671 arrhythmia alarms annotated by nurse scientists, 88.8 per cent were false positives. That percentage applies only to the annotated arrhythmia alarms and not to all 2.56 million.

Has alarm fatigue killed anyone?

The Joint Commission's Sentinel Event Alert 50 of 8 April 2013 recorded 98 alarm-related events between January 2009 and June 2012, of which 80 resulted in death, 13 in permanent loss of function and five in extended care. Contributing factors included alarm signals inappropriately turned off in 36 cases and improper settings in 21. The Commission states that its reporting is voluntary and represents only a small proportion of actual events, so no conclusion about frequency or trend can be drawn from those numbers.

What does alarm fatigue mean for AI oversight?

Most AI oversight designs create a queue: flag the risky outputs, route the uncertain ones to a reviewer, surface a confidence score. Every queue has a base rate, and the reviewer will learn it. If the flags are mostly false the reviewer will treat them as noise, correctly on average and catastrophically on the exception. Four questions separate a control from a counter: what is the true positive rate and does the reviewer know it, is the volume inside what one person can action, who is allowed to turn it down and is that written down, and does anyone sample the unflagged population to find the misses.

Is alarm fatigue the same as automation complacency?

They are related and distinct. Automation complacency is reduced monitoring of a system because it has been reliable, so the operator under-checks a system that is usually right. Alarm fatigue is the reverse case: the system is usually wrong when it speaks, and the operator learns to discount it. Both are rational adjustments to an observed base rate, and both fail on the exception.

In this hub

Definitions

The terms this field uses, defined against their primary sources.

Ask the evidence
What does the evidence actually show?What should our board be asking about this?Where does Rahim disagree with the consensus?
Bring this into your organisation

If this describes something happening in your teams, say so.

Keynotes, board sessions and advisory work, drawing on research across more than 200 organisations in 30 countries. Tell me the room, the date and the shift you need. A reply within 24 hours.

Start a conversation

Topics and audiences  ·  All research

Box of Amazing

Rahim’s free weekly letter on AI and human capability

If this was useful, the weekly letter is where the thinking happens first. Most of what ends up on this site starts there. Weekly essays on AI, capability and the future of work. Read by 25,000 people, every week since 2017. Free, and one click to stop.

Opens Substack to confirm. No pitch in it, unsubscribe in one click, and nobody follows up because you read something.

Running an event, or responsible for how AI arrives in your organisation? Keynotes  ·  Advisory for CEOs and boards  ·  Enquire