Alert fatigue
Also known as Alert overload, Notification fatigue
Definition
Alert fatigue is the reduced ability or willingness to respond carefully to alerts after repeated exposure to notifications that are noisy, low priority, or rarely actionable.
Why alert fatigue happens
An alert is useful when it asks someone to take a timely action. If a notification fires for a condition that is harmless, already being handled, or impossible to fix during the response window, people learn that the notification does not reliably deserve attention. Over time, they may acknowledge alerts mechanically or miss a serious event among routine noise.
Alert fatigue often comes from thresholds chosen without a service objective, duplicate alerts for the same symptom, or alerts that describe a cause that responders cannot directly change. A high request count may be normal during a launch, while a sustained error ratio may threaten users even when the absolute count is small.
Improve the signal
Start by defining what the responder should do when an alert fires. Tie paging alerts to a user-facing symptom or an SLO risk, and route lower urgency observations to a place that does not interrupt incident response. Group related events and include links to the relevant dashboard, runbook, and ownership information.
Review alerts after incidents and on a regular schedule. Track false positives, duplicate notifications, time to acknowledge, and alerts that received no action. These measures help reveal whether the system is producing actionable evidence.
Alert quality is a team practice
Reducing noise is not just a monitoring configuration task. Teams need clear ownership, realistic thresholds, and a process for retiring alerts that no longer protect a service. Engineering analytics can add another useful view by showing whether operational interruptions line up with changes in code or workflow. The goal is a smaller set of signals that people trust when the stakes are high.
How this relates to Weave
Weave helps connect engineering activity and quality signals so teams can investigate whether recurring alerts correspond to recent changes, review bottlenecks, or operational work. That context can help teams decide which alerts deserve action and which need redesign.
Explore Engineering intelligence