Skip to content
Monitoring dashboard
NetworkingMaintained2023

Network Monitoring & Alerting

Self-hosted monitoring that catches link degradation before users report it.

Overview

A monitoring stack covering link health, device availability, and latency trends, with alerting routed to the channels the team actually watches.

01The problem

Faults were reported by users rather than detected by the team. By the time a ticket arrived, the outage had usually been running for a while.

02The solution

Deployed polling and availability checks across the estate, baselined normal behaviour, and alerted on deviation rather than fixed thresholds. Escalation routing means the right person is notified at the right hour.

Screenshots

  • Link health over time
  • Alert routing rules

More work