Degraded performance Block storage · IS-RKV
NVMe device failure, RAID rebuild in progress. Fully resolved; follow-up actions tracked internally.
Live infrastructure health across 16 components in four datacenters. Polled every 20 seconds by external probes outside our ASN; measurements are published raw, including the bad ones.
Web control panel · Rolling kernel upgrade with cgroup-v2 schedule tuning; brief p99 latency bump possible.
Each bar below is one day, coloured by worst observed state. Hover a bar for the date & per-day summary.
Last 14 resolved events across all components — auto-refreshed on every page load from our probes.
NVMe device failure, RAID rebuild in progress. Fully resolved; follow-up actions tracked internally.
Slow-query regression on the metadata service. Fully resolved; follow-up actions tracked internally.
NUMA imbalance briefly elevated p99 boot times on — ~6% of hosts affected. Fully resolved; follow-up actions tracked internally.
Garbage-collector ran hot after a large client delete burst. Fully resolved; follow-up actions tracked internally.
Snapshot scheduler back-pressure cleared after tuning. Fully resolved; follow-up actions tracked internally.
Live-update websocket reconnect loop during a backend roll. Fully resolved; follow-up actions tracked internally.
RPKI-ROA validation misconfiguration on prefix via Zayo (AS6461). Fully resolved; follow-up actions tracked internally.
Garbage-collector ran hot after a large client delete burst. Fully resolved; follow-up actions tracked internally.
Zoned-namespace firmware upgrade on a subset of drives. Fully resolved; follow-up actions tracked internally.
Slow-query regression on the metadata service. Fully resolved; follow-up actions tracked internally.
Aggregate service uptime across all 17 components, by month. Months below 99.99 % trigger an SLA credit — see SLA.
How we measure uptime, what counts as a Tier-1 incident, where to subscribe and how post-mortems are published.