Degraded performance Network · NL-AMS
RPKI-ROA validation misconfiguration on prefix via Zayo (AS6461). Fully resolved; follow-up actions tracked internally.
Live infrastructure health across 16 components in four datacenters. Polled every 20 seconds by external probes outside our ASN; measurements are published raw, including the bad ones.
Block storage · RO-BUH · NVMe firmware upgrade on one storage shelf, rolling.
Each bar below is one day, coloured by worst observed state. Hover a bar for the date & per-day summary.
Last 14 resolved events across all components — auto-refreshed on every page load from our probes.
RPKI-ROA validation misconfiguration on prefix via Zayo (AS6461). Fully resolved; follow-up actions tracked internally.
Garbage-collector ran hot after a large client delete burst. Fully resolved; follow-up actions tracked internally.
Zoned-namespace firmware upgrade on a subset of drives. Fully resolved; follow-up actions tracked internally.
Slow-query regression on the metadata service. Fully resolved; follow-up actions tracked internally.
IOMMU group re-mapping required a short guest stun — ~3% of hosts affected. Fully resolved; follow-up actions tracked internally.
Hypervisor-level memory pressure investigation — ~9% of hosts affected. Fully resolved; follow-up actions tracked internally.
ICMP rate-limit mis-tune surfaced during a probe sweep from SwissIX Zurich. Fully resolved; follow-up actions tracked internally.
IPv6 routing anomaly on carrier GTT (AS3257). Fully resolved; follow-up actions tracked internally.
NUMA imbalance briefly elevated p99 boot times on — ~4% of hosts affected. Fully resolved; follow-up actions tracked internally.
L4 scrubbing capacity briefly saturated during a large amplification attack. Fully resolved; follow-up actions tracked internally.
Object-store index promotion briefly held the write lock. Fully resolved; follow-up actions tracked internally.
Aggregate service uptime across all 17 components, by month. Months below 99.99 % trigger an SLA credit — see SLA.
How we measure uptime, what counts as a Tier-1 incident, where to subscribe and how post-mortems are published.