System status
Everything is running
Component health for the Terrace platform, refreshed every 30 seconds from the same probes our on-call uses.
99.949%
30-day availability
253
monitored components
3
incidents this quarter
14min
median time to mitigate
Components
Component healthall systems operational
Live ingestoperational
60 days ago99.907% uptimetoday
Multi-angle packagingoperational
60 days ago99.888% uptimetoday
Low-latency deliveryoperational
60 days ago99.923% uptimetoday
Replay publishingoperational
60 days ago99.937% uptimetoday
Rights and territory checksoperational
60 days ago99.993% uptimetoday
Accounts and billingoperational
60 days ago99.924% uptimetoday
Incident history
Every incident since the platform launched, including the ones customers never noticed.
Elevated error rate in one metro
resolved09 Jul 2026 · impact: no data loss
14:52 UTC
Root cause identified: a faulty line card on one of two upstream ports. Traffic was drained from the affected device.
14:21 UTC
We are investigating elevated p99 latency reported by monitoring in a single facility.
15:40 UTC
The device was replaced and traffic re-balanced. Metrics are back to baseline; we are keeping the incident open for another hour to confirm.
Scheduled maintenance — border router firmware
resolved09 May 2026 · impact: none
02:00 UTC
Maintenance window opened. Requests were served by the standby control plane throughout.
03:12 UTC
Upgrade completed with no customer-visible impact.
About this page
Synthetic probes run from nine independent networks on three continents. A component counts as down when at least three probes on different networks fail two consecutive checks.
Anything that degrades a customer-visible SLO: availability, error rate, or the latency budget for the affected component.
Yes — the status API returns the same data as JSON, and every component exposes a Prometheus-compatible metrics endpoint.