All systems operational
Live operational status by service and region, with 90 days of history. Incidents are published here whether or not anyone noticed them.
Incident history
Everything from the last 12 monthsElevated latency on analytics ingest, US West
A backlog in the analytics ingest queue delayed dashboard freshness to roughly 20 minutes in US West. Call handling, scheduling, and all customer-facing paths were unaffected. Resolved by scaling the consumer group and adding a backpressure alert at 60% of queue capacity.
Carrier failover event, US East
A primary carrier degraded in US East. Automatic failover moved traffic to the secondary carrier within 90 seconds. Eleven calls experienced a longer connect time and none dropped. We have since reduced the failover threshold.
Delayed SMS delivery, all regions
An upstream messaging provider queued outbound SMS across all regions. Confirmations were delayed rather than lost, and all queued messages delivered on recovery. We added a second provider to the failover path in February.
