Skip to content
Upstash · Data and observabilityMay 12, 2026, 09:49 UTC

Fly.io Upstash Redis Service Distruption

CriticalNetworkUpdated 19h ago
May 12, 09:49 UTCMay 12, 13:48 UTC
Duration
3h 59m
Impact
Critical
Root cause
Network
Upstash, 90 days
5 incidents
Affected
Not listed by the vendor.
Status page

Final update

On May 12th and 13th at various times, a subset of Upstash Redis instances on [Fly.io](http://Fly.io) experienced intermittent hangs and elevated error rates. The Redis process would stall inside a logging syscall , alive but not making progress , which made the issue hard to spot from our usual telemetry. After investigating with Fly's team, we identified the root cause as a bad interaction between a recent guest kernel update on Fly's newer machines and an upstream Cloud Hypervisor bug \([cloud-hypervisor#7672](https://github.com/cloud-hypervisor/cloud-hypervisor/issues/7672)\) affecting log writes from inside the VM. We mitigated by disabling the affected logging paths, and Fly has since rolled out a hypervisor-side patch, fully resolving the issue. No data was lost. Sorry for the disruption.

Timeline

  1. Postmortem · May 15, 12:43 UTC
    On May 12th and 13th at various times, a subset of Upstash Redis instances on [Fly.io](http://Fly.io) experienced intermittent hangs and elevated error rates. The Redis process would stall inside a logging syscall , alive but not making progress , which made the issue hard to spot from our usual telemetry. After investigating with Fly's team, we identified the root cause as a bad interaction between a recent guest kernel update on Fly's newer machines and an upstream Cloud Hypervisor bug \([cloud-hypervisor#7672](https://github.com/cloud-hypervisor/cloud-hypervisor/issues/7672)\) affecting log writes from inside the VM. We mitigated by disabling the affected logging paths, and Fly has since rolled out a hypervisor-side patch, fully resolving the issue. No data was lost. Sorry for the disruption.
  2. Resolved · May 12, 13:48 UTC
    This incident has been resolved.
  3. Monitoring · May 12, 12:05 UTC
    A fix has been implemented and we are monitoring the results.
  4. Identified · May 12, 10:50 UTC
    The issue has been identified and the fix is being implemented.
  5. Investigating · May 12, 09:49 UTC
    Some regions are experiencing connectivity issues due to an ongoing network problem. We are currently investigating

More from Upstash

Full history

Also caused by network

All
StartedIncidentDuration
Sep 1901:36 UTCElevated errors in Ashburn, VA (IAD)Cloudflare0m
Sep 1411:59 UTCIncreased wait times for macOS jobsCircleCI1h
Sep 410:57 UTCIncreased errors on High Performance Edge Network - FRA RegionNetlify0m
Sep 214:18 UTCUpstream network issuesFly.io1h 39m
Sep 116:11 UTCService degradation in GCP us-central-1ClickHouse19m
Sep 114:44 UTCMultiple products in us-central1-b are experiencing network service degradation.Google Cloud4h 8m

From vendors' own status pages and disclosures. Times as reported. Logos via logo.dev; trademarks belong to their owners.

Weekly: the week's major outages, postmortems and breaches, Saturday mornings.