← Retour au blog
tech 20 May 2026

Incident Report: Railway Blocked by Google Cloud (Resolved)

On May 19, 2026, Railway experienced a platform-wide outage due to an incorrect suspension of its Google Cloud account. This article details the incident, its resolution, and future preventive measures.

Article inspired by the original source
Incident Report: Railway Blocked by Google Cloud (Resolved) ↗ blog.railway.com

Introduction

On May 19, 2026, Railway, a platform for app deployment, faced a significant hurdle: an incorrect suspension of its Google Cloud account led to a platform-wide service disruption. This incident impacted all Google Cloud-hosted infrastructure, including the dashboard, API, and portions of the network infrastructure.

Incident Timeline

The incident began on May 19, 2026, at 22:20 UTC and lasted until approximately 06:14 UTC on May 20. During these eight hours, Railway faced a total service outage, rendering its APIs and databases inaccessible. Users immediately encountered 503 errors and other error messages and were unable to log in.

What Happened

Google Cloud erroneously placed Railway's production account in a suspended status. This action took offline the compute infrastructure hosted on Google Cloud. Although workloads on Railway Metal and AWS environments remained up, Railway's reliance on a Google Cloud-hosted control plane API to populate routing tables caused a cascade of failures beyond Google Cloud. At peak impact, all Railway workloads across all regions were rendered unreachable.

Response and Recovery

The Railway team immediately began working on recovering the Google Cloud environment. Deployments were blocked platform-wide while individual services were restored. Once the entirety of the infrastructure was restored, a significant backlog of queued deployments was gradually processed to avoid overwhelming the platform.

Preventive Measures

Railway took responsibility for the architectural decisions that allowed an upstream provider's action to cascade into a platform-wide outage. The team is committed to adopting more robust redundancy practices to prevent future disruptions. This includes diversifying cloud providers and enhancing caching systems.

Conclusion

This incident highlighted Railway's critical dependency on Google Cloud and the importance of developing resilient systems. By learning from this event, Railway strives to fortify its platform to ensure increased reliability for its users.

Let's discuss your project in 15 minutes.

Google Cloud Railway Incident Report Platform Outage Service Recovery
Deepthix newsletter · 100% AI · every Monday 8am

An AI agent reads tech for you.

Our AI agent scans ~200 sources per week and ships the best articles to your inbox Monday 8am. Free. One click to unsubscribe.

Visit the newsletter page →

Want to automate your operations?

Let's talk about your project in 15 minutes.

Book a call