3 ms·
Curious how Spotify and other enterprise GCP users have mitigated GCPs repeated global outages. This isn’t their first nor will it be their last.
by baskire 6y ago
Curious how Spotify and other enterprise GCP users have mitigated GCPs repeated global outages. This isn’t their first nor will it be their last.
- bobviolier 6y agoThe Google Cloud was almost unaffected, during the outage you could not log into the cloud console or use tools like gcloud - but the services itself (vms, gke, pubsub, etc) kept working throughout the outage.
- wdb 6y agoThat's not correct, you wouldn't be able to connect these services after restart. E.g. you wouldn't be able to auth to Cloud SQL or Datastore. Stopping development
- jrockway 6y agoServices hosted on GCP didn't lose any traffic during this incident. This was a control plane outage, not a data plane outage. That means that if you needed to go in and adjust something to resolve your own outage, you were in trouble, but if your services were healthy, nothing bad happened. As the time of the outage increases, service health probably trends to zero without being able to manage the service... but for a few hours, it's not a disaster. Usually.