4 ms·
FTA: "say datacenter automatic failover component, in one part of the territory. Getting this component right may take months of your time. But it is OK. You ar
by suff 7y ago
FTA: "say datacenter automatic failover component, in one part of the territory. Getting this component right may take months of your time. But it is OK. You are building a new street in one of the suburbs, and this adds up to the big picture."
Wait, what?!
Driving metaphor aside... 'MONTHS of your time' to implement failover? People, this is WHY the cloud was invented, so we didn't have to spend months re-inventing. In AWS and even GCP, this takes a day or less if you know the (well documented easy to use) storage offerings. Seriously reconsider your selection criteria when you start saying things like this, because what I just heard is that my team just told me implementing failover is going to cost $60k. Guess how much easier that made my case to switch to another cloud-native offering? TCO over everything. Even Ballmer would agree with that - he made the same case for Windows, against Linux in the 90's.
- silverlake 7y agoHe's talking about building the underlying failover mechanism for CosmoDB. For a customer it's easy and automatic, but GCP and AWS and Azure have to build it first.
- rad_gruchalski 7y agoBut months?
- manigandham 7y agoTo build a globally distributed database offering as complex as Cosmos DB with multiple SLA requirements? Can you build it faster?
- suff 7y agoSurely you're referring to Microsoft's team's time to implement failover to the product offering, not time every customer spends on implementation??? FTA: "if _you_ are an engineer working on a small thing, say datacenter automatic failover component" (emphasis added). Pretty sure he's talking about EVERY customer spending months to turn on fail-over.
- samdixon 7y agoThe previous poster is correct, I think you are mistaken. It should read like this: "if you are an engineer at Microsoft working on a small azure component, say datacenter automatic failover component" At least this is how I read it since the OP works at MS it appears.
- 013a 7y agoI can't for the life of me think of a compelling reason why regional fallover is something most companies would want in a db. It sounds great, but the reality is: the last time Azure had a regional failure, it also took down other regions. AWS and Google also have similar horror stories. On paper, regions are geographically isolated, but the reality is that in all of these instances of failure there was some "super-region" that has some core global infra that every region relied on, and it went down (AWS and us-east-1, Azure and south central US). You don't actually want single-provider multi-regional, what you want is multi-provider, like what Anthos is trying to do. But that's a harder sell to CIOs, and Azure checks a lot of boxes on paper despite being just horrible. The other funny thing about Azure is how they champion the number of regions they have (54), but many of these regions only have one AZ (only 8 have more than one). So when they say that something is multi-region or does regional fallover, its like, "great, that's TABLE STAKES for getting HA on Azure". But with AWS, you have at least two AZs in every region, so its not as big of a deal. GCP is the same way, but there's some language in their docs even AWS poked fun at during Re:Invent last year where they say that AZs "often" have isolated power and networking. Not always, just often. So are they truly isolated?
- deleted 7y ago[deleted]
- manigandham 7y agoThis isn't about customers using the cloud. His perspective is as an engineer on the Azure team building these databases and other features for you to use. He's "inventing" the cloud so that you don't have to reinvent, as you state.
- dharmashuklaMS 7y agoHi, I am from Azure Cosmos DB Engineering Team. The article is referring to the implementation details of automatic failover mechanisms inside of Cosmos DB service. For the customers, automatic failover is available as a turnkey capability. Customers do not need to spend any time to implement automatic failover. Thanks.