4 ms·
Amazing metrics, the website performed flawlessly and their entire organization should get a giant round of applause
by tcarn 9y ago
Amazing metrics, the website performed flawlessly and their entire organization should get a giant round of applause
- delecti 9y agoAfter hearing pagers going off throughout that day, it's reassuring to hear someone on the outside say everything went flawlessly.
- rajathagasthya 9y agoI’m curious, how do you prep for things like the Prime Day? I imagine almost all parts of Amazon.com and AWS operate at their limits, especially resource-wise. What precautions do you take?
- lxmcneill 9y agoJeff covered this in the post (GameDays, excessive amounts of auditing, etc.). Regarding the resource limits you suggest, they mention metrics of 50+ pB of data movement and 3.34 trillion DynamoDB queries in 30 hours, all of which they elastically scale down after the event... so I'd say resource limitations are more in terms of humans on-deck and crisis management, rather than physical limitation of hardware. (edited to correct size - was 52 pB not 520 pB...)
- jlgaddis 9y agoYou might be interested to read the white paper that was linked in TFA.
- LiteskinKanye 9y agoWhat's the DevOps duty look like? Anytime you just see a page and root cause it and just tell your higher-up they are SOL
- mabbo 9y agoI can't speak for the whole company, but I've done 5 peaks (holiday seasons) as an SDE in the Warehouse/Delivery orgs. The key to remember about Amazon DevOps is that Developers are also DevOps. My team would usually share a dedicated DevOps team with 3 to 6 other teams- usually rotating front-line on-call duty between 2 or 3 people in the US and 2-3 in India. That DevOps person has the job of: what is the problem? Do my teams own this problem (if not, redirect to the right place)? Do I know how to immediately fix this? If not, for which of my teams do I page the on-call SDE?" When you have a great DevOps team, SDE oncall duty is a walk in the park. But DevOps people take time to become great- and many of them are also applying for transfers to SDE roles. Overall, each org decides how on-call/DevOps is. If time and effort and spent investing in stable software, it's easy. If other priorities get in the way, things can get bad.
- engi_nerd 9y agoSure, and for the second year in a row, the Amazon website told me a price for a deal, told me the deal was still available, and then wouldn't let me purchase it. Overall things went well for Amazon, but flawless? I think not.