14 ms·
Microsoft Azure Outage
- 6451937099 4y ago[dead]
- 6451937099 4y ago[dead]
- alkonaut 4y agoIt comes and goes. Teams and Azure DevOps some times works perfectly for a few minutes, then responds with all 503's for a few minutes.
- klaude 4y agoAnyone having problems with Azure too?
- ensocode 4y agoyes here. storage, db, apis - its not permanent but still persisting. It can be monitored at the azure status page as well https://status.azure.com/status https://status.azure.com/status
- kgdinesh 4y agoAt work, we all got kicked out of a teams meeting an hour back and sending/receiving e-mails on Outlook seems to be slow. Location: Chennai, India
- midasz 4y agoThis is going to be the most productive day ever
- maxaigner 4y agoReported issues with Teams, Microsoft 365, etc
- saikatsg 4y agoTeams is working now for me. However, all my notification preferences got reset!
- danjc 4y agoAuth via Microsoft ID is degraded, our platform is blipping (cache retries, message retries due to packet loss), access to the Azure portal is degraded and the Azure status page isn't loading consistently.
- neversaydie 4y agoSeeing problems with Azure DevOps in Western Europe here, can't open most pages/log in. Teams and Office appear to be working fine.
- LilBytes 4y agoNothing is working for me, Oceania/Australia. Including O365, Azure, Azure Devops.
- jupiterblues- 4y agoMinecraft, Asure, Office 365, etc... MS cloud services have issue
- NKosmatos 4y agoWhat's the point of having a status page if it doesn't indicate the issues? https://status.azure.com/en-us/status https://status.azure.com/en-us/status Azure, Teams, Outlook are almost down from Greece and Germany, and their status page shows that everything is fine :-)
- deleted 4y ago[deleted]
- berkut 4y agostatus.office.com had been down for 15 mins, but it's back up now...
- maushu 4y agoThe point is PR. Never trust a status page if it's not directly connected to the monitoring system.
- wrldos 4y agoThey never attach it to the monitoring because monitoring systems usually generate a lot of false positives which affect their published SLA.
- polack 4y agoThen they should have a "?" status that can be triggered by automated systems that acknowledge that it looks to be an issue but that they are manually investigating. If it's a false positive they just resolve it without it affecting SLA and if it's a real problem then us customers wouldn't have to debug our own stack for 2 hours before Microsoft informs us that they are the problem. EDIT: Wonder how many man-years of extra debugging work their non-working status page have caused the customers.
- deleted 4y ago[deleted]
- 4y ago
- saikatsg 4y ago> We've identified a potential networking issue and are reviewing telemetry to determine the next troubleshooting steps. You can find additional information on our status page at https://msft.it/6011eAYPc https://msft.it/6011eAYPc or on SHD under MO502273.
- ricardobayes 4y agoI'm so surprised by MS's strategy for using random domains and TLD's, this certainly don't make it easy for phishing avoidance.
- Tepix 4y agoMakes sense to use a different domain if everything is down because it could also effect DNS for the main domain.
- wiradikusuma 4y agoI think what the OP saying is, if you have multiple random domains, how would people know which ones are legit (or not)? Say I have mixxxrosoft.com, how would you know this is one of MS' official domains?
- tenplusfive 4y agoLuckily Microsoft also provides a service for that: Safelinks https://learn.microsoft.com/en-us/microsoft-365/security/office-365-security/safe-links-about?view=o365-worldwide https://learn.microsoft.com/en-us/microsoft-365/security/off... Also a personal favorite of mine: http://microsft.com http://microsft.com (not entirely sure if its just to prevent typosquatting or if this is actually used in some products)
- luckylion 4y agoI don't know whether it's a typo but https://support.microsoft.com/en-us/topic/contact-us-91f63b43-f11c-927d-be8e-fb1af4435293 https://support.microsoft.com/en-us/topic/contact-us-91f63b4... lists "EOC: criskgro@microsft.com (For CEE and MEA)" under the Microsoft Credit Services. It feels like a typo, but who knows. If they don't have anything in place to catch this type of error, it's probably a good idea to register every domain someone could accidentally type.
- markuman123 4y agorussia? shut down a service and halt the productivity of most companies in the west...because most companies moved to azure ad and teams.
- swarnie 4y agoI'm not sure Russia is as capable as you've all spent the last few decades making out....
- deathanatos 4y ago> russia? Oh please. Azure is plenty capable of taking themselves offline on their own.
- dustedcodes 4y agoAzure is the most developer hostile cloud environment. I have zero sympathy for people being affected by this because if you voluntarily use Azure then this is what you deserve. Sorry for being so miserable, but Azure has given me soooo much grief over the last 10 years that I'm just completely done with this shitshow of a platform.
- Tepix 4y agoWorks for me (not right now, but in general).
- dustedcodes 4y agoOf course, there's always someone who will say that -.- Did Windows ME and Windows Vista also work really great for you?
- Dalewyn 4y agoWindows ME was one of my favorite versions of Windows, not being ironic about it either. Its infamy has more to do with how Joe Average uses computers in general. As for Vista, while I did not use it in its day I can tell its problems were far more to do with crapass hardware manufacturers and their crapass drivers. Vista with access to 7's drivers and hardware runs just fine.
- jug 4y agoWindows ME did in fact work mostly fine here too, lol. Relatively speaking for Windows 9x performance, of course. I only used it for a year, not because I couldn't stand it but because such was the pace of major Windows updates back then. Windows Vista was honestly worse for me, not due to bugs but for being two years ahead the curve of hardware, and GPU vendors seemingly rolling their thumbs during betas and once WDDM¹ went live, they panicked and rolled out alpha quality work. So many driver crashes compounded with the heavy RAM requirements... Other than that, and with less of an UAC nazi, I could see an OS that was similar to what Windows 7 became if I squinted. Hardware had caught up, drivers were mature, and on top Microsoft optimized its performance. In hindsight, WDDM should've been an update to Windows XP that could be rolled out well in advance and let developers focus on a single thing rather than new OS compatibility on top, and deep changes like UAC. ¹ It was necessary work though: https://en.wikipedia.org/wiki/Windows_Display_Driver_Model https://en.wikipedia.org/wiki/Windows_Display_Driver_Model
- ChickeNES 4y agoDoes the Internet Archive use Azure? archive.org is throwing 503s
- voytec 4y agoTwo weeks ago they were affected by the Elasticsearch outage[1], too. [1] https://news.ycombinator.com/item?id=34337518 https://news.ycombinator.com/item?id=34337518
- ochrist 4y agohttps://downdetector.dk/ https://downdetector.dk/ indicates several MS products and services are having problems. Here is the status from MS on Twitter: https://twitter.com/MSFT365Status/status/1618149579341369345 https://twitter.com/MSFT365Status/status/1618149579341369345 Edit: Added this link which apparently is the new status page and seems to be updated: https://status.office365.com/ https://status.office365.com/
- dsign 4y agoThis makes you wonder if some centralization patterns, i.e. Azure AD, are not a national security problem?
- paganel 4y agoAt this point most probably, yes. Especially as more and more government entities/agencies are moving to the cloud, many of them to Azure (because of MS). I live in Eastern Europe, but I suspect that this migration is happening all around Europe and North America.
- laacz 4y agoCritical infrastructure cannot be reliant on a cloud (or internet availability, if possible). In most EU countries that's a law.
- ibejoeb 4y agoAzure AD is a nightmare. I don't know how many of you sign in to multiple tenants in the console, but it generally involves buying a new computer.
- gerdesj 4y agoIt involves a lot of private browsing sessions which is actually MS's recommendation! What a PITA.
- hansamann 4y agoDuckDuckGo.com - no search results showing up at all... are they on Azure?
- braymundo 4y agoDuckDuckGo is also affected (blank search results).
- hansamann 4y agoduckduckgo is not showing any results right now... are they on Azure, too?
- zidad 4y agoMost likely, because duckduckgo partially depends on Bing
- nosebear 4y agoI'm hearing from four different friends from four different companies in Germany that they can't really work right now.
- steve1977 4y agoIf they were relying on Outlook and Teams to be productive, they probably couldn't really work before either.
- hnarn 4y agoWhat a naive comment. As if the only truly important jobs exist in engineering and require nothing but git and a book on C.
- OscarDC 4y agoI interpreted this comment as more of a jab at how inefficient are outlook and teams themselves as applications. I don't know if it's the right interpretation to have, but I kind of agreed with it, considering huge issues I had with teams (curiously some of them are only there for linux users, weird when considering the fact that I only use teams' web page) - not saying I could do better though!
- choeger 4y agoYeah, what BS. Everyone knows that if you have a book on C, you can always quickly implement git yourself.
- vikramkr 4y agoI'm unsure what being in engineering has to do with using outlook and teams?
- steve1977 4y agoThat wasn't my point. But tools like Teams kill more productivity than they enable, at least in my experience. If anything, I was more productive yesterday, because I got disturbed less.
- oars 4y agoLinkedIn seems to be struggling as well. Lots of latency, page loads are taking 10-20 seconds for me.
- skc 4y agoHad a few dropped calls in Teams over here this morning (South Africa), otherwise our devops stuff is currently fine.
- hobofan 4y agoNot sure if it's directly related, but GitHub is also experiencing issues: https://www.githubstatus.com/ https://www.githubstatus.com/
- ricc 4y agoGH has been a Microsoft company since 2018...
- quickthrower2 4y agoGood to see GH is eating the dog food
- marvinblum 4y ago"We are investigating reports of issues with Actions. This looks related to Azure networking issue which is impacting multiple regions. We are seeing improvements and will continue to monitor this."
- spoils19 4y agoIt's good that Microsoft saved money via layoffs so that it balances out when customers leave Azure. Very forward thinking company.
- cryptonym 4y agoLeave to go where? On-premise and being miserable having to wait months to get a new server with poor automation, observability and worse outages? To another major cloud provider with similar pricing and outages? Cloud helped mostly with automation and scaling but if your system is that critical, you should consider a good CDN as load balancer and multi-cloud (or at least multi-region) for actual robustness.
- jakewins 4y agoAWS and GCP both have ~100% uptime in every region for VMs this month. Meanwhile the majority of Azure regions have had various outages in the same period: https://cloudharmony.com/status-of-compute https://cloudharmony.com/status-of-compute
- barbazoo 4y agoWow I didn't expect the difference to be so obvious.
- throwaway2037 4y agoIt is weird that this answer was downvoted. I agree. What a great page!
- azfubar 4y agoAlmost certainly due to Azure's broken policy where we have critical change advisory's that block deployments for huge periods of time towards the end of the year because of Black Friday and then holidays. Every team has basically been unable to deploy since the week before Thanksgiving when a surprise CCOA was pushed out by leadership at the behest of a certain big customer... then there was the World Cup and the winter holidays. Nobody could really deploy anything from a week before Thanksgiving until a week after the New Years... almost two months worth of batched changes and every team YOLO button pressing as soon as they could in January. And now layoffs so everyone is super unmotivated! Excellent stuff going on right now from Microsoft senior leadership.
- kornish 4y agoAh - so that's why GitHub Actions are unreliable right now.
- Benjamin_Dobell 4y agoGlad it wasn't just me. I was waiting over 10 minutes for a hosted runner.
- quickthrower2 4y agoSuch a late 2010s / 2020s problem :-(
- asim 4y agoCloud is the new power grid. When it goes down, we lose power to everything. Will we learn from the grid and decentralise some of the compute and cloud services?
- adql 4y agoOffice359 strikes again
- Yuioup 4y agoYou mean Office364
- wrldos 4y ago0<Office<365
- DoctorDabadedoo 4y agoEveryone deserves a break between Christmas and New Years, even the folks at MS! /s
- dx034 4y agoShows that all these availability zones and regions don't really help if an outage can knock out a whole cloud provider. And that's not specific to Microsoft. The only way to really ensure uptime is to use two providers. Sadly, that's basically only possible with on-prem/colocation where traffic is cheap.
- sofixa 4y agoIt's mostly Azure though that is badly designed to such an extent that multiple times there have been global outages. In general Azure availability, security (the only major cloud provider with not one but multiple cross-tenant security exploits) and usability are pretty terrible so it shouldn't be used for anything but saying "this is how it should not be done". GCP had a similar thing once, where a BGP update knocked out their Asian regions. AWS have never had a global outage. (And no, that time S3 in us-east-1 was down wasn't a global outage, the only customer code/workloads that were impacted was code interacting with S3 that didn't specify the region and had to rely on us-east-1 to determine it, and it didn't work anymore)
- wereallterrrist 4y agoSomeday someone will write a book about how AD, AAD, etc, exert the control they do at MS and go as unchecked (or at the time) as they do. AD's inability to execute made Azure a significantly less pleasant platform until they finally fixed accounts a couple of years back to properly do OAuth 2.0 with ARM. Maybe the book is just "AD brings in the money" but wow, they sure bring it down as well. Global outages like that always stink of AD.
- Andys 4y agoTo be fair, AWS once had a global Route53 outage, which was effectively a global outage for anyone using AWS for DNS.
- eurg 4y agoDo you have a link to an article about that? My google-fu is weak, and this sounds interesting - that should not happen to DNS - at all - and from the outside Route53 looks quite well managed. So what the heck did they do?
- osivertsson 4y agoMany games that use Azure PlayFab are down as well due to this. Both PlayFab services and PlayFab MPS game-server hosting are currently broken. https://status.playfab.com/ https://status.playfab.com/
- quickthrower2 4y agoDid we finally exhaust IPv4? /jk
- cube00 4y ago> The issue is causing impact in waves, peaking approximately every 30 minutes. Does anyone have any general ideas on what kind of outage manifests itself like this? Devices retrying to authenticate every 30 minutes and finding the service is down perhaps?
- urbandw311er 4y agoCan sometimes be scaling/monitoring loops. i.e. cluster comes up, provides some limited service, gets overloaded and drops below required performance metric, gets killed by monitoring/scaling system, repeat...
- kemals 4y agoThousandEyes public outage map shows the scale of the Office365 outage: https://www.thousandeyes.com/outages/ https://www.thousandeyes.com/outages/
- reset-password 4y agoI have some Azure services that are not able to consistently make outbound HTTP requests to my heartbeat monitoring service so I'm getting alert after alert this morning. This is just the nudge I needed, and I'll be moving the whole thing to Linode later this afternoon.
- idk1 4y agoDoes this mean they need to rebrand, because it's not up 365 days of the year? Maybe rebrand it to Microsoft 364.5?
- altairprime 4y agoThere’s 365.2425 days per year, so a six hour outage is just about 0.2425 hours, which suggests that they remain able to declare 365 when considering this specific outage only.
- ericpauley 4y agoI think the joke always went that they should rename it Microsoft 360.
- alkonaut 4y agoWouldn't it be quite simple to set up an unofficial status page that just pings some relevant services and if they have a disastrous outage at least, it shows it? Because I think it's clear that their status page is useless and "manual".
- generalizations 4y agoAnyone else remember the bad Windows Defender virus signature they put out on Friday the 13th a couple weeks ago? Microsoft is not having a good start to their year.
- mensetmanusman 4y agoHope the Leopard tanks aren’t running azure…
- stephencoyner 4y agoI did notice chatgpt was down earlier, but it could have been heavy usage caused
- ruffrey 4y agoIn the azure portal, it shows a "Routine Unplanned outage" - ??
- rossdavidh 4y agoWell points for honesty, at least. :)
- ugh123 4y agoI guess thats the 0.0001% of outage for an advertised 99.9999% uptime
- funnymony 4y agoAt least they have a sense of humor
- Eleison23 4y ago[dead]
- sli 4y agoEvery Azure product I've had to use has been lousy in every possible way. Azure DevOps at my last employer was a nightmare and nobody in the company liked it, not even the managers who decided on it.
- telcal 4y agoI use Azure DevOps daily and honestly have no issues, it works well. What didn't work for you?
- BLKNSLVR 4y agoI've been learning / using DevOps for the past four months and find it "quite good", and have previously used Jira, although not in great detail. I'm making the effort to learn it in increasing detail as it's the company-wide chosen system. I'm interested to know what made / makes it a nightmare for anyone else. (And I'm no fan of Microsoft as a whole)
- 6451937099 4y ago[dead]