Skip to content

Comment on Final Root Cause Analysis of Nov 18 Azure Service Interruptionparent

Comments

Only a guess but from how its worded it seems that the storage frontend that had already entered infinite loops may have taken the tem+ hours to restart.

Mark, the Azure CTO, gives a good breakdown of the time taken for each portion of the incident recovery in this video: http://channel9.msdn.com/posts/Inside-the-Azure-Storage-Outa...

That may help address these questions. Just FYI, I am an engineer in the Azure compute team.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.