As we rely on the Internet more and more, the big outages are hurting us each time we turn around.  In the last month, we saw outages to Amazon Web Services (AWS), Microsoft Azure, and now Cloudflare.  Each of these outages had an impact that was significant to businesses, users, and daily life.  Each of these was widespread due to the level of reliance that we have on each of these.  This Cloudflare issue was attributed to a configuration file that grew too large, crashing their system.  The previous ones were rumored to be as a result of a change for one, and a system code bug in the other.  No matter the cause, it’s important to know that these things do happen, and will continue to happen.  Always have a contingency plan for your internet based needs, whatever they are, from business to personal.  

 

CTO blames bot mitigation bug triggered by routine config change.

Source: Cloudflare’s CTO apologizes after error takes huge chunk of the internet offline — ‘we failed our customers and the broader internet’