Independent downtime-cost research, read by SRE and reliability teams.Sponsor this site →

Cloud Reliability Comparison

AWS vs Azure vs GCP: Outage Frequency & History Compared

Updated September 2026 · 28 documented major incidents across three clouds

Outage Frequency, In Brief

By documented major incidents, the three hyperscalers cluster around one to two significant outages per year each. Our database records 10 major AWS incidents since 2012, 9 for Azure since 2018, and 9 for GCP since 2016. Raw counts are not a clean reliability ranking - they reflect how long each provider has been tracked and how visible each event was. AWS outages concentrate in us-east-1 and affect the most customers per event; Azure incidents often hit Microsoft 365 and Teams; GCP fails less often but its global control plane lets a single failure scale worldwide within minutes (June 2025). Full side-by-side table, combined timeline, and SLA credit comparison below.

Outage Frequency & History, Side by Side

MetricAWSAzureGCP
Documented major incidents1099
Tracking windowsince 2012since 2018since 2016
Approx. frequency~1 / yr~1-2 / yr~1 / yr
Costliest incidentS3 Feb 2017 ~$150M (Cyence); Dec 2021 est. similarOct 2025 Front Door global outage (~8.5h)Jun 2025 Service Control / IAM (dozens of products)
Most recent majorJul 24 2026 us-west-2 network outage (~80 min; Apple Pay, Reddit, DoorDash)Feb 2026 multi-region control-plane failure (~12h)Sep 1 2026 us-central1 fiber-disconnection (~4h 11m, 19 products)
SLA credit range10% -> 30% (EC2, <99%)25% -> 100% (<99%)10% -> 50% (Compute, <95%)
Failure patternConcentrated in us-east-1; affects the most customers per eventOften hits Microsoft 365 and Teams - high enterprise visibilityRarer, but a global control plane means failures scale worldwide fast

Incident counts are the major outages documented on each provider page below; they are not exhaustive logs of every status-page event. No cloud provider discloses customer-impact figures for its own outages, and for eleven of the thirteen events in the timeline below no named analyst or insurance-modelling firm has published one either, so those read "None published" rather than carrying a guess. The two that do are October 2025 (CyberCube's preliminary insured-loss range) and July 2024 CrowdStrike (Parametrix, US Fortune 500 only). To size an event against your own revenue, the AWS and GCP chronologies show revenue at risk per $100M of annual revenue with the arithmetic shown. Updated October 2026.

Combined Outage Timeline (Most Costly, Recent)

DateProviderWhat HappenedPublished Cost
Sep 2026GCPus-central1-b (Iowa) fiber-disconnection during maintenance; all fibers pulled at once, 19 products degraded ~4h 11mNone published
Jul 24 2026AWSus-west-2 (Oregon) network path to Seattle metro failed ~80 min; cross-region traffic timed out; Apple Pay, Reddit, DoorDash, Hulu, PlayStation Network offlineNone published
Jul 16 2026AWSCloudFront VPC Origins control-plane failure; Frankfurt AZ (euc1-az2) capacity limit served global 5xx ~3.5h; Canvas, Blackboard, Hugging Face, UK National Lottery hitNone published
May 2026AWSus-east-1 (use1-az4) data-center cooling/thermal event; EC2/EBS, ELB, EKS, Redshift, MSK offline ~28h; Coinbase down ~7hNone published
Feb 2026AzureRemediation workflow disabled anonymous read on storage hosting VM extension packages; VM/VMSS/AKS provisioning failed multi-region ~12h, then Managed Identities token failures in East US & West USNone published
Nov 2025CloudflareConfig-propagation failure degraded a large share of the web (~2.4B users routed through Cloudflare)None published
Oct 2025AWSDynamoDB DNS race condition cascaded to EC2/Lambda/STS; ~70K orgs$38-581M insured (CyberCube)
Oct 2025AzureFront Door data-plane config change failed globally (~8.5h); cascaded to M365, Outlook, Copilot, Azure Portal, Xbox; Alaska Airlines, Costco, Starbucks hitNone published
Jun 2025GCPInvalid quota policy crashed Service Control; IAM auth failed across dozens of products; cascaded to Cloudflare, Spotify, Discord, OpenAINone published
Jul 2024CrowdStrike / AzureFalcon sensor update crashed Windows hosts globally, incl. Azure VMs; airlines, banks, hospitals hit$5.4B (US Fortune 500, Parametrix)
Jan 2023AzureGlobal WAN/BGP router config change broke connectivity across Azure and Microsoft 365 (Teams, Outlook, SharePoint) worldwide (~5h)None published
Dec 2021AWSus-east-1 EC2/ECS/Lambda/SNS outage (~7h), major global disruptionNone published
Feb 2017AWSS3 us-east-1 outage (~4h) disrupted a large portion of the internet$150M, S&P 500 (Cyence)

Per-provider chronologies: AWS, Azure, GCP. CrowdStrike was a software-update incident, not a cloud-provider failure, but it crashed Windows hosts including Azure VMs - see the CrowdStrike case study.

SLA Credits: What Each Cloud Actually Pays

Provider (service)Lower-tier creditMaximum creditClaim window
AWS (EC2)10% (<99.99%)30% (<99%)30 days
Azure (most services)25% (99%-99.99%)100% (<99%)Auto for some since 2024
GCP (Compute Engine)10% (99%-99.95%)50% (<95%)30 days

All credits are a percentage of the monthly fee for the affected service, applied to future invoices - never your revenue loss. Full SLA credit vs actual loss analysis, or calculate your own exposure.

Frequently Asked

Which cloud provider has the most outages?
By documented major incidents, all three sit around one to two significant outages per year. Our database has 10 major AWS incidents since 2012, 9 for Azure since 2018, and 9 for GCP since 2016. Raw counts favor whichever provider you have tracked longest - AWS events concentrate in us-east-1 and affect the most customers, Azure's often hit Microsoft 365 and Teams, and GCP's are rarer but scale globally fast.
What was the worst AWS, Azure, and GCP outage?
Only AWS has an outage with a published loss estimate, so the honest comparison is by breadth and duration. AWS: February 2017 S3 (~4h), which Cyence estimated cost S&P 500 companies ~$150M plus ~$160M for US financial-services firms, and December 2021 us-east-1 (~7h), a broader control-plane failure AWS never costed and no analyst firm has estimated. Azure: October 29, 2025 Front Door (~8.5h global), cascading across Microsoft 365, the Azure Portal, and Xbox, and January 25, 2023 global WAN (~5h) breaking Azure and Microsoft 365 connectivity worldwide. GCP: June 12, 2025 Service Control/IAM (~3h, us-central1 up to ~2h 40m), dozens of products, cascading to Cloudflare, Spotify, Discord, and OpenAI. Microsoft and Google have never published a customer-loss figure for any outage, and no analyst firm has published one for the Azure or GCP events here.
How do AWS, Azure, and GCP SLA credits compare?
All three pay a percentage of the affected service's monthly fee, never your revenue loss. AWS EC2: 10% below 99.99% up to 30% below 99%. Azure: 25% at 99%-99.99% up to 100% below 99%. GCP Compute Engine: 10% at 99%-99.95% up to 50% below 95%. Even the maximum credit covers a tiny fraction of the business loss from the same outage.

Updated 2026-04-27