Real-Time Status Monitor

Is it just me, or is something actually down?

⚡ Live data cache auto-refreshing in 60s
🔴Internet Health
97.34%
Widespread Issues
15 services with active issues Updated live

GitHub

Operational

All Systems Operational.

90-Day Uptime: 95.06% Checked: just now

GitLab

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Bitbucket

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

npm Registry

Operational

All Systems Operational.

90-Day Uptime: 99.56% Checked: just now

Vercel

Operational

All Systems Operational.

90-Day Uptime: 98.89% Checked: just now

Netlify

Operational

All Systems Operational.

90-Day Uptime: 98.61% Checked: just now

Cloudflare

Degraded

A fix has been implemented and we are monitoring the results.

90-Day Uptime: 94.61% Checked: just now

AWS

Outage

Region Availability.

90-Day Uptime: 90.00% Checked: just now

Google Cloud

Outage

Incident Report Summary On Tuesday, 1 September 2026, some customers in us-central1 experienced network service degradation and instance isolation for a duration of 4 hours and 11 minutes. To our customers whose businesses were impacted during this disruption, we sincerely apologize. This is not the level of quality and reliability we strive to offer you, and we are taking immediate steps to improve the platform’s performance and availability. Root Cause The event was triggered by the inadvertent physical disconnection of network fiber-optic cables during routine hardware maintenance. A technician was performing a scheduled capacity upgrade on data center routers that support a fraction of capacity in the us-central1-b zone and a small fraction of capacity in the us-central1-f zone. During the hands-on maintenance, optical fibers connecting the data center routers to the network fabric were unplugged and the optical transceivers were replaced with a transceiver that supports higher-density connectivity. This transceiver was not compatible with the transceivers still in place in the data center fabric, causing a loss of connectivity for each fiber. The network is designed with redundant routers in separate data center rooms within the same building so that failure of one router does not interrupt service. The maintenance was intended to upgrade one router at a time over several days, with traffic diversions at each step and verification between steps that the network has returned to a fully-connected state. However, a procedural error in the manually orchestrated upgrade process caused the complete list of transceiver replacements across all routers to be issued to the technician without instructions to sequence the work one router at a time. The maintenance workflow also did not include the expected human and software verification steps for detecting unintended disruption. As a result, the technician sequentially unplugged all fiber paths across the affected devices within 13 minutes. The speed and nature of the error prevented warnings of incorrect action from reaching the engineer before connectivity was lost. Additionally, a standing procedure to halt if light is detected on any fiber-optic cable after unplug was not followed. This resulted in compute capacity in the impacted zone being isolated from the network. Customers were unable to reach their virtual machines, and those virtual machines could not establish connections outside their zone. Remediation and Prevention The issue was detected immediately by automated network loss monitoring systems as well as proactive probes, which rapidly engaged the network engineering and incident response teams. To mitigate the immediate customer impact, engineering teams actively moved traffic away from the impacted infrastructure. Concurrently, hardware operations technicians on site identified the disconnected optical links and physically re-inserted the original optical transceivers. Once the physical links were fully restored, traffic flow rates normalized and the traffic was redirected back to return the capacity to service. Google is committed to preventing a repeat of this issue in the future and is completing the following actions: Reinforce training in us-central1 region and globally around safe maintenance techniques, including the mandatory requirement to verify whether light is on every unplugged fiber. Finish the migration of network upgrade workflows to a fully automatically sequenced orchestration system. Most common regional network workflows were completed in 2025, and zonal workflows are in progress. Complete the deployment of work stop alerting for actions resulting in the unintended disconnection of live fiber cables in Google data centers. Complete the deployment of the automated system that will move regional traffic away from faulty zones, reducing the time from approximately 19 minutes in this instance to around 5 minutes. Detailed Description of Impact On Tuesday, 1 September, from 07:41 to 11:52 US/Pacific, a portion of the us-central1-b and us-central1-f zones experienced severe network degradation and resource isolation. 1. Multiple products affected: The network disruption to/from a portion of us-central1-b and us-central1-f zones impacted all zonal GCP Products for some customers in those zones. 2. Regional impact: Regional products using capacity in the affected data center were affected until traffic diversions were fully in place at 08:00 US/Pacific. Google Kubernetes Engine and Cloud Run / Google App Engine had longer impacts, discussed below. 3. failure rate: Traffic flow drop rates for resources hosted in the affected area reached 100% during the peak of the incident, resulting in unreachable virtual machines and elevated lost data packets. Affected Services and Features Google Compute Engine: Inability for users in the affected portion of us-central1-b and us-central1-f zones to access virtual machines externally, and inability for VMs to reach remote resources. Cloud SQL, AlloyDB for PostgreSQL, Google Cloud Bigtable, Cloud Filestore: Data access and connectivity severed for instances localized strictly to the impacted infrastructure. Cloud Spanner, Cloud Firestore: Elevated Remote Procedure Call (RPC) failure rates, write timeouts, and response delays spikes for the nam5 multi-region instances due to severed connectivity to components running in the impacted zone. Virtual Private Cloud (VPC), Cloud NAT, Apigee, Google BigQuery, Google Cloud Dataflow, Looker, Google SecOps SOAR, Hybrid Connectivity, Cloud Interconnect: Elevated network lost data packets, connectivity drops, and API (the engine that lets different apps talk to each other) timeouts in the impacted zone. Google Kubernetes Engine (GKE): Between 07:40 and 09:15 US/Pacific on 2026-09-01, customers experienced errors and timeouts on requests to their GKE cluster control planes (Kubernetes API (the engine that lets different apps talk to each other) servers) in us-central1. This could have caused new or updated workloads to fail scheduling and Kubernetes API (the engine that lets different apps talk to each other) operations to fail. Cloud Run / Google App Engine: Temporary response delays spikes and pending queue aborts as backend workloads automatically evacuated and shifted to healthy capacity. A subset of workloads in the specifically impacted area experienced degradation until 11:52 US/Pacific. Customer Impact Customers with resources hosted in the specific affected clusters within us-central1-b and us-central1-f experienced a complete loss of network connectivity starting at 07:41 US/Pacific. During this time, virtual machines and associated services became unreachable from the internet and from other Google Cloud regions, and those internal resources could not initiate outbound connections. The issue was detected almost immediately and engineers took traffic diversion actions at 07:45 and 08:00 US/Pacific to reduce impact to regional products. By 08:50 US/Pacific, the majority of physical connections were restored and traffic began recovering, and the traffic diversion actions were removed at 09:19 US/Pacific. Full recovery across nearly all affected services and long-running operations was confirmed by 11:52 US/Pacific. Because the network isolation was strictly limited to specific clusters within us-central1-b and us-central1-f, multi-zonal deployments correctly utilizing redundancy across other zones in us-central1 were largely able to bypass the physical hardware failure and continue serving traffic. Customers with projects in both us-central1-b and us-central1-f would not have seen multi-zonal impact - at most one of their zones would have capacity in the affected data center.

90-Day Uptime: 90.00% Checked: just now

Azure

Operational

Everything is looking good

90-Day Uptime: 100.00% Checked: just now

OpenAI

Operational

All Systems Operational.

90-Day Uptime: 93.00% Checked: just now

Anthropic

Operational

All Systems Operational.

90-Day Uptime: 96.56% Checked: just now

Hugging Face

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Replicate

Operational

Under Maintenance.

90-Day Uptime: 98.72% Checked: just now

Stripe

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

PayPal

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Shopify

Operational

All Systems Operational.

90-Day Uptime: 98.22% Checked: just now

Discord

Operational

All Systems Operational.

90-Day Uptime: 99.39% Checked: just now

Slack

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Twilio

Degraded

Twilio customers may be experiencing User Authentication Identity Gateway API (the engine that lets different apps talk to each other) failure increase and Gateway API (the engine that lets different apps talk to each other) high response delays for Verizon in the United States. Our team has identified the cause, and is working to resolve the issue. We will provide another update in 1 hour or as soon as more information becomes available.

90-Day Uptime: 95.28% Checked: just now

SendGrid

Degraded

Twilio customers may be experiencing User Authentication Identity Gateway API (the engine that lets different apps talk to each other) failure increase and Gateway API (the engine that lets different apps talk to each other) high response delays for Verizon in the United States. Our team has identified the cause, and is working to resolve the issue. We will provide another update in 1 hour or as soon as more information becomes available.

90-Day Uptime: 98.39% Checked: just now

Mailchimp

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Notion

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Atlassian (Jira/Confluence)

Operational

All Systems Operational.

90-Day Uptime: 99.06% Checked: just now

Figma

Operational

All Systems Operational.

90-Day Uptime: 99.94% Checked: just now

HubSpot

Operational

All Systems Operational.

90-Day Uptime: 99.89% Checked: just now

Zoom

Operational

All Systems Operational.

90-Day Uptime: 100.00% Checked: just now

Supabase

Degraded

We have rolled out the deployment process changes to eu-central-2 and continue to monitor. We will provide additional updates as we make this change to other regions. Once this change is implemented across all regions, we will release a new Supabase version that impacted users can upgrade to from their dashboard. Completing that upgrade will resolve this issue.

90-Day Uptime: 95.28% Checked: just now

Render

Operational

All Systems Operational.

90-Day Uptime: 97.94% Checked: just now

DigitalOcean

Operational

All Systems Operational.

90-Day Uptime: 97.72% Checked: just now

Sentry

Operational

All Systems Operational.

90-Day Uptime: 95.61% Checked: just now

Datadog

Operational

All Systems Operational.

90-Day Uptime: 99.17% Checked: just now

Railway

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Fly.io

Operational

All Systems Operational.

90-Day Uptime: 97.72% Checked: just now

Groq

Operational

All Systems Operational.

90-Day Uptime: 99.94% Checked: just now

ElevenLabs

Operational

All Systems Operational.

90-Day Uptime: 97.33% Checked: just now

Perplexity

Operational

Perplexity is up and working normally.

90-Day Uptime: 100.00% Checked: just now

Midjourney

Operational

Midjourney is up and working normally.

90-Day Uptime: 100.00% Checked: just now

Google Gemini

Major Outage

Google Gemini is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

xAI / Grok

Major Outage

xAI / Grok is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

Plaid

Operational

All Systems Operational.

90-Day Uptime: 98.50% Checked: just now

Adyen

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Square

Operational

All Systems Operational.

90-Day Uptime: 98.50% Checked: just now

Braintree

Major Outage

Braintree is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

Zendesk

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Dropbox

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Salesforce

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Linear

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Microsoft 365

Major Outage

Microsoft 365 is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

Google Workspace

Major Outage

Google Workspace is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

YouTube

Major Outage

YouTube is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

Netflix

Major Outage

Netflix is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

TikTok

Major Outage

TikTok is not reachable. The service may be down or experiencing a major outage.

90-Day Uptime: 90.00% Checked: just now

Instagram

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Facebook

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

WhatsApp

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Spotify

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

Twitch

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now

X (Twitter)

Operational

All Systems Operational

90-Day Uptime: 100.00% Checked: just now


Recent Incidents Feed

Investigation in progress.

Investigation in progress.

Investigation in progress.

Investigation in progress.

Investigation in progress.

Investigation in progress.

Investigation in progress.

Investigation in progress.

Resolved. Services returned to baseline response profiles.

Resolved. Services returned to baseline response profiles.

Why Choose Is It Down?

Official status pages often take 10–30 minutes to declare an outage. Our **Lag Detector** monitors user reports to flag disruptions before official APIs update.

We translate confusing JSON statuses like `degraded us-east-1 queue backlog` into plain-English.

View Reliability Leaderboard

Frequently Asked Questions

IsItDown is a real-time service status aggregator and outage tracker. We monitor API portals, cloud hosting infrastructure, payment gateways, and developer tools to give you an immediate view of internet-wide health.
Official status pages often take 10 to 30 minutes to declare an incident. We combine active API status polling (our Lag Detector) with real-time user report submissions (our Community Outage Pulse) to flag disruptions before they are officially declared.
Technical status logs can be highly cryptic (e.g., 'us-east-1 connection backlog mitigated'). We translate these technical messages into simple, plain-English statements so you know exactly which services are impacted.
We record daily uptime records for all monitored platforms. Operational days score 100% uptime, minor degradations score 95%, and major outages score 60% uptime. The scores are weighted and averaged over the past 90 days.
Yes, you can subscribe to receive email alerts for any service. Navigate to the service's status page, enter your email address, and we will automatically notify you when status changes occur.