Skip to content

Grafana Labs outage history and postmortems

Data and observability · 74 incidents in 90 days · 99.91% implied uptime · last 2d ago

Status page Compare

Incidents

270 on file
StartedIncidentDuration
Sep 2221:27 UTCKubernetes Observability Billing & Usage Incorrect16h 39m
Sep 2214:57 UTCIRM Access Issues for a Small Group of Users20h 10m
Sep 2119:16 UTCMimir write request errors0m
Sep 1712:34 UTCElevated Latency Managing Cloud Provider Integrations in GCP US Central0m
Sep 1510:00 UTCIncreased Execution Time for Browser Checks in Synthetic Monitoring1h 17m
Sep 1507:49 UTCIntermittent Metric Write Errors in GCP US Central (prod-us-central-0)0m
Sep 1413:04 UTCCloudWatch request timeouts in prod-us-central-00m
Sep 1223:00 UTCFailures Submitting Self-Serve Configuration Changes for Grafana Cloud Logs in AWS Sweden (prod-eu-north-0)0m
Sep 911:27 UTCInvestigating elevated database load in AWS Germany2h 24m
Sep 818:05 UTCIssues with Geomap Tiles4h 36m
Sep 816:14 UTCHigh Latency in prod-us-east-24h 19m
Sep 417:03 UTCDegradation of Hosted Grafana in US Central Region3h 39m
Sep 416:45 UTCUS Central Region Instability6h 15m
Sep 413:36 UTCAlert rule creation, deletion, and update degradation in prod-us-east-22h 56m
Sep 322:13 UTCTempo and Mimir Read and Write Failures2h 42m
Sep 309:32 UTCHigh Latency in prod-ap-south-13h 17m
Sep 214:14 UTCElevated Latency in Grafana Cloud Logs (prod-ap-south-1)0m
Sep 115:25 UTCInvestigating issues in US Central (prod-us-central-0, prod-us-central-5)4h 54m
Aug 2817:12 UTCPartial Logs Write Outage1h 29m
Aug 2800:58 UTCSome Grafana UI features may be unavailable or reverting to legacy behaviour1h 42m
Aug 2712:09 UTCElevated error rates affecting metrics writes in prod-us-central-057m
Aug 2701:54 UTCMimir Writes Incident in prod-us-central-00m
Aug 2510:39 UTCIncident Management unavailable in US Central23m
Aug 1822:53 UTCCloud Logs read path outage on eu-west-20m
Aug 1315:25 UTCK6 Test Outage1h 22m
Aug 1013:30 UTCMetrics: Elevated Error Rates Reads/Writes0m
Aug 819:48 UTCAlerting expressions pipeline failing when recovery settings0m
Aug 517:57 UTCSome Cloud Test Runs Terminated0m
Aug 511:41 UTCo11y requests too large for nods in the prod-eu-west-2 region.3m
Aug 419:11 UTCSome Grafana Instances Unavailable6h 21m
Aug 410:34 UTCLogs latency increase within prod-eu-west-33h 15m
Aug 107:25 UTCK6 - Cloud test-run issues1h 6m
Jul 3120:33 UTCPartial Read Outage for Loki in prod-us-east-40m
Jul 3110:51 UTCDegraded Performance: Stack Provisioning Failures within certain reigons (PDC Setup)1h 22m
Jul 3014:45 UTCPDC Authentication Issues2h 55m
Jul 3013:15 UTCIssues with Billing/Usage Dashboard Metrics and Panels.2h 2m
Jul 2921:46 UTCIRM Performance Degradation in EU Region2h 53m
Jul 2918:49 UTCGrafana Cloud non-billing usage metrics gaps in select regions0m
Jul 2819:40 UTCPartial OTLP Write Outage in prod-us-east-33m
Jul 2417:02 UTCWrite Outage1h 18m
About Grafana Labs

The last year, day by day

No incident Minor Major CriticalDeeper colour: longer outage

Downtime, added up

Data and observability median
2h 52m

Grafana Labs: 2h 1m weighted downtime in range, category median 2h 52m.

74 incidents, 24 major or critical, 15d 19h of incident time (4d major or critical, 11d 19h partial), longest 6d 2h, implied uptime 99.91%, between Jun 28, 2026 and Sep 25, 2026.

Stated causes

262 incidents with no cause stated.

Across fru.dev:Company profilePaydaysChangefeedUptime

Questions

Is Grafana Labs down right now?

Check Grafana Labs's own status page at https://status.grafana.com for the live state. This page shows Grafana Labs's incident history from that page; the most recent incident started 2026-09-22: "Kubernetes Observability Billing & Usage Incorrect".

How often does Grafana Labs have outages?

Grafana Labs posted 74 incidents in the last 90 days and 270 in the last year, 24 of them major or critical in the last 90 days. The median incident lasted 1h 54m.

When was Grafana Labs's last major outage?

2026-09-22: "IRM Access Issues for a Small Group of Users", lasting 20h 10m.

What usually causes Grafana Labs incidents?

Where Grafana Labs explained the cause, the most common was third-party dependency (3 incidents). Most status updates do not name a cause.

Sources: vendors' own status pages, published postmortems and SEC 8-K Item 1.05 filings, read daily. Times as reported. Logos via logo.dev; trademarks belong to their owners.

Outages by email

Saturday mornings: the week's major outages, new postmortems and disclosed breaches, only in weeks that had some.

Double opt-in. Unsubscribe any time.