Skip to content
Cloudera · Data and observabilityMar 3, 2025, 22:12 UTC 1 year ago

EU Control Plane Service Disruption

CriticalDeploymentUpdated 3h ago
Mar 3, 22:12 UTCMar 4, 00:40 UTC
Duration
2h 28m
Impact
Critical
Root cause
Deployment
Cloudera, 90 days
1 incidents
Affected
Cloudera Management ConsoleCloudera Data Platform (EU) - CDP Management Console
Status page

Final update

On March 03, 2025, a service disruption occurred within our EU Control plane. This disruption resulted from a configuration change implemented during a routine production cluster upgrade, which inadvertently triggered a synchronization error, leading to service unavailability. ‌ Upon detection of the issue, our engineering teams promptly allocated additional resources to the affected cluster. This action facilitated the restoration of services to an operational state. The root cause of the disruption was the result of an oversight during the standard upgrade procedure, which resulted in the synchronization failure. ‌ We sincerely apologize for any inconvenience this service disruption may have caused. We have implemented corrective measures and refined our upgrade protocols to mitigate the risk of similar incidents in the future. We are committed to maintaining the highest standards of service reliability and appreciate your understanding.

Timeline

  1. Postmortem · Mar 19, 16:53 UTC
    On March 03, 2025, a service disruption occurred within our EU Control plane. This disruption resulted from a configuration change implemented during a routine production cluster upgrade, which inadvertently triggered a synchronization error, leading to service unavailability. ‌ Upon detection of the issue, our engineering teams promptly allocated additional resources to the affected cluster. This action facilitated the restoration of services to an operational state. The root cause of the disruption was the result of an oversight during the standard upgrade procedure, which resulted in the synchronization failure. ‌ We sincerely apologize for any inconvenience this service disruption may have caused. We have implemented corrective measures and refined our upgrade protocols to mitigate the risk of similar incidents in the future. We are committed to maintaining the highest standards of service reliability and appreciate your understanding.
  2. Resolved · Mar 4, 00:40 UTC
    Current Status: Our teams have successfully deployed a fix for the issue and confirmed that the issue has been resolved. If you are still experiencing issues or have any questions please raise a support case with us. A root cause analysis (RCA) will be published within seven business days. Customer Experience: During this window customers may experience issues logging into the console and potential slowness accessing some services.
  3. Monitoring · Mar 4, 00:27 UTC
    Current Status: Our teams have successfully identified the source of the issue and have implemented a solution, which is currently under monitoring. Should you continue to experience issues logging into the console, we kindly request that you submit a support case to us for further assistance. We will keep you updated once we confirm that the issue is resolved on our end. Customer Experience: During this window customers may experience issues logging into the console and potential slowness accessing some services.
  4. Identified · Mar 3, 23:26 UTC
    Current Status: Our teams are actively working on a permanent solution to fully restore the service. Please expect another update in 60 mins. Customer Experience: During this window customers may experience issues logging into the console and potential slowness accessing the experiences.
  5. Identified · Mar 3, 22:12 UTC
    Current Status: Our teams have identified the source of the issue and have partially restored the services. We continue to work on a restoring the service and will have another update in the next 60mins. Customer Experience: During this window customers may experience issues logging into the console and potential slowness accessing the experiences. Incident Start time: 19:58 UTC March 3rd, 2025

More from Cloudera

Full history
StartedIncidentDuration
Sep 323:18 UTC4 weeks agoCloudera Management Console not accessible in US Control Plane3h 29m
May 1118:12 UTC4 months agoIntermittent Performance Issues - US Control Plane0m
Mar 1822:03 UTC6 months agoIntermittent Management Console Access Issues Across US, EU, and AP Regions0m
Sep 2510:23 UTC1 year agoFreeIPA connectivity issues6h 38m
Sep 2419:42 UTC1 year agoIntermittent Performance and Access Issues with the Cloudera Management Console21h 19m
Aug 1314:33 UTC1 year agoDataHubs, DataLakes and FreeIPA are unreachable in US region45m

Also caused by deployment or rollout

All
StartedIncidentDuration
Sep 1106:22 UTC3 weeks agoIssues with Workers VPC hostname route resolution on 2026-09-11Cloudflare0m
Aug 2014:43 UTC6 weeks agoIntermittent failures creating agent tasksGitHub9h 54m
Aug 2000:31 UTC6 weeks ago[Medium] Issue with Microsoft Office Integration and Box EditBox1h 13m
Aug 622:22 UTC8 weeks agoTrouble Using Search Bar For Some AdminsSlack2h 44m
Jul 2220:20 UTC2 months agoSome users may have experienced errors when accessing Amplitude.Amplitude0m
Jul 2110:23 UTC2 months agoJob runs failing at the git clone stepdbt Labs2h 29m

Outages by email

Saturday mornings: the week's major outages, new postmortems and disclosed breaches, only in weeks that had some.

Double opt-in. Unsubscribe any time.