IncidentSeverityMajorStatusResolvedMonitoring
False alerts after cluster memory shortage
–
Last night our Kubernetes cluster reached its memory limit (out of memory). This affected the monitoring system and may have triggered false alerts.
I'm actively working on resolving the issue. Please note that during this time the monitoring may continue to send false alarms that do not reflect an actual problem with your services.
Your services themselves are not affected by this issue. I will let you know as soon as the situation is fully resolved. Thank you for your understanding.
I'm actively working on resolving the issue. Please note that during this time the monitoring may continue to send false alarms that do not reflect an actual problem with your services.
Your services themselves are not affected by this issue. I will let you know as soon as the situation is fully resolved. Thank you for your understanding.
Timeline
Updates on this entry
Platform and monitoring is stable again
The cluster has now been fully rescaled. This required moving several pods around, which temporarily triggered additional notifications from the monitoring service across the various channels.
All services are running stably again and the monitoring is now reporting correctly.
Please accept my apologies for any inconvenience caused, and thank you for your understanding.
All services are running stably again and the monitoring is now reporting correctly.
Please accept my apologies for any inconvenience caused, and thank you for your understanding.