Search posts, tags, users, and pages
Prateek Srivastava
SRE & DevOps Engineer | Kubernetes, Multi-Cloud, Linux | Writing about real production systems
⚡ TL;DR: An EKS worker node running 30+ pods crashed at 08:43 IST on Aug 16. EC2 Auto Scaling didn't act. EKS Node Auto Repair was never enabled. Cluster Autoscaler had no headroom (Min=Max=6). Our me
No responses yet.