Hints: service-down¶
Hint 1¶
Check the basics: are pods running?
- kubectl get pods -n grokdevops
- kubectl get events -n grokdevops --sort-by='.lastTimestamp' | tail -10
Hint 2¶
The issue is likely at the pod or deployment level. Check:
- Is the deployment scaled to 0? (kubectl get deployment -n grokdevops)
- Are pods in a crash loop? (look at STATUS and RESTARTS columns)
- Are pods stuck in Pending? (check node resources or scheduling constraints)
Hint 3¶
Look at the deployment spec and recent Helm changes:
- helm history grokdevops -n grokdevops
- kubectl get deployment grokdevops -n grokdevops -o yaml | head -60
- Was there a recent change to image, command, or resource limits?
Hint 4¶
Most likely causes: wrong image tag, broken probe, resource exhaustion, or a failed Helm upgrade. Check the specific error in kubectl describe pod and cross-reference with the matching runbook.