An SRE team spends 30% of their time manually restarting crashed microservices. Which automation approach BEST eliminates this toil?
-
A
Implement liveness probes and automatic pod restarts in Kubernetes
-
B
Document the restart procedure in a runbook
-
C
Create a Slack alert for on-call engineers
-
D
Schedule weekly manual health checks