Health checks flap
After a deploy, a service's instances keep flapping in and out of the load balancer's healthy pool — the dashboard shows hosts being marked unhealthy, removed, then re-added a minute later, over and over. User-facing error rates are spiky as capacity bounces. The new build added a deeper health check that also pings a downstream dependency. How do you triage the flapping?
What a strong answer looks like
Stop the bleeding first (mitigate), then form hypotheses from real signals. Separate root cause from symptom, communicate status as you go, and close with what prevents a repeat.
0:00 of about 20 min
Which questions mattered is sealed until you submit. Telling you now would just be handing over the edge cases.
Run or narrate your approach, then ask the coach.