Code RoomRegional quota exhaustion on surge
HardPrep Room Coding #2846

Regional quota exhaustion on surge

On-callReliability & on-callSenior–Staff~30 min

You run the same service in three regions behind latency-based DNS. During an outage at a peer CDN, traffic shifts and us-east-1 takes a 3x surge while eu-west-1 and ap-southeast-1 stay flat. us-east-1 fails to scale: new instances launch but a chunk never reach service and the ASG log shows a service-quota error. The other two regions are healthy and well under the same nominal limits. Your global capacity dashboard (which aggregates across regions) shows plenty of headroom. How do you triage and mitigate?

What a strong answer looks like

Stop the bleeding first (mitigate), then form hypotheses from real signals. Separate root cause from symptom, communicate status as you go, and close with what prevents a repeat.

0:00 of about 30 min
Which questions mattered is sealed until you submit. Telling you now would just be handing over the edge cases.