You are on-call for an infrastructure service that has a large number of dependent systems. You receive an alert indicating that the service is failing to serve most of its requests and all of its dependent systems with hundreds of thousands of users are affected. As part of your Site Reliability Engineering (SRE) incident management protocol, you declare yourself Incident Commander (IC) and pull in two experienced people from your team as Operations Lead (OLJ and Communications Lead (CL). What should you do next?
Willard
10 months agoHubert
10 months agoLeigha
11 months agoTy
11 months agoMargery
11 months agoMauricio
11 months agoDarrel
11 months agoDaniel
11 months agoCarmen
11 months agoSharee
11 months agoDexter
11 months agoMiss
11 months ago