DevOps/CloudHigh signal
Incident Response
Also written as Incident Management
The structured process a team follows when something breaks in production — detecting, communicating, mitigating, and resolving the issue.
Think of it like
Like a hospital's triage process during an emergency — a defined sequence of who does what, so the response isn't improvised in the moment.
Junior or senior?
Junior sounds like
Has been paged into an incident without owning it.
Senior sounds like
Has run point — declaring severity, coordinating responders, communicating status.
Ask them
“Tell me about the worst incident you were part of. What was your specific role in resolving it?”