DevOps/CloudHigh signalEmergingAround since 2011

Chaos Engineering

Also written as Chaos Monkey

The practice of deliberately injecting failures into a production or production-like system (killing a server, cutting network access) to test how well it holds up, before a real failure happens unplanned.

Think of it like

Like a fire drill that actually sets off a small controlled fire, to see whether the sprinklers and exits really work — instead of hoping they do when it counts.

Junior or senior?

Junior sounds like

Talks about resilience in the abstract.

Senior sounds like

Has actually broken something in production or staging on purpose to test it, and can describe what they learned.

Ask them

“Has your team ever deliberately broken something in production (or staging) to test resilience? What did you learn?”