An organization has a large and complex infrastructure which often leads to incidents that impact users.
Which is the BEST technique to use to improve this situation?
Large and complex environments often fail in unexpected ways. Chaos engineering is a technique used to improve resilience by deliberately testing how systems behave under failure conditions. This helps organizations expose weaknesses, improve recovery mechanisms, and reduce user-impacting incidents over time.
A is not the best fit for reducing operational fragility. C improves quality in delivery pipelines, but it does not directly test operational resilience in complex live-like environments. D is a useful team agreement on completion criteria, but it is too limited for this scenario.
B is best because chaos engineering directly targets the uncertainty and fragility that often come with complex infrastructures, which is very consistent with HVIT's resilience focus.
Currently there are no comments in this discussion, be the first to comment!