Chaos engineering sounds expensive. Netflix built Chaos Monkey to randomly kill production servers. Google runs DiRT (Disaster Recovery Testing) across their entire infrastructure. Amazon does game days where they intentionally take down services.

You're building a Node.js API. You don't have a platform team. You don't have a chaos infrastructure. But you still need to know: what happens when your dependencies get slow?

The good news is that 80% of the value of chaos engineering comes from one question, and you can answer it locally in five minutes.

The one question that matters

What does my application do when a dependency responds slowly or not at all?