We test our code.

We write unit tests, feature tests, integration tests and end-to-end tests. We run static analysis. We review pull requests. We build CI pipelines specifically to stop bad code reaching production.

Then we add an LLM to our application, write a system prompt and...

Hope for the best?

I've been building more AI-powered functionality recently, and this is something I've become increasingly interested in.