We test our code.
We write unit tests, feature tests, integration tests and end-to-end tests. We run static analysis. We review pull requests. We build CI pipelines specifically to stop bad code reaching production.
Then we add an LLM to our application, write a system prompt and...
Hope for the best?
I've been building more AI-powered functionality recently, and this is something I've become increasingly interested in.






