Built for the WeMakeDevs × TrueFoundry Agent Harness Hackathon.
There are two kinds of "AI for incident response," and both of them are wrong.
The first acts on its own. It sees latency spike, decides it knows why, and rolls back your deploy at 3am. When it's right, it's magic. When it's wrong — and it will be wrong, because production is where confident reasoning goes to die — you now have two incidents.
The second just pages a human and summarizes some logs. Safe, and nearly useless. The twenty minutes of mechanical work still land on the person who just woke up.
I wanted to find out whether you could have the first one's speed with the second one's safety. Not by making the model more careful — you cannot prompt your way to a safety guarantee — but by making the unsafe action structurally impossible until a human says yes.






