I’m a backend engineer, and my cofounder is a doctor training in emergency care.

Rounds began with something she kept returning to in our conversations. An exam gives you the relevant information. A patient gives you an opening complaint, and you decide what to ask, what to examine, which investigations matter, and when you know enough to commit.

We wanted to simulate that reasoning process. I assumed the conversational patient would be the easy part.

The first prototype felt impressive for about five minutes.

Then we questioned it more aggressively. Ask about the same symptom twice and part of the history might change. Request a troponin and the model could invent a perfectly plausible value. Phrase a leading question carefully enough and the patient might hand over the diagnosis.