Most agent frameworks give each agent its own context window and call it memory. That works right up

until you run more than one agent, and then it quietly becomes the most expensive design decision in

the system.

We run a fleet where different agents are deliberately backed by different models — one family

handles long-form drafting, another handles structured extraction, a couple run on a local path with