Most agent frameworks give each agent its own context window and call it memory. That works right up
until you run more than one agent, and then it quietly becomes the most expensive design decision in
the system.
We run a fleet where different agents are deliberately backed by different models — one family
handles long-form drafting, another handles structured extraction, a couple run on a local path with







