Originally published on the InsForge blog, written by Tony Chang (CTO & Co-Founder). Reposted here with permission.

In December we published the first MCPMark benchmark results comparing InsForge MCP, Supabase MCP, and Postgres MCP across 21 real-world database tasks using Claude Sonnet 4.5. InsForge came out ahead on accuracy, speed, and token efficiency.

We reran the benchmarks. This time on Claude Sonnet 4.6, the latest model from Anthropic. InsForge MCP achieves 28% higher Pass⁴ accuracy while using 2.4x fewer tokens than Supabase MCP. The efficiency gap has widened.

Updated Results: Claude Sonnet 4.6

Same 21 MCPMark Postgres tasks, 4 runs per task, strict Pass⁴ scoring.