Poolside made its first models available to a broader audience in April 2026 with Laguna M.1 and XS.2. Until then, the company had focused on government and public-sector customers. XS.2 was also its first open model under the Apache 2.0 license. Laguna S 2.1 is the third version in the series released in roughly three months.

Laguna S 2.1 beats much larger open models

With thinking enabled, Laguna S 2.1 scores 70.2 percent on Terminal-Bench 2.1, which tests models on long-running terminal tasks. It ranks just behind Tencent's Hy3 (295B-A21B) and ahead of much larger open models, including DeepSeek-V4-Pro-Max, Nemotron 3 Ultra, and Thinking Machines Lab's debut model. The overall leaderboard is led by OpenAI's GPT-5.6 Sol, Anthropic's Claude Fable 5, and Kimi K3.

On Terminal-Bench 2.1, Laguna S 2.1 lands in the upper midfield, just behind leading closed-weights models and ahead of many larger open-weights systems. | Image: Poolside

Poolside says Datacurve's DeepSWE benchmark offers a better comparison because its scores are spread across a wider range. Laguna S 2.1 scores 40.4 percent, while some open-weight models with more than one trillion parameters remain below 10 percent. It also ranks near the top of its class on SWE-Bench Multilingual, SWE-Bench Pro, and SWE Atlas.