Poolside Releases Laguna S 2.1: A Small Coding Model Outperforming 1T+ Parameter Giants
mervenoyann · x · 2026-07-22
Poolside has announced Laguna S 2.1, claiming it is the most capable agentic coding model in its weight class.
On the Terminal-Bench 2.1 benchmark, the model scored 70.2, competing directly with models 5–25x its size and outperforming several of them. On the highly challenging long-horizon DeepSWE benchmark, Laguna S 2.1 scored 40.4, beating multiple open-source models with over 1 trillion parameters. The company has also released the full trajectory of every trial in the final evaluation set on GitHub for transparency.
Related event: Poolside Releases 118B Open-Weight Coding Model Laguna S 2.1(35 posts)→
More from Models
- Meta's Muse Agent has built-in invite code logic, hinting at free-usage expansion — testingcatalog · 2026-09-11
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11