Poolside releases Laguna S 2.1, a 118B open model with 1M-token context
willccbb · x · 2026-07-22
Poolside released Laguna S 2.1, its most capable model so far.
Key details:
- 118B total parameters, 8B active per token
- Up to 1M token context window
- Supports both thinking and non-thinking modes
- The team says it can compete with models many times larger while still fitting on a single NVIDIA DGX Spark
- Weights are available now under OpenMDW-1.1 on Hugging Face
The image also shows benchmark comparisons where Laguna S 2.1 leads several coding and multilingual evaluation suites, including Terminal-Bench, SWE-Bench Multilingual, SWE-Bench Pro, DeepSWE, SWE Atlas, and Toolathlon Verified.
Related event: Poolside Releases 118B Open-Weight Coding Model Laguna S 2.1(35 posts)→
More from Infra
- 80% of the DIY LLM inference hype posters have already quit — it's brutally hard systems work — abhijithneil · 2026-09-11
- Hugging Face's Ultra Scale Playbook: a free book on training LLMs on GPU clusters — mdancho84 · 2026-09-11
- Is inference latency becoming the biggest bottleneck for production AI agents? — Euphoric_Sea632 · 2026-09-11
- LLM Serving Metrics Thread: Why TPOT and Uptime Make or Break User Experience — abhijithneil · 2026-09-11
- PlanetScale launches sharded Postgres: 768 servers acting as one, 1PB scale — dhruv2038 · 2026-09-11
- Can a 7900 XTX 24GB run Qwen locally? Reddit seeks ROCm tok/s benchmarks — thenomadexplorerlife · 2026-09-11