Ornith-1.5 released: Self-improving loop matches Claude Opus 4.8 performance

alexcovo_eth · x · 2026-08-20

DeepReinforce released Ornith-1.5, a family of open-source models (9B Dense / 35B MoE / 397B MoE) featuring a revolutionary self-improvement training loop (self-proposing tasks, scaffolding, RL rollouts). It achieves SOTA results on Terminal-Bench (86.1) and SWE-Bench Verified (86), rivaling Claude Opus 4.8. The 35B model outperforms Qwen 3.6, while the 9B quantized version runs on phones at just 1.5GB. Supports FP8/GGUF/MLX under MIT license.

Related event: Open-Source Ornith-1.5 Family Launches, 397B MoE Rivals Claude Opus 4.8(24 posts)→

Original post →

More from Models

Models channel →