Open-source Ornith-1.5 hits SOTA with self-improving strategies
aftahi_ai · x · 2026-08-19
Ornith-1.5, a family of open-source LLMs (9B Dense, 35B MoE, 397B MoE), has been released with a focus on end-to-end self-improvement. The models not only solve tasks but also generate new tasks and build their own scaffolds.
It achieves state-of-the-art performance on several benchmarks:
- SWE-Bench Verified: 86 (a significant jump)
- Terminal-Bench 2.1: 86.1
- Tool Decathlon: 71.2
The author claims performance comparable to Claude Opus 4.8 across reasoning, agentic, and coding tasks.
More from Models
- Ling-3.0 Open Source: 6 Base Checkpoints Released with WSM Training Method — bclavie · 2026-08-19
- Grok 4.6 lands on Amazon Bedrock with 500k context and configurable reasoning — SpaceXAI · 2026-08-19
- empero-ai's Qwen3.8-9B-Distill trends on Hugging Face — empero-ai · 2026-08-19
- z-lab's Qwen3.8-27B DFlash2: block-diffusion draft model for fast inference — z-lab · 2026-08-19
- GGUF quantized Qwen3.8-9B-Distill lands for llama.cpp local runs — empero-ai · 2026-08-19
- Replit Free Mode powered by GPT-5.6 Luna: $20/mo enough to deploy 10+ apps — amasad · 2026-08-19