Open-weight models cut the cyber gap to 4–7 months as Kimi K3 hits demand limits
rohanpaul_ai · x · 2026-07-21
Rohan Paul’s daily newsletter rounds up several notable AI items:
- Open-weight models now trail the closed frontier by only 4–7 months on long-horizon cyber capability, down from 6–10 months for much of 2025.
- An AI-advice study found that model suggestions can make people less willing to say “I don’t know”, even when accuracy is incentivized.
- Kimi K3 reportedly nearly doubled its nearest rival, Claude Fable 5, on a demanding benchmark for autonomous legal work.
- Kimi K3 also fixed 15 critical security bugs that Codex and Fable reportedly refused because of cyber guardrails.
- The newsletter highlights a framing that AI agents fail first through broken context, not isolated execution errors.
- It also notes that U.S. serving providers like Modal, Fireworks, and Baseten can host Kimi K3 at roughly one-tenth the cost of Chinese competitors thanks to access to Nvidia and AMD chips.
- Demand is so strong that new Kimi K3 subscriptions are currently blocked.
- Finally, it points to an essay arguing why one OpenAI senior employee thinks open-weight models are decelerationist.
More from Models
- Newer models need a different prompting style, and old tricks can make outputs worse — emollick · 2026-07-21
- GLM-5.5 is said to arrive in 4 weeks with open weights — tanay_mehta · 2026-07-21
- Fable 5 is credited with a 3-variable counterexample to the Jacobian conjecture — Various-Affect4841 · 2026-07-21
- Ben’s Bites roundup highlights Kimi K3, Fable 5, Cursor costs and self-driving companies — Ben's Bites · 2026-07-21
- Apple-π benchmark asks whether video models reason about physical laws or just mimic motion — liuziwei7 · 2026-07-21
- OpenAI models look best at drawing and understanding complex SVG diagrams — mimi10v3 · 2026-07-21