Nanbeige4.2-3B claims 63.6 on SWE-Bench Verified with a Looped Transformer
UsedMorning9886 · reddit · 2026-07-22
Nanbeige Lab released Nanbeige4.2-3B, a compact agentic model built on a Looped Transformer architecture.
The model has 3B non-embedding parameters and the post claims strong agentic results, including 63.6 on SWE-Bench Verified and 46.9 on SWE-Bench Pro. The author says it outperforms larger Qwen3.5 and Gemma 4-class models on coding and daily workflow benchmarks when tested in OpenClaw with Lyzr Control Plane, while noting that the benchmark claims should be treated cautiously until independently verified. The release is pitched as a lightweight local personal assistant.
Related event: Nanbeige4.2-3B Released with Looped Transformer for Agent Capabilities(5 posts)→
More from coding & agent
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- Codex tip: use Sol with Astra and Luna sub-agents to save usage — pvncher · 2026-09-11
- agents-best-practices: a provider-neutral Agent Skill for designing and auditing agentic harnesses — tom_doerr · 2026-09-11
- Cognition's SWE-2 uses a KKT duality argument in RL to shift the effort Pareto curve — YouJiacheng · 2026-09-11
- First-ever Three.js Conference lands in Paris, with a panel on AI-shortened design workflows — OdinLovis · 2026-09-11