Nanbeige4.2-3B claims 63.6 on SWE-Bench Verified with a Looped Transformer
UsedMorning9886 · reddit · 2026-07-22
Nanbeige Lab released Nanbeige4.2-3B, a compact agentic model built on a Looped Transformer architecture.
The model has 3B non-embedding parameters and the post claims strong agentic results, including 63.6 on SWE-Bench Verified and 46.9 on SWE-Bench Pro. The author says it outperforms larger Qwen3.5 and Gemma 4-class models on coding and daily workflow benchmarks when tested in OpenClaw with Lyzr Control Plane, while noting that the benchmark claims should be treated cautiously until independently verified. The release is pitched as a lightweight local personal assistant.
Related event: Nanbeige4.2-3B Punches Above Its Weight in Agentic Coding(2 posts)→
More from coding & agent
- Voice agents are pushing LangChain tracing and observability beyond text workflows — Hacubu · 2026-07-22
- DeepChat open-sources a local-first desktop client for multi-model agent workflows — Roger_M_Taylor · 2026-07-22
- Grok Build Update: Speech-to-Text Lets You Brief Agents by Voice — tetsuoai · 2026-07-22
- Codex’s repeated resets become a running joke about product addiction — tinyfool · 2026-07-22
- Tencent's Design Agent Platform Miora Goes Public: Generating Brand VI in Minutes — 数字生命卡兹克 · 2026-07-22
- Agents need realistic environments for API keys, security hurdles and RL hacks — 1a3orn · 2026-07-22