35B Open-Weight Model Tops Code Search at 100x Lower Cost Than Frontier Models
ypatil125 · x · 2026-09-05
Applied Compute partnered with turbopuffer to RL post-train Qwen3.6-35B-A3B to search code across 9,000 GitHub repos via a precomputed index. The small open-weight model tops the needle-in-a-haystack task outright at 2-10x lower latency and roughly 100x lower cost than frontier models, showing that targeted post-training can make small models viable search agents for large codebases.
Related event: RL-Trained 35B Open Model Cuts Code Search Costs 100x(2 posts)→
More from coding & agent
- Replacing Custom API Integrations with an MCP Bridge: Managing Ad Campaigns via Hermes and Discord — emberlitpublishing · 2026-09-05
- Bolt launches parallel task execution for faster dev workflows — tristanbob · 2026-09-05
- Failure modes found only by running coding agents unattended for months — Fragrant_Yoghurt1135 · 2026-09-05
- Open-source Claude Code course goes viral: 15 modules from first project to production — adnan_hashmi · 2026-09-05
- Manager agent + workers in git worktrees: orchestration that survived overnight runs — Fragrant_Yoghurt1135 · 2026-09-05
- Agent Substrate brings instant suspend/resume and 10x density to K8s AI agents — davemccollough · 2026-09-05