Dev builds adaptive harness that auto-scales LLM reasoning effort up and down
milpster · reddit · 2026-09-06
A Reddit developer proposes an alternative to manually fixing thinking-effort levels: modify your harness so the system adaptively raises or lowers reasoning effort between low and xhigh based on success and failure streaks.
The posted logs show it downshifts on "stable successful streaks" and escalates on "meaningful failures", combined with phase tracking (EXPLORE/DEBUG/IMPLEMENT). Subjectively it speeds up problem solving considerably, but no reliable quality evaluation yet — the author is considering a GPQA Diamond before/after test.
More from coding & agent
- Juggling ChatGPT, Claude and agents, users hunt for a unified AI memory setup — utkuaytac · 2026-09-06
- A 3-person team can't make multi-agent coding work: three context-sharing attempts failed — Practical_Sink401 · 2026-09-06
- Villager sim game built locally on 16GB VRAM with Qwen3.8-27B Q3 quant at 75 tok/s — Fancy-Snow7 · 2026-09-06
- mitsuhiko uses AI agent to build Java-style virtual threads for Python — mitsuhiko · 2026-09-06
- Two-person team spent 12 days and $400+ credits making a 6-min AI short film with Codex + Seedance 2.5 — APPSO · 2026-09-06
- AI Coding Will Prevent Expertise: The 'Expert Novice' Paradox in Dev Skills — bibryam · 2026-09-06