Gemma 4 Remains Weak on Agentic Tasks
ABLPHA · reddit · 2026-07-20
The OP reports that Gemma 4 is still "lazy": even with preservethinking enabled, updated Unsloth GGUFs, and a new chat template, it still halts after just a few tool calls when performing agent tasks in Hermes. It typically outputs something like "I did A, but B/C happened, so next I'll do D" and then stops.
In contrast, models like Qwen 3.6 27B, DeepSeek V4 Flash, and GPT-OSS 120B can continuously advance tasks. Gemma 4 feels more suited for chatting and world-building discussions rather than multi-turn agentic workflows. The post essentially asks if Gemma 4 is inherently unfit for this type of task.
More from coding & agent
- Cheaper OpenAI Agents API alternative: sandbox service undercutting E2B by 46% — airesearch12 · 2026-09-11
- His agent kill switch ran for months before he found it was wired to nothing — AnvilandCode · 2026-09-11
- Kernel's Browser Agents Can Now Pay Online Using Aliases, Never Touching Card Data — jeff_weinstein · 2026-09-11
- OpenAI opens up agent sandboxes: BYO or pick from Cloudflare, E2B, Modal, Vercel and more — threepointone · 2026-09-11
- SocialCrawl MCP lets agents search Reddit, YouTube, TikTok, X with one API key — dooddyman · 2026-09-11
- Astra builds a surprisingly polished Catan game in three.js, reusing past UI and 3D assets — FinanceYF5 · 2026-09-11