Gemma 4 Remains Weak on Agentic Tasks

ABLPHA · reddit · 2026-07-20

The OP reports that Gemma 4 is still "lazy": even with preservethinking enabled, updated Unsloth GGUFs, and a new chat template, it still halts after just a few tool calls when performing agent tasks in Hermes. It typically outputs something like "I did A, but B/C happened, so next I'll do D" and then stops.

In contrast, models like Qwen 3.6 27B, DeepSeek V4 Flash, and GPT-OSS 120B can continuously advance tasks. Gemma 4 feels more suited for chatting and world-building discussions rather than multi-turn agentic workflows. The post essentially asks if Gemma 4 is inherently unfit for this type of task.

Original post →

More from coding & agent

coding & agent channel →