Gemma 4 dev-agent comp locks everyone to one 31B model — is the code graph the intended edge?

politefella0 · reddit · 2026-09-26

The Gemma 4 dev-agent competition allows only gemma-4-31b-it-qat-w4a16-ct for every agent and subagent, with optional per-agent PEFT LoRA adapters, prompts, sub-agents, and sandboxed skills. Scoring is SWE-Bench-style PASS/FAIL within a 12-hour total budget including sandbox setup. Over 4,000 entrants and 600+ submissions are in; it closes Dec 2.

The author's key observation: the harness provides getcodeneighbors, searchsimilarcode (cosine over pre-computed embeddings), and getcodesubgraph, while the paper track explicitly calls out Graph Reasoning and Code Comprehension with a released graph+embedding dataset — strong hints that graph-based repo comprehension is the intended winning path.

Three concrete debates follow:

Original post →

More from coding & agent

coding & agent channel →