Unsloth's Gemma 4 mmproj Breaks with Newer llama.cpp, Causing Silent Multimodal Failures

Top_Speaker_7785 · reddit · 2026-08-06

A developer found that Unsloth's Gemma 4 mmproj file became incompatible with newer llama.cpp builds, causing multimodal features (image analysis, voice transcription) to silently fail, outputting <unused49> garbage tokens. The root cause is third-party quantizers not staying in sync with llama.cpp's internal format. Switching to ggml-org's official GGUF fixed the issue. The post details the debugging process, root cause analysis, and fixes, and raises community questions.

Original post →

More from coding & agent

coding & agent channel →