Extending Jev Mode to Images: Constrained llama.cpp Outputs as Image Selections
opUserZero · reddit · 2026-09-29
Codacus built Jev mode (constrained decoding) for llama.cpp; the author extended the idea to images with an agent plus harness — the constrained answer is directly an image selection, skipping the decode step entirely: no captioning pause, just a decision based on one image or a group.
Use cases: ask the same question over a batch of images for classification, or hand the model 20 images and have it pick the one containing a rubber duck. PR and a YouTube explainer (made with Codacus's RenderDiv framework) are linked.
More from coding & agent
- Millions of person-hours wasted building AI harnesses, erased by new model releases — sebpaquet · 2026-09-29
- Unverified Claim: Anthropic Engineers Share All Claude Sessions, Teammates Can Steal Each Other's Tasks — YouJiacheng · 2026-09-29
- Open-source self-driving sim repo auto-galleries 38 demos, adds Claude Code PR review skill — 4310sy · 2026-09-29
- Celesto: open-source persistent microVM computers for AI agents, boots in 500ms — aniketmaurya · 2026-09-29
- A practical guide to adopting AI in your organization: management skills over prompt tricks — chribonn · 2026-09-29
- RSI Arena: 8 AI agents get 1,000 GPU-hours each to train a better model live — my_cat_can_code · 2026-09-29