Legal case-hallucination benchmark shows Gemma4:26b slightly beating Claude
MRGWONK · reddit · 2026-09-24
The author built an MCP tool for legal research (syfert.com/mcp) and a case hallucination benchmark around it, finding Gemma4:26b slightly outperforming Claude Fable 5.1 on legal research tasks.
More from Models
- Boltzbit's BAST paper claims LLMs can learn 1000x faster by generating weights from live data — ahuja_priyank · 2026-09-24
- Fireworks launches Ember-1: Kimi K3-based model cuts reasoning tokens by ~40% — sophiamyang · 2026-09-24
- Liquid AI's DSpark speculative decoding makes LFM2.5-VL-3B up to 3.13x faster — helloiamleonie · 2026-09-24
- Users can't get ChatGPT to brighten a photo without regenerating faces — Objective-Bit-797 · 2026-09-24
- Opus 5.5 builds a 40-second animation and explains how it did it — zck · 2026-09-24
- DeepSeek Harness Quietly Ships Official GUI Client for Windows and Mac — vista8 · 2026-09-24