FLock ships 1.88B THIS/THAT model 1.2, claims 87.8% on its own decision test
matlabulous · x · 2026-09-24
FLock released THIS / THAT model 1.2, a tiny 1.88B-parameter model that generates zero tokens and scored 87.8% on the company's own "super complex decision test" — up from 40.6% (v1.0) and 77.5% (v1.1) — claiming it now beats Claude Opus 5 and GPT-5.6 on that benchmark. The model is free via the FLock API. Note the comparison is on a proprietary in-house test, not public benchmarks.
More from Models
- Why Competing With Meta's Muse Is Hard: 40K Ratings at 4.88, Plus Distribution — FinanceYF5 · 2026-09-24
- OpenAI's MentalHealthBench: GPT-6 Astra Scores 57.3 vs GPT-4o's 32.1 — rohanpaul_ai · 2026-09-24
- Opus 5.5 Shows Off UI Flair: Checkmark Icons and Menu Polish on Its Own — vista8 · 2026-09-24
- Claude Opus 5.5 roasts every AI model and makes the whole video itself — bookwormengr · 2026-09-24
- Leaked naming: GPT-6 Sol equals GPT-5.6 Terra, Luna degradation confirmed — PawelHuryn · 2026-09-24
- Viral 'Opus 5.5 update' post fuels Anthropic release speculation — rudrank · 2026-09-24