Inkling Hands-On Test Fails to Impress
emollick · x · 2026-07-16
After testing the new open-weight model Inkling, Emollick found its performance unstable in his tests: even when set to xHigh, it struggled to work reliably, with its chain of thought 'diverging' on simple requests.
Replies also noted that Inkling's current performance falls far short of some frontier Chinese open-source models, pointing out that it even failed the Lem Test, a test that has supposedly been passed by all frontier models since DeepSeek r1 / Sonnet 3.5.
Related event: New Open-Source Model Inkling Falls Short in Tests(2 posts)→
More from Models
- Looking for a classifier of software engineering task shapes to pick models per task — StewartalsopIII · 2026-09-11
- DeepSeek V4 Pro API to continue after Sept 2026, billing unchanged — teortaxesTex · 2026-09-11
- DeepSeek V4.1 Flash Hits 98% of GPT-6 Astra's Score at 1.4% of the Cost in Third-Party Benchmark — ayushtweetshere · 2026-09-11
- TheZvi Polls: Has Your Coding Model Choice Changed Since Fable 5.1 and Astra? — TheZvi · 2026-09-11
- antirez Weighs In on Anthropic Banning Minors From Using Claude — antirez · 2026-09-11
- Engram's random reads don't suit SSDs; CPU-memory over NVLink could serve all 72 GPUs — bookwormengr · 2026-09-11