Cheap LLMs middling at academic null-finding detection, researchers say
RexDouglass · x · 2026-09-19
RexDouglass asks whether Qwen excelled at the task; jonmellon replies that most cheap models they tried were middling at detecting null findings in academic abstracts, showing the task is nontrivial.
More from Models
- Noam Brown: GPT-6 Astra does have observable chain of thought, calls it fragile — burny_tech · 2026-09-19
- Jev's popularity signals the AI crowd is open to models beyond LLMs — BLUECOW009 · 2026-09-19
- QuixiAI picks gemma-4-26B-A4B-it as base model for OpenJev — QuixiAI · 2026-09-19
- Open-weight Jev replica based on Qwen3.8 27B with 265k context drops tomorrow — TheZachMueller · 2026-09-19
- Jev beats a Sonnet 5-powered retriever on accuracy at a fraction of the cost — IanArawjo · 2026-09-19
- Alibaba open-sources medical AI model that detects cancer and nearly 150 conditions — giveen · 2026-09-19