Swapping in Exa search makes DeepSeek V4.1 Flash consistently flagship-grade, tester finds
bookwormengr · x · 2026-09-15
A hands-on review of DeepSeek V4.1 Flash with a custom harness: (1) replacing the default search with Exa pushed answer quality from 10% dud rate to consistently Astra High/Fable level — search API quality matters far more than expected; (2) the model held its ground on a Huawei 7.2T NPO Engine spec debate and assembled a precise annotated diagram by cutting and pasting from multiple sources; (3) at 300+ tokens/s with fast Exa search, PTC and Agent Team modes deliver fast deep research that beat the author's patience-limited ChatGPT/Claude Deep Research workflows (Agent Team is a bit flaky). Verdict: maybe flash is all you need.
More from Models
- Claude Fable 5.1 holds #1: 1M-token input, ties GPT 6 Astra on new benchmarks — DeepLearningAI · 2026-09-15
- Tests show OpenAI hasn't changed Codex quotas: 800M+ Astra tokens per cycle — daniel_mac8 · 2026-09-15
- AlphaSense tests Fable and Astra as financial researchers: top score but only on par with Opus-5 — CShorten30 · 2026-09-15
- Grok 4.7 may drop today from xAI, unless delays strike again — mark_k · 2026-09-15
- Rumor: OpenAI may soon release GPT-6 Sol and Luna versions — soumitrashukla9 · 2026-09-15
- Princeton eval finds reasoning models fail structurally equivalent task variants, lacking systematicity — princetonu · 2026-09-15