Swapping in Exa search makes DeepSeek V4.1 Flash consistently flagship-grade, tester finds

bookwormengr · x · 2026-09-15

A hands-on review of DeepSeek V4.1 Flash with a custom harness: (1) replacing the default search with Exa pushed answer quality from 10% dud rate to consistently Astra High/Fable level — search API quality matters far more than expected; (2) the model held its ground on a Huawei 7.2T NPO Engine spec debate and assembled a precise annotated diagram by cutting and pasting from multiple sources; (3) at 300+ tokens/s with fast Exa search, PTC and Agent Team modes deliver fast deep research that beat the author's patience-limited ChatGPT/Claude Deep Research workflows (Agent Team is a bit flaky). Verdict: maybe flash is all you need.

Original post →

More from Models

Models channel →