Hands-on: Qwen 3.8 Flash Next beats new Siri at PDF payment sums
sleight42 · reddit · 2026-10-09
The author ran a small hands-on test: given a PDF of medical provider statements and asked to sum payments, Siri first grabbed the value on page one, then after being corrected still summed it wrong across a few pages. The same document handed to Qwen 3.8 Flash Next via Hermes extracted the text, summed it, and then used vision to double-check itself. Conclusions: Siri is still weak, Qwen 3.8 xxs is quite good at administrative agent tasks, and that removes one more reason to buy a new iPhone.
More from Models
- Text-Only Qwen3.5 2B/4B/9B MLX 4-bit Packages Released, 2B Is Just 1GB — sachasayan · 2026-10-09
- Why multilingual LLMs are hard: character decoding and BPE are the hidden bottleneck — ivan_bezdomny · 2026-10-09
- Anthropic launches Cyber Mission to defend critical infrastructure and open-source software — AnthropicAI · 2026-10-09
- Best model was cheapest: open-weights model ran 669 clinical decisions for 1.7 cents — antoine_chaffin · 2026-10-09
- Emad Mostaque: OpenAI Burned $10-20M Compute Solving Navier-Stokes, Prices Falling Fast — rohanpaul_ai · 2026-10-09
- Musk touts Grok Bot upgrades: Opus 5.5 on demand, full X access, big speed gains — elonmusk · 2026-10-09