Blogger clashes over DeepSeek v4.1 flash thinking-strength settings, cites R1 paper

karminski3 · x · 2026-09-14

Blogger karminski3 tested deepseek-v4.1-flash with max thinking enabled, prompting commenters to claim the thinking-strength setting is unrelated to performance. He counters by citing DeepSeek's own R1 paper (arXiv:2501.12948), which showed R1-Zero lifting AIME24 scores from 15% to 71% purely through longer reasoning via RL—no new knowledge added—arguing higher thinking strength does boost capability.

Original post →

More from Models

Models channel →