DeepSeek engineer who wrote v4.1's core Attention operator reflects on building his own replacement
WebAssemblyMan · reddit · 2026-09-15
A DeepSeek engineer who says he wrote v4.1's main Attention operator published a long reflection on recursive self-improvement:
- Pace: ChatGPT to reasoning models (o1, R1) took 2 years; reasoning to tool-using agents 1.5 years. In his own specialty, AI went from doc-lookup helper to independently reading CUDA/PTX/SASS and optimizing operators in one year.
- His read: AI-written operators will likely match or beat his within 6-12 months; AI thinks 300 tokens/second and writes code in 20 seconds—he can't.
- Why keep going: it's fun, and "if I must be revolutionized, I want it to be by myself"—deliberately slowing down wouldn't stop competitors anyway.
- Outlook: he expects to stay employed but likely forced into a new career, giving up work he loves.
Related event: DeepSeek Engineer Writes He Is Accelerating the AI That Replaces Him(4 posts)→
More from AGI Musings
- Rival AI Labs Rally Behind Anthropic's Slowdown Proposal — The AI Daily Brief · 2026-09-15
- AI arms-race narrative misleads policy, says researcher: cooperation, not competition, is the only path — GregCook2011 · 2026-09-15
- One hour of vibe coding flips 'AI is useless' skeptics, from professors to execs — pwlot · 2026-09-15
- Derek Thompson: No innocents in the frontier AI debate, it's partisans plus self-interest — deanwball · 2026-09-15
- AI doomerism goes mainstream: high schoolers refuse homework because 'we'll all die soon anyway' — BraydonDymm · 2026-09-15
- KP: Personal AI agents will build 1000x more apps than humans; AX beats UI — thisiskp_ · 2026-09-15