Tang Jie on LLMs: Post-Training is More Crucial
dotey · x · 2026-07-18
This repost summarizes Tang Jie's extensive thoughts on LLM development in 2025. The core message: pre-training remains important but is no longer the only protagonist; what truly determines a model's real-world applicability is mid-to-late stage training that activates capabilities for long-tail scenarios and practical tasks.\n\nHe also makes a strong assertion: the first principle of AI applications shouldn't be "building new apps," but rather "replacing human labor." Following this logic, when building applications, one should prioritize identifying which roles and workflows are best suited for AI takeover, rather than just chasing superficial product innovation.\n\nThe post also emphasizes that many current models have become "over-specialized" just to ace benchmarks, making them unstable in complex real-world scenarios. Consequently, post-training, alignment, and training oriented towards real-world tasks will become increasingly critical.
More from AGI Musings
- IG reel debunking the 'winning the AI race against China' fallacy hits 500k likes — louisvarge · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- AI researcher memes agent-swarm tinkering with He Jiankui's embryo-editing quote — dejavucoder · 2026-09-11
- nabla_theta: happy to be wrong if the AI utopia arrives with little ex ante risk — nabla_theta · 2026-09-11
- OpenAI researcher: space operas now need ambiguously aligned superintelligences for realism — jachiam0 · 2026-09-11
- Great Forecasters Aren't Magic: Bare Probabilities Need Model-Based Derivation — Sam_kuyp · 2026-09-11