DeepSeek V4 Could Continue Pretraining with MOPD Reusing Domain Experts

teortaxesTex · x · 2026-07-31

A discussion suggests DeepSeek V4 could continue pretraining from intermediate checkpoints, not just post-training; MOPD allows cheap reuse of domain experts. Also, a user noticed V4 Flash API fingerprint update, with knowledge cutoff updated to Nov 2025.

Original post →

More from Models

Models channel →