DeepSeek Found Using SmolVLM to Curate Pretraining Image-Text Data

Hugging Face researcher Elie Bakouch discovered that DeepSeek used SmolVLM to strictly score and curate image-text pretraining data, noting with pride that his team's model was adopted in industry. He later clarified the finding came from an earlier DeepSeek technical report.

2026-09-10 ~ 2026-09-10 · 3 related posts