Continued pretraining vs RAG: an accuracy and performance comparison on Qwen 3.5 4B

funJS · reddit · 2026-09-12

The author compares continued pretraining (CPT) of a Qwen 3.5 4B model against a RAG implementation on the same base model, measuring accuracy and performance to quantify the benefit of internalizing knowledge vs on-the-fly retrieval. Detailed results in the linked article — useful for anyone deciding between CPT and RAG for vertical domains.

Related event: CPT vs RAG: Testing Knowledge Internalization on Qwen 3.5 4B(3 posts)→

Original post →

More from Models

Models channel →