Experiment: continued pretraining vs RAG on Qwen 3.5 4B for accuracy and performance

funJS · reddit · 2026-09-12

The author ran a hands-on comparison on Qwen 3.5 4B: continued pretraining (CPT) to internalize domain knowledge versus RAG on the base model, measuring accuracy and performance to quantify internalizing knowledge vs retrieving it on-the-fly. Full findings are in the linked write-up.

Original post →

More from Research

Research channel →