Baseten’s paper writes 247 fake facts into Qwen3 and still can’t make them stick

gerardsans · x · 2026-07-26

What the paper tests

Main findings

Conclusion

The author’s interpretation is blunt: fine-tuning is not a reliable substrate for durable knowledge storage. Weight updates create local patches on top of a pretraining distribution, but the pretraining prior keeps reasserting itself.

Original post →

More from Models

Models channel →