Code Model Security Flaws and MoE Fine-Tuning Fragility

Research indicates that code models are inherently unsafe due to public training data, and simple safety fine-tuning is insufficient. Furthermore, MoE models are more fragile than Dense models during SFT, showing high sensitivity to hyperparameters and routing instability.

2026-08-12 ~ 2026-08-12 · 2 related posts