DeepSeek V4.1 Flash's 190B Engram Lookup Table Confuses Builders: Can It Store Knowledge Without Fine-Tuning?
No_Afternoon_4260 · reddit · 2026-09-15
- A Reddit user asks for an ELI5 of DeepSeek V4.1 Flash's architecture: what do the 500B+ weights and the 190B+ "engram" lookup table actually bring?
- Known traits include aggressive prefill optimization and a very large KV cache.
- Key open question: can the engram table be updated with new knowledge—or grown—without fine-tuning the model?
More from Models
- A common jailbreak: third-person role-play gradually blurs lines to widen the model's Overton window — BlancheMinerva · 2026-09-16
- cocktail-peanut predicts Jev will go open weights, seeing more potential locally than as an API — cocktailpeanut · 2026-09-16
- Speculation mounts OpenAI's mysterious 'new model' is a fresh pretrain, not an RL run — teortaxesTex · 2026-09-16
- GPT-6 Astra hits 68.7% on DrugDiscoveryBench, benchmark authors call it a step function — KexinHuang5 · 2026-09-16
- Zero scores 2.5% on Grade School Math vs base model's 62.2% — creators say it's not an assistant — maxsloef · 2026-09-16
- ChatGPT generates suggestive image, then refuses to swap its wolves for cats — comFX87 · 2026-09-16