Storing model-native numerical memory outside a frozen LLM: replicated from Qwen to Mistral, 127/128 Top-1

Nearby_Indication474 · reddit · 2026-10-07

After completing AKBASCORE MAM on Qwen2.5-7B-Instruct, the author independently localized and replicated the same memory mechanism on Mistral-7B-Instruct-v0.3: 127/128 correct memories at Top-1 (99.22%), counterfactual retrieval 125/128 (97.66%), shifted-pointer control 0/128. No weight changes, no fine-tuning, no LoRA, no learned router, and no gold memory ID is supplied at retrieval time.

Key ideas:

The author argues the deeper question than 99.22% is what exactly the model is reading: this path stores machine-native numerical states derived from the model itself, diverging from conventional text/embedding retrieval.

Original post →

More from Models

Models channel →