AINewsnow

I stored model-native numerical memory outside a frozen LLM and retrieved the associated memory through its own attention Qwen → Mistral replication, 127/128 Top-1

Two different 7B transformers. Two different internal coordinates. The same memory mechanism. I started AKBASCORE MAM on Qwen2.5-7B-Instruct. I have now independently localized and replicated the mechanism on Mistral-7B-Instruct-v0.3. Final Mistral result: 127/128 correct memories at Top-1 — 99.22%…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-07 10:11 · r/LocalLLM
    I stored model-native numerical memory outside a frozen LLM and retrieved the associated memory through its own attention Qwen → Mistral replication, 127/128 Top-1

More stories

  1. Mistral Large 4 beats Qwen 3.8 Max and Kimi K3 on Terminal-Bench — r/singularity
  2. Introducing Mistral Large 4 — Mistral AI News
  3. Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of China — Wired AI
  4. What to know about Mistral's ML4 as it bets on EU sovereignty in the US-China open-weight AI race — Euronews Next
  5. Qwen Flash Next on Single B200 or B300, any pointers ? — r/LocalLLM
  6. I mapped every major Qwen release from 2023 to 2026: 44 models, from Qwen-7B to the 2.4T open weights (with sources) — r/machinelearningnews
  7. Europe finally takes the lead — r/LocalLLM
  8. Q (@qtnx_) on X - Mistral Large 4 is still doing RL runs, keep seeing improvements (vs preview version). Release at the end of the month — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →