Imagine a massive library where every book is shelved by its first few words. You want to correct a wrong date in one entry, but the same shelf holds dozens of other books that start the same way. Change the shelf label and you corrupt everything else on it. Leave it alone and the wrong fact persists. EngramEdit is a method for changing individual shelf labels surgically — correcting the targeted book across all its paraphrases while leaving neighbors untouched. The paper's committed claim: you can perform decoupled factual edits on conditional-memory architectures (specifically DeepSeek's Engram system) by computing target memory representations across multiple phrasings and then jointly updating the shared n-gram embeddings, with a penalty that scales with embedding reuse to protect unrelated knowledge. The Transformer backbone stays frozen throughout. This matters because conditional memory architectures like DeepSeek Engram already expand LLM capacity by looking up learned embeddings from input n-grams, adding capacity without proportional compute. But nobody had shown you could cleanly edit those embeddings for knowledge updates without cascading side effects. The hard problem is that different surface expressions of the same fact activate different n-gram routes, while shared embeddings create crosstalk between unrelated facts. EngramEdit's architecture is straightforward: first, compute target memory representations that make the model predict the corrected fact across a set of diverse paraphrases. Then, jointly optimize the shared n-gram embeddings to match those targets across all expressions and edits simultaneously, with a regularization penalty weighted by how frequently each embedding is reused elsewhere. This is a constrained optimization over embedding space, not fine-tuning the Transformer weights. The method belongs to the family of knowledge-editing techniques (ROME, MEMIT, MEND), but it operates on a fundamentally different substrate — external conditional memory rather than internal MLP weights. The results are striking on headline numbers: near-perfect editing success, and nearly 3× the strongest baseline's accuracy under chain-of-thought prompting for multi-hop reasoning over edited facts. Unrelated knowledge and general capabilities are reported as 'largely preserved' even as edits accumulate. However, all validation appears to be the authors' own experiments on their own setup. No pre-registration, no independent replication, and the comparison baselines are not clearly named with version numbers in the abstract. The 'near-perfect' and '3×' numbers are compelling but need independent confirmation. The milestone question is clear: how many simultaneous edits can accumulate before crosstalk starts degrading general performance? The paper says edits accumulate without major degradation, but the scale tested is not specified in the abstract. The real unlock is thousands-to-tens-of-thousands of concurrent edits — the point where this becomes a practical alternative to periodic retraining. That number, and the degradation curve around it, is what the field needs next. The obvious experiment not run: testing on a different conditional-memory architecture (not DeepSeek Engram), or at substantially larger scale with adversarial edit sets designed to maximize crosstalk. The honest read is likely (a) — they had access to one architecture and one compute budget, and generalization studies are expensive. If this only works on Engram's specific n-gram design, the contribution narrows considerably.