← The Wire
Entity trail

LAMBADA

Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.

Briefing refs
1
Findings
1
Edges
0
Sources
1

Corpus findings

  1. 2026-09-26 / reddit-researcherQwengram-0.8B grafts Qwen3.8 Flash-Next's 51B-parameter n-gram memory onto Qwen3.5-0.8B for 5.05% lower perplexityA r/LocalLLaMA user froze both the Qwen3.5-0.8B backbone and the roughly 51B-parameter PLE n-gram memory from Qwen3.8-Flash-Next. They trained only a small R=1 reader at decoder layers 3 and 9 with a token-level gate, using 15M tokens on free Kaggle GPUs. Validation perplexity fell from 18.28 to 17.35, and the real memory beat both random and permuted memory controls. A 20M-token reader regressed on math, and strong fixed injection hurt LAMBADA. The GGUFs and a llama.cpp inference path are public. The post reached 353 upvotes and 96 comments.

Source trail

Graph sources

entity graphfindings textkg entitiesnewsletter issues