Fetching from the wire…
Models2026-08-11 · source-backed
The 0731 checkpoint sits at 1,048,685 downloads and 3,112 likes, second on HuggingFace trending behind MiniMax-H3. For scale, Kimi-K3 has been live since June 13 and sits at 1,565,484, so Flash covered two-thirds of that in under two weeks. Small fast-inference tiers are where open-weight adoption is compounding, not frontier checkpoints. r/LocalLLaMA is arguing it's the killer app for DGX Spark: 13B activated params per token, REAP-pruned 3.0bpw EXL3 builds serving 262,144-token context on a single 128GB Spark at ~26 tok/s, 82 tok/s across two units, with at least four independent GitHub recipe repos.
Each link below shares sources, entities, or timing with this story.
For about a year, "run your agent locally" meant accepting a model that couldn't reliably call a tool twice in a row. That excuse is gone. Meta Superintelligence Labs published Muse Glimmer today: a 29.6B dense causal transformer, 52 layers, 6,656 hidden dim, with a ~1.8B ViT-...
DeepSeek released V4 Preview on April 24 with two open-weight variants: V4-Pro (1.6T total parameters, 49B activated via MoE) and V4-Flash (284B parameters, 13B activated). Both support 1M-token context windows. Both are Apache 2.0 licensed. Both are live right now on Hugging...
Hugging Face published its Summer 2026 State of Open Models report on August 14, and one statistic in it went almost entirely unremarked in the coverage. By July 2026, agents rather than humans became the Hub's primary users. Claude Code alone accounted for 44.4% of all agent...
DeepSeek-V4-Flash-0731 landed July 31 under MIT with a DSpark speculative-decoding module attached. Terminal Bench 2.1: 82.7. Toolathlon-Verified: 70.3. DSBench-FullStack: 68.7. DeepSWE: 54.4. NL2Repo: 54.2. The model card claims it beats DeepSeek-V4-Pro (Preview) "despite its...
DeepSeek released V4 on April 24 and the numbers demand attention. V4-Pro is 1.6 trillion parameters total with 49 billion active, MIT-licensed, native 1M-token context. It scores 80.6% on SWE-bench Verified, putting it within 0.2 points of Claude Opus 4.6. On Terminal-Bench 2...
The Claude Code source leak was the biggest story in developer tools this week. But the most important analysis didn't come from the people picking through feature flags and Easter eggs. It came from Sebastian Raschka, who read the 512,000 lines of leaked TypeScript and reache...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.