Fetching from the wire…
Public story · 2026-08-03 · high
The repo moved 8.25 terabytes of declared weights through Hugging Face's Xet system in about 692 kilobytes, a ratio near 11,900,000 to one.
Why now: The repository was still live and unflagged at 906 downloads as of August 3.
Vacuum-16t declares 16,501,264,351,232 parameters across 3,841 tensors on Hugging Face, per its model card.
Every one of those parameters is a zero, and the listing still passes Hugging Face's verification and shows 906 downloads. If a real lab padded a safetensors header the same way to inflate a leaderboard number, nothing in that verification step would catch it.
The entire 8.25-terabyte file traces back to one deduplicated 64-kilobyte block of nulls. The upload moved through Hugging Face's Xet system in about 692 kilobytes, a compression ratio near 11,900,000 to one. It carries an MIT license and a declared context window of 4,294,967,296 tokens.
That's possible because Hugging Face computes a model's parameter count from the safetensors file header, the metadata block listing tensor shapes and dtypes. It never reads the tensor data underneath.
A header can declare any shape it wants, and nothing checks that the bytes match. The mechanism, not the joke, is what Hugging Face would need to fix.
Each link below shares sources, entities, or timing with this story.
Announced July 27 with Microsoft, IBM, Red Hat, Palantir, CrowdStrike, Cloudflare, Databricks, Hugging Face, LangChain, Nous Research, Reflection AI, Thinking Machines Lab, SpaceXAI and the Linux Foundation. Huang's framing is pointed: during the Hugging Face incident "closed...
This is a supply-chain fact, and most people are still treating it as a geopolitics argument. Sequoia published "America's Open-Model Paradox" on July 24 with the number that reframes the whole conversation: Qwen's share of open-model fine-tunes went from 1% in January 2024 to...
Blaizzy/nativ (1,163 stars, Swift, MIT, macOS 26+) comes from the mlx-vlm author and bundles that server into a SwiftUI app that discovers MLX models already in your HF cache. It exposes OpenAI-compatible chat, Responses, image, audio and model endpoints plus Anthropic Message...
DeepSeek released V4 Preview on April 24 with two open-weight variants: V4-Pro (1.6T total parameters, 49B activated via MoE) and V4-Flash (284B parameters, 13B activated). Both support 1M-token context windows. Both are Apache 2.0 licensed. Both are live right now on Hugging...
1. Use claude agents --json to build session dashboards. Claude Code v2.1.145 outputs all live agent sessions as structured JSON with status, model, elapsed time, and parent relationships. Pipe it into a tmux status bar widget or session picker script for switching between bac...
Huang used his inaugural X post on July 24 to publish "Open Weights and American AI Leadership," a three-page letter on Nvidia's own servers signed by 25 companies including Meta, Microsoft, IBM, Mistral, Mozilla, Hugging Face, a16z, Palantir and the Linux Foundation. Within a...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.