Fetching from the wire…
OSS2026-09-11 · source-backed
DeepSelect is the TopK kernel behind the indexer in DeepSeek Sparse Attention, and DeepSeek says it runs 2-20x faster than torch.topk. DeepJIT is a header-only C++20 runtime that compiles kernels at runtime, with one interface for both NVIDIA CUDA and Huawei Ascend. deepseek-recipe converts Messages, Chat Completions and Responses requests into V4 and V4.1 prompts for self-hosting. The Ascend support tells you the most: DeepSeek is building its kernels to run on Huawei hardware as well as NVIDIA.
Each link below shares sources, entities, or timing with this story.
DeepSeek released V4 on April 24 and the numbers demand attention. V4-Pro is 1.6 trillion parameters total with 49 billion active, MIT-licensed, native 1M-token context. It scores 80.6% on SWE-bench Verified, putting it within 0.2 points of Claude Opus 4.6. On Terminal-Bench 2...
Published September 7, it puts OpenAI Codex in the agent picker with a copy-ready ~/.codex/config.toml panel pointing Codex CLI and Desktop at Manifest over the Responses API (GitHub). Two compatibility fixes make it work: Responses-API role: "developer" instruction messages f...
Created August 26, a TypeScript harness for authorized penetration tests, bug bounty work, labs and CTFs, keeping sessions local. Every layer is replaceable from configuration, and it supports OpenAI Chat Completions and Responses, Anthropic Messages, DeepSeek and any OpenAI-c...
DeepSeek released V4 Preview on April 24 with two open-weight variants: V4-Pro (1.6T total parameters, 49B activated via MoE) and V4-Flash (284B parameters, 13B activated). Both support 1M-token context windows. Both are Apache 2.0 licensed. Both are live right now on Hugging...
IndexCache identifies that DeepSeek Sparse Attention reduces core attention to O(Lk) but its lightning indexer retains O(L²) complexity. By exploiting high cross-layer index stability to reuse token selection indices, IndexCache delivers significant wall-clock throughput impro...
Bloomberg reports DeepSeek will permanently maintain the V4-Pro discount that was set to expire end of May. Input pricing from $1.74 to $0.435/M tokens, output from $3.48 to $0.87. Enabled by migration to Huawei Ascend 950 accelerators and the 1.6T-parameter MoE architecture a...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.