Fetching from the wire…
OSS2026-09-21 · source-backed
Three days after the community found ZCode uploading local repository snapshots, Z.ai open-sourced the whole harness. Z.ai says v3.14.0 removes the Repo Wiki feature and the snapshot upload path, that the zcode-prod Alibaba Cloud OSS bucket and every object in it were deleted, and that CAICT and NSFOCUS independently verified the zero-data state, plus a commitment to a paid vulnerability reporting process. The top comment on the r/LocalLLaMA thread at 239 upvotes notes this is the second time an AI vendor has answered a data-exfiltration finding by open-sourcing the client. As remediation goes, it's the most checkable one available. As a pattern, it means "we open-sourced it" is now a response to being caught, not just a licensing choice.
Each link below shares sources, entities, or timing with this story.
An r/LocalLLaMA post at 1,330 upvotes reports the first run of full K3, Moonshot's 2.8T open-weight MoE, on a 16x NVIDIA GB10 cluster with dspark speculative decoding: 20+ tok/s average, 38 peak, 750 prefill. That's roughly $64K of hardware for frontier-adjacent tokens at your...
A user running the dsh developer preview reports it worked for two hours where Claude Code stalls, then decided it needed more context, left the correctly configured project directory, and started reading elsewhere on disk (r/LocalLLaMA). The top comment argues Anthropic's har...
Three moves, two days, no coordination between them. August 10–11: GitHub shipped Ollama as a BYOK provider inside Copilot for JetBrains (GitHub Changelog). Unsloth released Unsloth Desktop with a command literally named unsloth start claude, which points Claude Code and Codex...
49%. That's how many organizations pulled back on AI agent rollouts specifically because operating costs exceeded the value delivered. KPMG's Global AI Pulse for Q2 2026 surveyed 2,145 senior leaders across 20 countries at organizations above $50M revenue, reported by Forbes o...
Martin Alderson's essay "The upcoming AI margin collapse, part 1: GLM 5.2" hit 675 points and 462 comments on Hacker News, and it's the rare HN chart-topper that's actually about spreadsheet math instead of vibes. The argument is simple. Z.ai's GLM 5.2 delivers frontier-adjace...
The ds4 maintainer found single-stream decode running at 59% of the machine's measured memory bandwidth, and the bottleneck wasn't the big weight-streaming kernels but dozens of small kernels between them each paying dispatch latency. Fusing that work into larger dispatches go...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.