Fetching from the wire…
Models2026-09-22 · source-backed
At a closed-door investor meeting on September 21 across its Beijing and Hangzhou offices, CEO Liang Wenfeng said the company is training a 2-trillion-parameter model and intends to build an 8-trillion one, which would be nearly three times Kimi K3's 2.8T. Attendees surrendered phones and were given paper and pens. Analysts quoted in coverage peg the 2T model as DeepSeek V4.1 Pro, expected mid-to-late October. The chain traces back to a single X post, so the sourcing is thin. (r/LocalLLaMA)
Each link below shares sources, entities, or timing with this story.
DeepSeek posted a community notice: once V4.1 Flash launches around September 10 Beijing time, and until a V4.1 Pro exists, every V4 Pro request routes to V4.1 Flash and bills at Flash unit pricing. The stated reason is that Flash has surpassed Pro on performance, cost, speed...
OpenAI posted first benchmark results for Jalapeño, its Broadcom co-developed inference ASIC, claiming 1.5x to 1.9x more throughput per kilowatt and 1.7x to 3.6x lower end-to-end latency than Nvidia GB200 and GB300 rack systems, measured on the SemiAnalysis InferenceX suite. T...
RTK has almost 80,000 GitHub stars and a simple promise. It sits between your coding agent and the shell, trims noisy command output before the model reads it, and claims 60-90% savings. Quesma ran it on Terminal-Bench 2.1 and found costs went up. With RTK on, average cost per...
The aggregate parameter count Chinese labs shipped in 30 days now exceeds what the whole open-weight ecosystem produced in the first half of 2026. r/LocalLLaMA Nikkei Asia reported August 15 that Z.ai positions GLM-5.3 as a direct rival to Anthropic's Mythos on coding and secu...
DeepSeek dropped V4 in mid-June as an open-weight model with a 1-million-token context window, priced at $1.74 per million input tokens, posting near-parity with GPT-5.4 on math and Q&A benchmarks (MindStudio). That's the headline number. The architecture underneath is more in...
A 232-upvote r/LocalLLaMA thread builds an open-weights argument on Ramp's mid-August corporate spending data: Fable 5, the most capable and expensive model in the lineup, accounts for 11% of those companies' Anthropic spend. (r/LocalLLaMA) The thread's read is that Qwen, GLM...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.