Reddit
DeepSeek told investors behind locked doors it is training a 2T model and plans an 8T one
At a closed-door investor meeting on 2026-09-21 across its Beijing and Hangzhou offices, DeepSeek CEO Liang Wenfeng said the company is training a 2-trillion-parameter model and intends to build an 8-trillion-parameter one, which would be nearly three times Moonshot's 2.8T Kimi K3. Attendees had to surrender phones and electronic devices and were given only paper and pens. Analysts quoted in coverage peg the 2T model as DeepSeek V4.1 Pro, expected mid-to-late October; the r/LocalLLaMA thread (197 upvotes, 135 comments) traces the claim back to a single X post from @wallstengine.
↳ Follow the thread