Fetching from the wire…
Public story · 2026-09-12 · high
Shenzhen Suqiao Intelligent Technology lists the card at 65% of a stock 5090's price, but the memory type on the listing doesn't match Nvidia's spec.
Why now: Tom's Hardware reported the listing on September 11, and it's already split opinion on r/LocalLLaMA.
A shop called Shenzhen Suqiao Intelligent Technology is selling a modified RTX 5090 with 96GB of memory on Alibaba for $3,888, about 35% cheaper than a stock US 5090 despite tripling the VRAM. Tom's Hardware reported the listing on September 11, and TechPowerUp corroborated it the same day, per Tom's Hardware.
For anyone running local LLMs, VRAM sets how big a model fits on one card. Tripling it at two-thirds the price of a stock 5090 would matter a lot.
The listing identifies the memory as GDDR6X. The retail RTX 5090 uses GDDR7. r/LocalLLaMA fixated on the mismatch, and no one outside the shop has verified what's actually on the board.
That leaves the memory bandwidth unknown. A 96GB card running at GDDR6X specs, if that's really what's in it, would hold bigger models than a stock 5090 but move data slower per pin while doing it.
Until someone puts one of these cards on a bench and measures memory bandwidth, the 96GB number is a capacity claim. Nothing about speed is confirmed.
Each link below shares sources, entities, or timing with this story.
Alibaba released Qwen3.6-27B on April 22. Dense architecture. Open weights. 77.2% on SWE-bench Verified, within 3.7 points of Claude Opus 4.6. On SkillsBench, it scores 48.2% versus its own 397B MoE predecessor's 30.0%. That's a 77% improvement with 14.8x fewer parameters. Let...
Ollama cut v0.34.0-rc1 on September 5 at 23:49 UTC, and the headline item changes the shape of the local-versus-hosted decision rather than the performance of either side: Ollama-hosted open models can be selected directly inside ChatGPT Desktop, with setup driven from the Oll...
Moonshot AI dropped Kimi K2.6 today and the numbers are hard to ignore. One trillion parameters total, 32 billion active per token across 384 experts, 256K context window, and native multimodal input. It scores 58.6 on SWE-Bench Pro versus GPT-5.4's 57.7 and Claude Opus 4.6's...
Tom's Hardware tracked the storefront price from $8,565 at March 2025 launch to $13,250 in June to $16,000 now. The cause named across Tom's, VideoCardz, TechPowerUp, and ThinkComputers is GDDR7 memory, not silicon: the card carries 32 × 3GB modules, so every per-module increa...
Alibaba post-trained its flagship in place on September 1, keeping the 2.4T-parameter base and 1M context. All eight published coding benchmarks improved: TerminalBench 3.0 from 11.3 to 29.0, DeepSWE 1.1 from 56.6 to 69.3, QwenSWEbench V2 from 55.1 to 70.0, JobBench from 53.4...
Thirteen hours of model time. Ninety-eight prompts. Two weeks. Five consumer devices reverse-engineered, and the results are not subtle. The author of schlarp.com had Claude Opus 5 work through firmware for an Insta360 Link webcam, an ASUS ROG Swift PG42UQ monitor, a Shure MV7...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.