Fetching from the wire…
Research2026-09-09 · source-backed
Executing 327 quantized code-capable artifacts (305 from the official Ollama library across 15 model lines, 22 from top HuggingFace community repos) through a calibrated 15-task smoke suite found five silently defective: four Qwen2.5-Coder-3B conversions and one phi3.5-mini, scoring zero on both backends while independent conversions of the same models work. That's 1.6% of official artifacts. Two produce output whose surface statistics sit inside the healthy range, so nothing short of execution catches them. arXiv 2609.05881 The released quantcheck tool is the acceptance gate model registries currently lack.
Each link below shares sources, entities, or timing with this story.
The July 6 release delivers nearly 90% faster Gemma 4 token generation through multi-token prediction with automatic draft-length tuning, on by default, output-preserving, no config (Ollama). It also adds MLX-engine support for more model families and flash attention for older...
Three moves, two days, no coordination between them. August 10–11: GitHub shipped Ollama as a BYOK provider inside Copilot for JetBrains (GitHub Changelog). Unsloth released Unsloth Desktop with a command literally named unsloth start claude, which points Claude Code and Codex...
v0.33.0-rc2, published August 21, adds a Claude integration letting you toggle individual Ollama models for use inside Claude from the menu bar. The caching note is the better find: Ollama was moving Claude Code's "tokens left" countdown system message to the front of the prom...
Piotr Wilam crossed Python and Rust with Qwen2.5-Coder-7B and DeepSeek-Coder-V1-6.7B, inventorying grammatical concepts (58 Python, 57 Rust) identically in all four cells. Which concepts earn dedicated circuitry is set by the task — the models agree at Spearman rho = 0.638 for...
Alibaba released Qwen3.6-27B on April 22. Dense architecture. Open weights. 77.2% on SWE-bench Verified, within 3.7 points of Claude Opus 4.6. On SkillsBench, it scores 48.2% versus its own 397B MoE predecessor's 30.0%. That's a 77% improvement with 14.8x fewer parameters. Let...
This Rust harness (+2,585 stars) competes on resource footprint rather than features: 27.8 MB PSS for a single session with local embedding disabled, claimed 13.9× less than Claude Code and 6× less than jcode's own embedding-enabled mode. Time-to-first-frame 14.0ms against a c...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.