Fetching from the wire…
Models2026-07-26 · source-backed
owensong's VITS-family English TTS has 9,356,513 deployable parameters and a 37.53MB FP32 footprint producing 24kHz mono (Hugging Face). Reported: 66.2% preference in community blind listening tests, 4.395 UTMOS22 naturalness, 3.99% semantic error rate across multiple ASR evaluators. It hit 180 points on HN with only 15 comments, which usually means people upvoted and went to try it. The training corpus pipeline and full optimization recipe are withheld, so it's a usable artifact rather than a reproducible one.
Each link below shares sources, entities, or timing with this story.
owensong released Inflect-Nano-v2 (3,966,721 deployable params) and Inflect-Micro-v2 (9,356,513) under Apache-2.0, VITS-family end-to-end text-to-waveform with 128 latent channels, 3 encoder layers, 4 flow coupling blocks, 24 kHz mono. Nano-v2 runs at 0.0933 RTF (10.72x real-t...
Published August 25, it's IBM's first family of dense decoder-only reasoning models, with the 30B flagship claiming state-of-the-art resolve rates on SWE-Bench Pro and Terminal-Bench. The 8B and 30B went through an agentic training curriculum on real sandboxes for software eng...
For about a year, "run your agent locally" meant accepting a model that couldn't reliably call a tool twice in a row. That excuse is gone. Meta Superintelligence Labs published Muse Glimmer today: a 29.6B dense causal transformer, 52 layers, 6,656 hidden dim, with a ~1.8B ViT-...
DeepSeek released V4 Preview on April 24 with two open-weight variants: V4-Pro (1.6T total parameters, 49B activated via MoE) and V4-Flash (284B parameters, 13B activated). Both support 1M-token context windows. Both are Apache 2.0 licensed. Both are live right now on Hugging...
Blaizzy/nativ (1,163 stars, Swift, MIT, macOS 26+) comes from the mlx-vlm author and bundles that server into a SwiftUI app that discovers MLX models already in your HF cache. It exposes OpenAI-compatible chat, Responses, image, audio and model endpoints plus Anthropic Message...
Released August 4 under Apache 2.0, reframing moderation as policy-adaptive question answering: write your rule in plain language, get a calibrated safety score from a single token, no retraining, one interface for text and images. Mistral claims it matches open guard models u...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.