Fetching from the wire…
Models2026-07-25 · source-backed
owensong released Inflect-Nano-v2 (3,966,721 deployable params) and Inflect-Micro-v2 (9,356,513) under Apache-2.0, VITS-family end-to-end text-to-waveform with 128 latent channels, 3 encoder layers, 4 flow coupling blocks, 24 kHz mono. Nano-v2 runs at 0.0933 RTF (10.72x real-time) on four CPU threads; Micro-v2 at 0.1593 RTF (6.28x). The trade-off is deliberate: English only, one fixed male voice, and the training-corpus pipeline stays private, so it's open-weight not open-source. 410 upvotes on r/LocalLLaMA. Usable on-device TTS now fits in under 10MB of parameters, which puts voice output inside embedded budgets that ruled it out a year ago.
Each link below shares sources, entities, or timing with this story.
Shared entities / Same source domain / Shared topic / Earlier coverage
Both cover Apache, TTS; reported by the same outlet (huggingface.co); overlapping topics (apache-2, deployable).
Shared entities / Earlier coverage
Both cover Apache, English, TTS; earlier Apache coverage from 2026-03-26.
Both cover Apache, LocalLLaMA, Nano; earlier Apache coverage from 2026-03-02.
Shared entities / Same source domain / Earlier coverage
Both cover Apache, LocalLLaMA; reported by the same outlet (huggingface.co); earlier Apache coverage from 2026-04-24.
Both cover Apache, LocalLLaMA; reported by the same outlet (huggingface.co); earlier Apache coverage from 2026-04-02.
Shared entities / Earlier coverage / Downstream implication
Both cover CPU, LocalLLaMA; earlier CPU coverage from 2026-05-10; traces where this leads (implications).
Both cover CPU, LocalLLaMA; earlier CPU coverage from 2026-04-10; traces where this leads (what it means).
Shared entities / Same source domain
Both cover Apache, LocalLLaMA; reported by the same outlet (huggingface.co).