Fetching from the wire…
Models2026-08-29 · source-backed
English and Chinese, with reference-free voice design from a natural-language description, reference-guided cloning and low-latency streaming, currently first among open-weight models on the Artificial Analysis TTS leaderboard. The 6GB/4x-realtime figure comes from testers using the audio.cpp dev branch. Caveats from the same thread: the supported tag set is limited and not always followed, quality varies noticeably by seed, and the multilingual variant on the playground hasn't been released. (r/LocalLLaMA)
Each link below shares sources, entities, or timing with this story.
Shared entities / Same source domain / Shared topic / Earlier coverage
Both cover LocalLLaMA, VRAM; reported by the same outlet (reddit.com); overlapping topics (analysi, artificial).
Shared entities / Same source domain / Earlier coverage / Tension
Both cover LocalLLaMA, VRAM; reported by the same outlet (reddit.com); earlier LocalLLaMA coverage from 2026-07-27.
Shared entities / Same source domain / Earlier coverage / Downstream implication
Both cover LocalLLaMA, VRAM; reported by the same outlet (reddit.com); earlier LocalLLaMA coverage from 2026-04-10.
Shared entity: LocalLLaMA / Same source domain / Shared topic / Earlier coverage / Tension
Both cover LocalLLaMA; reported by the same outlet (reddit.com); overlapping topics (analysi, artificial).
Shared entities / Earlier coverage
Both cover English, LocalLLaMA, TTS; earlier English coverage from 2026-07-25.
Shared entities / Same source domain / Earlier coverage
Both cover LocalLLaMA, VRAM; reported by the same outlet (reddit.com); earlier LocalLLaMA coverage from 2026-08-28.
Both cover Chinese, LocalLLaMA; reported by the same outlet (reddit.com); earlier Chinese coverage from 2026-08-27.
Both cover LocalLLaMA, VRAM; reported by the same outlet (reddit.com); earlier LocalLLaMA coverage from 2026-08-23.