Fetching from the wire…
Public story · 2026-09-12 · high
The Hugging Face upload comes packaged in four export formats, though nobody's verified its accuracy claims yet.
Why now: The model went up September 9, and by September 12 it had passed 3,000 downloads without an independent benchmark.
A finetune called Orukeet went up on Hugging Face September 9, built on Nvidia's Parakeet-TDT model, and it already has 3,066 downloads. That's a quick pickup for anyone building offline speech-to-text tools in languages other than English.
It's licensed CC-BY-SA-4.0 and tagged for 25 languages, from Bulgarian to Ukrainian.
The upload packages the same weights four separate ways: nemo, GGUF, ONNX and sherpa-onnx. That should spare anyone downstream from converting the model themselves.
The quality claim is thinner. Orukeet surfaced because someone recommended it inside OpenWhispr, and that poster said outright they hadn't tested it themselves. The idea that it beats stock Parakeet is secondhand, and the poster flagged that the doubt runs deepest on Macs.
If Orukeet doesn't outperform Nvidia's own model on a real transcription benchmark, this is a distribution story, not an accuracy one. Four export formats for one untested finetune already exist. That says the local speech-to-text crowd cares more about where a model runs than whether it's better yet.
Each link below shares sources, entities, or timing with this story.
For about a year, "run your agent locally" meant accepting a model that couldn't reliably call a tool twice in a row. That excuse is gone. Meta Superintelligence Labs published Muse Glimmer today: a 29.6B dense causal transformer, 52 layers, 6,656 hidden dim, with a ~1.8B ViT-...
DeepSeek released V4 Preview on April 24 with two open-weight variants: V4-Pro (1.6T total parameters, 49B activated via MoE) and V4-Flash (284B parameters, 13B activated). Both support 1M-token context windows. Both are Apache 2.0 licensed. Both are live right now on Hugging...
Ollama cut v0.34.0-rc1 on September 5 at 23:49 UTC, and the headline item changes the shape of the local-versus-hosted decision rather than the performance of either side: Ollama-hosted open models can be selected directly inside ChatGPT Desktop, with setup driven from the Oll...
TechCrunch strings together three deals: Nvidia's reported $13 billion Hugging Face acquisition, its $6 billion Poolside arrangement, and Stripe's acquisition of OpenRouter for over $7 billion about two weeks before August 28. The thesis is acquirers hedging against frontier-l...
The system (arXiv 2609.10712) uses no formal prover, no tools and no internet access. Three Nemotron 3 Ultra checkpoints run a generate-verify-refine loop, and together they scored 30 of 42 at IMO 2026, the gold threshold. NVIDIA posted the math SFT and RL checkpoints on Huggi...
The day's highest-scoring r/LocalLLaMA post points out the deal takes the llama.cpp and ggml copyright along with the team Hugging Face hired in February 2026, including Georgi Gerganov (r/LocalLLaMA). The top reply at 957 upvotes is "If it happens, we shall fork and move on....
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.