RedditNVFP4 Support Coming to llama.cpp for Local InferenceGitHub·high signalTrue NVFP4 support arriving via GitHub PR 19769. Unlocks larger MoE models on 48GB or less memory at native FP4.SourceSource pageGitHub↳ Follow the threadShared entity / Stack layerBaseRT Claims 6.4x Faster Than llama.cpp and 3.9x Faster Than MLX on Apple SiliconProduct HuntPolicy dependency / Stack layernexu-io/open-design Hits 80.3K Stars as the Apache-2.0 Answer to Claude Design — Local Desktop App, 25 CLI Backends, Real File ExportGitHubPolicy dependency / Stack layerTwo Non-JS Ecosystems Get Native AI SDKs the Same Week — goai for Go and cruby's llm for RubyGitHubStack layer / Threat patternConstrain output style to cut tokens — and brevity may raise accuracy, not lower itGitHubStack layer / ContrastOpen Interpreter Reboots as a Coding Agent for Open Models Including Kimi K3GitHub TrendingStack layer / Follow-up threadHKUDS Follows nanobot With DeepTutor — 'Lifelong Personalized Tutoring' Built on RAGGitHub TrendingStack layerOmniRoute Trends at 20.9k Stars: One MIT-Licensed Endpoint Fronting 268+ Providers and 500+ ModelsGitHub Trending (primary repo)Policy dependency / Stack layerXiaomi Open-Sources Xiaomi-Robotics-1, a VLA Model Trained on 100K Hours of Real TrajectoriesHacker News / Xiaomi Robotics