The case against Ollama topped r/LocalLLaMA with 1,075 upvotes and 328 comments
r/LocalLLaMA (1,075 upvotes, 328 comments)·high signal
The post argues Ollama went over a year without crediting llama.cpp in its README while a license-compliance issue sat 400+ days without a maintainer response, that llama.cpp runs 1.8x faster (161 vs 89 tokens/second) with 30-50% CPU gaps, and that the mid-2025 move to a custom GGML backend reintroduced broken structured output, vision failures and assertion crashes. It also cites CVE-2025-51471 for token exfiltration via malicious registries and the DeepSeek-R1 distill naming. Top comments converge on LM Studio, llama.cpp and Unsloth Studio, with the recurring complaint being that Ollama will not use models already on disk.