Fetching from the wire…
Public story · 2026-08-17 · high
Version 1.1.10 shipped 32 MLX benchmarks on an Apple M4 Pro, the project's first, and fixed an Ollama bug that mismarked model families as installed.
Why now: llmfit published its first MLX benchmarks and shipped v1.1.10 on August 17.
Local-inference tool llmfit shipped version 1.1.10 with the project's first MLX benchmark results, run on an Apple M4 Pro, per its GitHub release.
That matters for anyone picking a local-inference runtime on Apple Silicon. MLX and llama.cpp's Metal backend rarely get compared on identical hardware, and 32 runs now do exactly that.
Those 32 results merged into llmfit's benchmark suite as new entries, giving developers a same-chip reference point instead of scattered numbers from different machines. The GitHub release doesn't say which backend won more of the 32 runs, only that the entries now exist side by side.
Version 1.1.10 also expanded llmfit's MCP server with RamaLama runtime discovery, and added the Qwen3.8 model family alongside vision capability exposure for MiniMax M3.
A separate fix stops a single sized Ollama install from marking its whole model family as installed.
The benchmark numbers matter less than the precedent they set. Once one project publishes a same-hardware MLX-versus-Metal comparison, framework maintainers lose the excuse that unfavorable results came from mismatched test rigs. Watch whether other local-inference tools start citing llmfit's numbers instead of running their own one-off comparisons.
llmfit shipped this comparison on August 17.
Each link below shares sources, entities, or timing with this story.
Ollama uses MLX / Shared entities / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (Ollama uses MLX); both cover GitHub, MCP, Metal, Ollama; reported by the same outlet (github.com).
Linked by a graph relationship (Ollama uses MLX); both cover Apple Silicon, MLX, Ollama; reported by the same outlet (github.com).
Ollama uses MLX / Shared entities / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Ollama uses MLX); both cover GitHub, Ollama, Qwen3; overlapping topics (hardware, model).
Ollama uses MLX / Shared entities / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (Ollama uses MLX); both cover MiniMax M3, Ollama; reported by the same outlet (github.com).
Linked by a graph relationship (Ollama uses MLX); both cover GitHub, Ollama; reported by the same outlet (github.com).
Apple released MLX / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (Apple released MLX); both cover MLX, Ollama, Qwen3; overlapping topics (against, model).
Ollama uses MLX / Shared entities / Same source domain / Earlier coverage
Linked by a graph relationship (Ollama uses MLX); both cover Apple Silicon, MLX, Qwen3; reported by the same outlet (github.com).
Ollama uses MLX / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (Ollama uses MLX); both cover Apple Silicon, Ollama, Qwen3; overlapping topics (apple, model).