Reddit
llama.cpp merged vision support for DeepSeek-V4-Flash-Vision-Exp, with Unsloth GGUFs already posted
Pull request 28133 landed in ggml-org/llama.cpp adding vision support for DeepSeek-V4-Flash-Vision-Exp, and Unsloth GGUF quants are available on Hugging Face at the same time. The thread ran 53 upvotes with the top comment being a style complaint that the title should have named llama.cpp explicitly. This closes the gap between the DeepSeek-V4-Flash and GLM-5.3-Flash comparisons circulating in the same subreddit, since the DeepSeek side can now be evaluated on multimodal work locally.
↳ Follow the thread