Reddit
r/LocalLLaMA Runs a Three-Way Flash-Tier Bakeoff: AntLing 3.0 Flash vs MiniMax M2.7 vs Step 3.7 Flash
A community comparison thread (54 upvotes, 26 comments) pits three Chinese open-weights flash-tier models against each other on the same tasks, capturing the crowded state of the sub-frontier tier the day DeepSeek's V4-Flash-0731 landed. For reference on the two verifiable entrants: Step 3.7 Flash pairs a 196B language backbone with a 1.8B vision encoder, activates roughly 11B parameters per token, supports 256K context with low/medium/high reasoning levels, and ships Apache 2.0 on Hugging Face; MiniMax M2.7 is MIT-licensed at $0.28/$1.20 per 1M input/output tokens. Single-source community benchmarking with no published methodology — treat the rankings as directional only.
Source
↳ Follow the thread