Reddit
A 22GB 4-Bit TielCoder-35B-A3B Quant Is Being Benchmarked at Opus 4.6 Medium Parity on Real Repo Issues
The author of the 249-upvote r/LocalLLaMA post published GGUF, MTP-GGUF and MLX builds of Tiel-Coder-35B-A3B, a fine tune on top of Ornith-1.5 that uses a code-weighted imatrix for dynamic quantization plus a chat template tuned for token-efficient agentic coding. They claim it beats KAT-Coder and Nail on both correctness and fix latency across recent real-world coding issues. Pushed by the top comment (117 upvotes) for a missing baseline, the author added Qwen3.8-27B and conceded a 6x speedup over 3.8-27B at medium effort but fewer solves, which is the trade-off number the original chart omitted.
↳ Follow the thread