An independent Aider eval finds ThinkingCap and Swift halve Qwen3.8-27B's time per case with no accuracy loss
r/LocalLLaMA (101 upvotes); huggingface.co/bottlecapai·medium signal
A community Aider eval (2 runs, Q8_0, llama.cpp 0.5.0) found that bottlecapai's new ThinkingCap-Qwen3.8-27B matched vanilla exactly: 27.1% first-try and 77.6% retry pass. It used 7,436 median tokens vs 12,547 and 777 s per case vs 1,481. Swift scored 30.8%/75.7% at 750 s per case. Two independent anti-overthinking fine-tunes now show about 2x wall-clock savings on local coding agents without measurable regression.