Reddit
Opus 4.7 Review: 'Both Great and So Frustrating' — Better Coding, Stricter Prompts, Higher Cost
A practitioner review on r/ClaudeAI (50 upvotes, 19 comments) describes Opus 4.7 as simultaneously impressive and maddening. SWE-Bench scores are up ~8 points over 4.6, with 13% higher resolution on a 93-task coding benchmark. But the new tokenizer bumps token counts 12-18%, web research quality regressed, and Opus 4.7's stricter instruction following breaks prompts that relied on 4.6's gap-filling behavior. For builders: upgrade for coding-heavy workloads, stay on 4.6 for research-heavy ones.
Source
↳ Follow the thread