The r/singularity read on Astra is token efficiency, not capability, because thinking tokens are no longer emitted
r/singularity·medium signal
A 536-upvote thread argues the underreported Astra result is cost per task rather than benchmark position, comparing token consumption against competitors on the same work. The corrective in the comments is that Astra reasons internally and emits no visible thinking tokens, so part of the apparent efficiency is a change in what gets billed and displayed rather than in what gets computed. Anyone comparing spend across models this week should check whether the counter includes reasoning tokens before concluding anything.