Fetching from the wire…
Skills2026-06-08 · source-backed
For non-reasoning models on math and logic, sampling multiple CoT paths and majority-voting lifts GSM8K by +17.9% over a single greedy pass. Highest-ROI accuracy lever for reasoning-heavy prompts. But skip it entirely on native reasoning models, they already reason internally and gain nothing, so you'd just pay 5x for the same answer.
Each link below shares sources, entities, or timing with this story.
OpenAI acquired Roi / Shared entity: CoT / Shared topic / Earlier coverage
Linked by a graph relationship (OpenAI acquired Roi); both cover CoT; overlapping topics (already, chain-of-thought, model, reasoning).
OpenAI acquired Roi / Shared entity: CoT / Shared topic / What happened next
Linked by a graph relationship (OpenAI acquired Roi); both cover CoT; overlapping topics (model, reasoning).
OpenAI acquired Roi / Shared entity: CoT / Shared topic / Earlier coverage
Linked by a graph relationship (OpenAI acquired Roi); both cover CoT; overlapping topics (chain-of-thought, reasoning).
headroom benchmarked against GSM8K / Shared entity: ROI / What happened next
Linked by a graph relationship (headroom benchmarked against GSM8K); both cover ROI; picks up the ROI thread on 2026-08-13.
headroom benchmarked against GSM8K / Shared entity: GSM8K / What happened next
Linked by a graph relationship (headroom benchmarked against GSM8K); both cover GSM8K; picks up the GSM8K thread on 2026-07-28.
OpenAI acquired Roi / Shared entity: ROI / What happened next
Linked by a graph relationship (OpenAI acquired Roi); both cover ROI; picks up the ROI thread on 2026-07-27.
Shared entity: CoT / Shared topic / Earlier coverage
Both cover CoT; overlapping topics (accuracy, chain-of-thought, model, reasoning); earlier CoT coverage from 2026-06-07.
Both cover CoT; overlapping topics (answer, chain-of-thought, model, reasoning); earlier CoT coverage from 2026-03-07.