Fetching from the wire…
Skills2026-06-07 · source-backed
On o3/o4-mini-class models with built-in reasoning, explicit CoT prompting gained only ~2.9 to 3.1% accuracy while adding 20 to 80% latency and cost. Net loss. Reserve CoT for non-reasoning models (it lifts Gemini Flash 2.0 ~13.5%), and for reasoning models supply only the objective and constraints. Source: Wharton GenAI Labs
Each link below shares sources, entities, or timing with this story.
Shared entity: CoT / Shared topic / What happened next / Tension
Both cover CoT; overlapping topics (accuracy, chain-of-thought, model, reasoning); picks up the CoT thread on 2026-06-08.
Google released Gemini Flash / Shared topic
Linked by a graph relationship (Google released Gemini Flash); overlapping topics (cost, flash, gemini, latency, model).
Google released Gemini Flash / Shared topic / Tension
Linked by a graph relationship (Google released Gemini Flash); overlapping topics (cost, flash, gemini, latency); pushes against this story (but).
Google released Gemini Flash / Shared topic
Linked by a graph relationship (Google released Gemini Flash); overlapping topics (built-in, flash, gemini, model).
Shared entity: CoT / Shared topic / What happened next
Both cover CoT; overlapping topics (chain-of-thought, explicit, model); picks up the CoT thread on 2026-08-06.
Shared entity: Gemini Flash / Shared topic / What happened next
Both cover Gemini Flash; overlapping topics (cost, gemini, model); picks up the Gemini Flash thread on 2026-07-31.
Shared entity: CoT / Shared topic / Earlier coverage / Downstream implication
Both cover CoT; overlapping topics (cost, reasoning); earlier CoT coverage from 2026-03-11.
Shared entity: Gemini Flash / Shared topic / Earlier coverage
Both cover Gemini Flash; overlapping topics (flash, gemini, model); earlier Gemini Flash coverage from 2026-05-03.