Fetching from the wire…
Public story · 2026-07-22 · high
AWS's fix, Self-Distilled Reasoning, adds the missing chain-of-thought tokens back in before fine-tuning starts.
Why now: AWS published this guidance as part of its Nova fine-tuning documentation, flagged in coverage dated July 22, 2026.
Fine-tuning strips reasoning out of models trained on ordinary answer pairs, AWS says. The stakes: teams pay for a reasoning model, fine-tune it on their own data, and quietly get less reasoning back.
AWS calls this reasoning suppression, laid out in a post about tuning its Nova models. Reasoning models work through visible thinking tokens before they answer.
Most fine-tuning datasets skip straight from question to answer, with no reasoning trace recorded in between. Train on those pairs, and a model learns to imitate the shortcut instead of the thinking.
AWS's fix is Self-Distilled Reasoning. It generates synthetic chain-of-thought tokens for datasets that don't have any, then fine-tunes on the augmented version instead of the raw pairs.
Built for Nova, but the problem it patches isn't specific to Nova, AWS says. Any reasoning model fine-tuned the ordinary way is exposed to it.
Nobody sets out to fine-tune the reasoning out of a model. It happens by accident, because the training data looks fine and nothing in a standard job flags whether the traces are missing.
Anyone fine-tuning a reasoning model on support tickets or resolved cases, with no thinking steps recorded, is running that experiment right now.
Each link below shares sources, entities, or timing with this story.
First-party pricing, counts toward AWS commitments, Codex via CLI and IDE plugins for VS Code, JetBrains, and Xcode, across commercial and GovCloud (AWS). This removes the procurement and compliance wall for AWS shops that couldn't touch OpenAI under existing contracts. Distri...
Six clients. One manifest. Zero vendor lock. Vercel published Agent Plugins 1.0.0 on August 6, an openly licensed spec that bundles Agent Skills and MCP servers behind a single portable manifest. The shape is deliberately boring: a plugin.json requiring only schemaVersion and...
The agent business model showed up this week, and it's rent. At Knowledge 2026, ServiceNow launched Action Fabric, a layer that every external AI agent must pass through to read data or run workflows inside the ServiceNow platform, metered on action-based pricing (ServiceNow N...
Amazon signed a definitive agreement with the Amsterdam team behind DuckDB. DuckDB and the rest of the Duck Stack stay MIT-licensed under the nonprofit DuckDB Foundation, and creators Hannes Mühleisen and Mark Raasveldt keep leading technical direction from Amsterdam (AWS). AW...
The full family, Sol, Terra, and Luna, is generally available on Amazon Bedrock with IAM and VPC controls (LLM Boss, AWS). Sol targets coding, biology, and cybersecurity agentic work. Terra runs everyday tasks at about half GPT-5.5's cost, and Luna optimizes for speed. The thr...
AWS made the managed AgentCore harness generally available on June 18. You define model, tools, skills, and memory with CreateHarness, then run it with InvokeHarness. It ships multi-model support (Bedrock, OpenAI, Gemini, LiteLLM), mid-session context preservation, built-in br...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.