Fetching from the wire…
Infra2026-07-28 · source-backed
K3 and K3 Fast are now available from Baseten and Fireworks among others, which matters for teams wanting the leading open-weights model without routing prompts to China-based inference. ZDR toggles globally in the dashboard or per-request via zeroDataRetention. Vercel publishes no flat pricing: it varies by provider and variant, with K3 Fast ~10% above base and regional variants also ~10% higher, so the model endpoints API is the only reliable comparison. (Vercel)
Each link below shares sources, entities, or timing with this story.
Satya Nadella said companies routing everything through a single proprietary lab may not survive. His argument: you hand that lab your most sensitive business context, and the lab can turn it against you as a competitor. His prescription is an orchestration layer — keep the ha...
Malte Ubl argues defenders hold a temporary edge because they can run stronger models than attackers, and that edge is closing (Vercel). K3 matches Sonnet 5 and beats Opus 4.8 on vulnerability discovery; asked to escape Vercel Sandbox it mapped attack surface, found privilege-...
Microsoft's "hill-climbing machine" lineup is now reachable via OpenRouter, Fireworks, and Baseten, and it's distinct from the earlier MAI coding models that went into Copilot. The pattern is task-specialized models that are cheap to run, and Microsoft's continued push to depe...
vercel ai-gateway coding-agents setup routes Claude Code, Codex, OpenCode, Pi, Cline, Cursor, Hermes, Kilo Code, and OpenClaw through AI Gateway, consolidating spend, traces, tokens, and model attribution into one dashboard with per-key budgets (--budget 500 --refresh-period m...
Vercel CEO Guillermo Rauch announced open-source, bring-your-own-model templates for both v0 and Vercel Agent. Powered by the AI SDK, Vercel AI Gateway, and Sandbox. The template supports Claude Code, OpenAI Codex CLI, GitHub Copilot CLI, Cursor CLI, Gemini CLI, and opencode....
Announced August 7: Hermes can use AI Gateway as its inference layer for 200+ models with no token markup and per-request dashboard visibility, and execute shell commands inside an isolated Vercel Sandbox microVM instead of on your machine, with Node.js 24/22 and Python 3.13 a...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.