Vercel AI Gateway Adds a Unified 'Fast Mode' Abstraction Across Every Model
Vercel Changelog·medium signal
Vercel shipped a beta unified fast-mode abstraction on AI Gateway: set speed to fast and the gateway serves the fast tier wherever a model offers one, using one request shape for every provider. This removes a real annoyance — each lab exposes latency tiers differently, so today you either hardcode per-provider flags or forgo the tier. For anyone routing across Anthropic, OpenAI and xAI behind one gateway, it makes latency a config value instead of a branch.