Sources
GPT-6 Astra's public numbers: $10/$50 per million tokens, 1,050,000-token context, and only ~55 tok/s peak throughput
OpenRouter's model page for GPT-6 Astra lists input at $10.00/M, output at $50.00/M, cache read $1.00/M, cache write $12.50/M and web search at $10.00 per 1K calls, with a 1,050,000-token context window and up to 128,000 completion tokens. It supports tools, tool_choice and JSON-schema structured outputs, and accepts PDFs, images and text. The number to plan around is throughput: P50 best latency 2.89s and peak 55 tokens per second across providers, which makes Astra slow as well as expensive relative to the 5.6 line.
Source
↳ Follow the thread