Fetching from the wire…
Public story · 2026-09-12 · high
The upgrade path is Gemini 3.1 Pro and 3.6 Flash at higher list prices, and one migration reported regressions.
Why now: Google's October 16 cutoff is five weeks away, leaving teams still on 2.5 with a short window to test the replacements.
Google will retire gemini-2.5-pro and gemini-2.5-flash on October 16, 2026, a date pushed back from an original June cutoff. The listed upgrades, Gemini 3.1 Pro and 3.6 Flash, cost more on the price list. Teams relying on 2.5 Pro for long-document work don't get a simple swap.
The specific loss, per a Hacker News thread on the change, is 2.5 Pro's document handling. One poster says it fits thousand-page inputs into about 300,000 tokens, and that no competitor matches that at comparable pricing. If true, matching it requires more than picking a newer model number.
A second commenter in the same thread reports regressions after moving workloads to the 3.x Flash line, without naming which workloads or what broke. One report isn't a benchmark. Anyone running similar work on 2.5 Flash should test 3.6 against real inputs before the cutoff.
The thread pulled only 19 points, small enough to read as an early warning rather than a wave of migration complaints. Two months is the entire notice period Google gave for a hosted model with no same-tier replacement listed.
Each link below shares sources, entities, or timing with this story.
Hardcoded model IDs now return API failures, and teams moving to 2.5-flash face a second mandatory cutover by October 16 (Google). Three major versions in under a year. The lesson isn't which model to pick, it's to stop hardcoding model IDs entirely. Abstract the selection or...
"Gemini is Cooked but GCP is Cooking" argues Google quietly shelved 3.5 Pro, which industry chatter placed at roughly Opus 4.5 level, shipping Gemini 3.6 Flash as a bridge the authors call worse than Muse Spark 1.2, Grok 4.5, and tier-1 Chinese open-source models. The hard num...
A June 18 model-tracking roundup reports Google set Gemini 2.5 Flash as the default across its consumer Gemini products, prioritizing latency and cost. This is single-source as of writing, so flag it pending the official Gemini blog, but it lines up with Google's other June 18...
The flagship slipped past July 17, the third postponement since an original June 2026 date. Bloomberg-sourced reporting attributes it to coding benchmarks failing to match GPT-5.6, plus hallucination and output-consistency problems, with a retraining data refresh aimed at codi...
Google pushed Flash to general availability and rolled out Gemini in Chrome (Windows/Mac for AI Pro/Ultra in the US), Gemini Omni globally to subscribers 18+, and a US Daily Brief (Google Gemini). The Flash GA is the builder-relevant piece: frontier-ish quality at speed and pr...
GA and stable for production across the Gemini API, Enterprise, and Antigravity, and now the default in the Gemini app and AI Mode in Search globally. Google pitches frontier-level intelligence at ~4x the speed of comparable models, priced at $1.50/$9 per 1M tokens, 1M-token c...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.