Fetching from the wire…
Public story · 2026-08-03 · high
A GitHub tool, sol-advisor, got 853 stars in two days doing the same split for OpenAI's GPT-5.6 tiers.
Why now: The read comes from two data points in the same coverage cycle: a 123-upvote Reddit report and a GitHub repo that hit 853 stars in two days, not a vendor pricing update.
A Reddit user switched to Sonnet 5 and watched Max-subscription rate limits vanish, per a 123-upvote post in r/ClaudeAI. They hadn't touched the model in three months before making the switch.
On a fixed-price Max plan, the constraint is how fast you burn quota, not per-token cost. That's a different optimization problem than the one pricing coverage usually solves. It flips the usual advice to save the expensive model for hard problems.
A GitHub project called sol-advisor picked up 853 stars in two days doing the same math for OpenAI's GPT-5.6 lineup. Implementation work runs in the cheap Luna and Terra tiers. Then a Sol-tier review, started with no prior context, has to sign off before anything ships. The review is a mandatory gate, per the repo.
That structure doesn't need GPT-5.6 or Claude specifically. Any provider with tiered models and a subscription cap creates the same incentive: do the typing cheap, spend the expensive model's context on judgment.
Quota burn rate, not per-token price, is becoming the variable that decides which model builders default to. Watch whether Anthropic or OpenAI start routing requests across tiers automatically. Developers are already hand-rolling the split themselves, in a Reddit thread and a two-day-old GitHub repo. Patterns that spread this fast tend to get productized.
Each link below shares sources, entities, or timing with this story.
Created August 1, ~403 stars/day, and it's a Shell repo, not a framework. It encodes a named architect role (Sol), two parallel implementation lanes (Luna and Terra), and a review pass by a fresh-context reviewer that cannot be skipped. The forced-fresh-context review is the t...
GPT-5.6 Luna went to $0.20 input / $1.20 output per million tokens on July 30. That's an 80% cut. Terra dropped 20%. Luna's input now undercuts Gemini 3.1 Flash-Lite ($0.25/$1.50) and sits at one-fifth of Claude Haiku 4.5's $1 input. Simon Willison covered the announcement and...
OpenAI's flagship of the newly previewed Sol/Terra/Luna family hits that speed tier on Cerebras hardware in July, with pricing from $5/$30 per 1M for Sol down to $1/$6 for Luna. The rollout stays partly gated to "trusted partners" at the US government's request, so the Cerebra...
Spotify's Portal team published Xirp on August 10: a vendor-neutral agentic development environment that manages concurrent sessions across Claude Code, Gemini CLI, and Codex, each session isolated in its own git worktree so dozens of agents can work the same codebase without...
The July 18 release notes bundle fixes that restore the full window, which means it had been silently degraded for some unspecified period. If you benchmarked those models in Codex over the past few weeks and found long-context performance underwhelming, you may have been meas...
July 9, across VS Code, Visual Studio, Copilot CLI, the cloud agent, github.com, GitHub Mobile, JetBrains, Xcode, and Eclipse. Sol is the high-reasoning tier at $5/1M in, $30/1M out, gated to Pro+/Max/Business/Enterprise. Terra is the balanced default at $2.50/$15. Luna is fas...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.