OpenAI launches GPT-6 Sol and Luna with API prices cut in half, confirmed as permanent rates

2026-09-23·8 min read

On Tuesday, September 22, 2026, OpenAI released two new members of the GPT-6 family, GPT-6 Sol and GPT-6 Luna, cutting API prices by roughly half against the previous GPT-5.6 generation. According to VentureBeat, an OpenAI spokesperson confirmed to the publication that these are permanent prices, not promotional or introductory rates. CNBC reports the two models are new tiers sitting below the GPT-6 flagship Astra: Sol targets more complex workloads, Luna targets high-volume repetitive tasks. The same morning, Anthropic launched Claude Opus 5.5, putting the two labs' new models right on top of each other in the price table.

Start with the numbers, because the entire weight of this story sits in the numbers. Per VentureBeat, GPT-6 Sol is priced at $2 per million input tokens and $10 per million output tokens; the current GPT-5.6 Sol rate is $4 input and $20 output, so the cut is exactly 50% in both directions. GPT-6 Luna is priced at $0.10 input and $0.50 output per million tokens; GPT-5.6 Luna currently costs $0.20 input and $1.20 output, a 50% cut on input and roughly 58.3% on output. CNBC reports both companies attributed the reductions to improvements in inference efficiency and prompt caching. In other words, this is not a subsidy to buy share; the same work genuinely costs less compute.

The second thing worth pinning down is how the three tiers divide the work. CNBC reports Sol sits below Astra and is meant for more complex workloads such as coding, while Luna targets what the company calls high-volume tasks, for example extracting information or summarizing documents. In VentureBeat's framing, Sol aims at the complex work developers and knowledge workers perform repeatedly, such as building features, reviewing code, debugging and analyzing data; Luna covers more tightly defined jobs such as summarization, extraction and answering straightforward questions; Astra remains for the most complex, multi-faceted projects that cross several kinds of media, along with harder scientific and mathematical problems. OpenAI says both new models improve on their respective 5.6 predecessors on benchmarks while staying less capable than the flagship Astra released earlier this month.

The third angle is who these prices sit next to, and that is more informative than the size of the cut. Per VentureBeat's comparison, Sol's $2 input and $10 output land exactly on Anthropic's Claude Sonnet 5, whose introductory $2/$10 pricing Anthropic made permanent in August. Anthropic's Claude Opus 5.5, launched the same Tuesday morning, is priced at $4 input and $20 output, meaning Sol's output rate is half of the new Opus; Anthropic's top general-access tier, Claude Fable 5.1, costs $10 input and $50 output per million tokens, five times Sol. Further down, Luna's $0.10 and $0.50 undercut Sonnet 5 by 95% on both input and output. On Google's side, Gemini 3.8 Flash, introduced earlier this month, costs $0.75 input and $3.75 output under introductory pricing that rises to $1.50 and $7.50 on January 1, 2027, so even after the increase Gemini 3.8 Flash stays cheaper than Sol. The pricing structures differ, though: Google time-limited its current Flash rate, while OpenAI says Sol's $2/$10 has no promotional expiration.

The fourth thread is timing, and this backdrop matters more than the discount itself. CNBC reports this is the first release from either frontier lab since the debate over whether the industry should slow down reached a fever pitch. The report notes that former Anthropic researcher Jacob Coxon set off that debate by posting on X on September 8 that he had quit his job, warning that both labs were gambling with our lives, after which OpenAI chief executive Sam Altman and Tesla and SpaceX chief executive Elon Musk joined Anthropic chief executive Dario Amodei's call. CNBC also points out that both companies are working to satisfy customers looking for more cost-effective models and seeking to rein in AI spending, under pressure from firms offering cheaper open-weight models, naming three Chinese companies: Alibaba, Moonshot AI and DeepSeek.

The fifth thread is what this means for people actually building products. VentureBeat's framing is that for enterprises this segmentation matters because the economics of AI agents increasingly depend on how many model calls a workflow makes, how much context gets replayed and whether a company really needs a frontier model for every step. That is where the Sol and Luna tier lands: move the simple, repetitive, well-defined calls to the cheap tier, keep the steps that genuinely need complex reasoning on the expensive one, and the bill difference will come from how you allocate calls rather than from model capability gaps. Anthropic's numbers cross-check the point: cache reads on Opus 5.5 fall to $0.20 per million tokens, 60% below Opus 5's $0.50, and cache reads are exactly where the bulk of cost lives in agentic and coding workloads. In other words, how efficiently you reuse context may matter to your bill as much as which model you pick.

Two things remain to watch. First, a permanent price is not the same as a free lunch: both new models stay below Astra in capability, and which step you can downgrade and which you cannot is something only your own evaluation run will tell you, since public benchmarks are best used to rule out clearly unsuitable options. Second, the fact that calls for a slowdown coexist with this release cadence is itself the story: two labs shipping a new model each on the same day indicates the pace of competition has not slowed because of public debate. Every fact and figure here comes from VentureBeat, CNBC and The New Stack reporting, with no speculation added.

🤔 Frequently Asked Questions

How much do GPT-6 Sol and Luna actually cost?

Per VentureBeat, GPT-6 Sol costs $2 per million input tokens and $10 per million output tokens, while GPT-6 Luna costs $0.10 input and $0.50 output. For comparison, GPT-5.6 Sol was $4 and $20, and GPT-5.6 Luna was $0.20 and $1.20. That makes Sol 50% cheaper in both directions, and Luna 50% cheaper on input and roughly 58.3% cheaper on output.

Is this price cut a limited-time promotion?

No. According to VentureBeat, an OpenAI spokesperson confirmed to the publication that the GPT-6 Sol and Luna rates are permanent prices, not promotional or introductory pricing. The report also notes that both companies attributed the cuts to improvements in inference efficiency and prompt caching. Note the contrast: Google's current Gemini 3.8 Flash rate is explicitly time-limited introductory pricing and rises on January 1, 2027.

How should you choose between Sol, Luna and Astra?

As positioned in the reporting: Astra is the flagship, reserved for the most complex projects spanning multiple kinds of media and for harder scientific and mathematical problems; Sol sits below Astra and targets recurring complex work such as coding, code review, debugging and data analysis; Luna is the high-volume tier for more tightly defined jobs such as summarization, extraction and answering straightforward questions. VentureBeat's caveat is that the point of tiering is to classify the calls inside a workflow, not to pick one model for an entire product.

Why did both companies ship new models on the same day?

CNBC's explanation is that both labs face pressure from cheaper open-weight models, while customers are looking for more cost-effective options and trying to rein in AI spending, naming three Chinese companies: Alibaba, Moonshot AI and DeepSeek. The report also notes this is the first release from either lab since the debate over slowing the industry intensified, a debate that started with former Anthropic researcher Jacob Coxon's resignation post on September 8.

🛠️ Recommended Tools

  • AI Token CounterThe value of a price cut depends entirely on your own call volume. Measure the token count of your prompts and context first, then multiply by Sol's $2 and $10 or Luna's $0.10 and $0.50 to see what a downgrade actually saves. Guessing at token counts is the most common form of self-deception in cost optimization.
  • AI Code ReviewerSol is positioned squarely at coding, which is exactly where tiering experiments are easiest: run your diffs through the cheap tier first, then hand only the questionable parts to the expensive one for a second pass. A week of that comparison gives you a more reliable answer than any leaderboard, because it is an answer measured on your own repository.
  • CSV to JSON ConverterLuna is built for extraction and summarization, tasks that usually end with structured data going into a store. Fix the downstream format as a JSON schema and prepare sample records first, and swapping model tiers becomes a config change. If you keep reshaping the format by hand, the money a price cut saves gets spent again on rework.

Summary

On September 22, 2026, OpenAI released GPT-6 Sol and GPT-6 Luna, cutting API prices by roughly half against the prior GPT-5.6 generation: Sol costs $2 per million input tokens and $10 per million output tokens, exactly half of GPT-5.6 Sol's $4 and $20, while Luna costs $0.10 input and $0.50 output against GPT-5.6 Luna's $0.20 and $1.20, a 50% cut on input and roughly 58.3% on output. An OpenAI spokesperson confirmed to VentureBeat that these are permanent prices rather than promotions, with the reductions attributed to inference efficiency and prompt caching. Sol targets more complex workloads such as coding, Luna handles high-volume tasks such as summarization and extraction, and Astra remains for the hardest problems. Anthropic launched Claude Opus 5.5 the same morning at $4 input and $20 output, making Sol's output rate half of it, while Google's Gemini 3.8 Flash remains explicitly time-limited introductory pricing. CNBC notes this is the first release from either lab since the debate over slowing the industry intensified, against a backdrop of customers reining in AI spending and competition from open-weight models. Every fact and figure here comes from VentureBeat, CNBC and The New Stack reporting, with no speculation added.

Sources: VentureBeat: OpenAI releases GPT-6 Sol and Luna models, slashing API costs 50% or more
CNBC: Anthropic and OpenAI roll out cheaper models in first release since call for slowdown
The New Stack: OpenAI releases GPT-6 Sol and Luna — and cuts token prices in half