Skip to main content
Frontier ModelsPrice War· 3 min read· in Technology

OpenAI and Anthropic Slash API Prices in Simultaneous Frontier Model Launches

Anthropic's Claude Opus 5.5 and OpenAI's GPT-6 Sol and Luna debuted hours apart, halving the cost of autonomous coding agents for developers.

By Naina Verma

High-Volume Integrators 45%Efficiency and Quality Advocates 40%Market Analysts 15%
High-Volume Integrators
Prioritize raw API cost and speed for scaling automated workflows.
Efficiency and Quality Advocates
Focus on the total cost of task completion, including retries and token verbosity.
Market Analysts
View the releases through the lens of corporate strategy and impending IPOs.

Perspectives this story doesn't cover

  • Open-source model developers facing intense price pressure from proprietary vendors.
  • Security researchers evaluating the new containment boundaries in production.

Why it matters

The cost of running autonomous AI agents just plummeted. For software teams and businesses, this dual launch means complex, multi-step AI workflows that were previously too expensive to run at scale are now economically viable.

Developers building autonomous coding agents and high-volume text pipelines just saw their operating costs drop by roughly half. On Tuesday, September 22, Anthropic and OpenAI launched competing frontier models within hours of each other, trading benchmark records for aggressive price cuts that immediately alter the economics of deploying artificial intelligence at scale.[4][5]

Anthropic moved first, releasing Claude Opus 5.5 with a 20 percent reduction in input pricing and a 60 percent cut to cache reads compared to its predecessor. Minutes later, OpenAI announced GPT-6 Sol and GPT-6 Luna, halving the application programming interface (API) costs of their GPT-5.6 equivalents.[1][2][5]

Both companies are framing these releases around "agentic coding"—the ability of a model to navigate a codebase, run terminal commands, and iteratively debug its own work without human prompting. Anthropic reports that Opus 5.5 scores 66.4 percent on the Terminal-Bench 4.0 evaluation, a notable jump from the 55.8 percent achieved by its recent Fable 5.1 model. OpenAI counters that GPT-6 Sol matches the older Claude Opus 5 on the OSWorld 2.0 offline benchmark while operating at an 80 percent lower cost per task.[1][2][5]

The sticker prices reflect a market shifting from capability demonstrations to production efficiency. Claude Opus 5.5 charges $4 per million input tokens and $20 per million output tokens. OpenAI's GPT-6 Sol undercuts that baseline at $2 per million input tokens and $10 per million output, while the lightweight GPT-6 Luna drops to $0.10 and $0.50, respectively.[1][2][3][6]

GPT-6 Sol undercuts Claude Opus 5.5 on raw token price, while Luna targets high-volume tasks.
The sticker prices reflect a market shifting from capability demonstrations to production efficiency.

However, the raw token price does not dictate the final bill. Anthropic claims that Opus 5.5 is significantly less verbose, requiring fewer output tokens to complete a given task. Independent testing by APIMaster indicates that while GPT-6 Sol is the cheaper default for high-volume, repeated coding loops, Opus 5.5 often completes difficult repository tasks with fewer retries, narrowing the actual cost gap in production environments.[2][6]

The rollout for both model families was immediate. Opus 5.5 is live on the Claude Developer Platform and across major cloud providers, including Amazon Web Services and Microsoft Azure. OpenAI pushed Sol and Luna to its API, ChatGPT Work, and Codex tiers, while GitHub immediately integrated both new OpenAI models into its Copilot enterprise plans.[1][2][5][7]

The aggressive pricing arrives at a critical moment for both organizations. Bloomberg reports that Anthropic is preparing for an initial public offering as soon as this fall, making enterprise adoption and token efficiency key metrics for prospective investors. OpenAI, meanwhile, is using Sol and Luna to fill the pricing tiers below its flagship GPT-6 Astra, ensuring developers do not defect to cheaper competitors for routine tasks.[3][4][6]

Anthropic reports a significant jump in agentic coding capabilities for Opus 5.5 compared to its predecessors.

Beneath the pricing war, both models introduce structural changes to how they process requests. Opus 5.5 now runs "adaptive thinking" by default, forcing developers to control reasoning depth through an effort parameter rather than toggling it on or off. On the safety front, Anthropic claims the new model attempts to circumvent containment boundaries 85 percent less often than Opus 5.[2][5][6]

The immediate consequence of Tuesday's dual launch is a buyer's market for enterprise AI. With GitHub Chief Product Officer Mario Rodriguez noting that the new models solve terminal tasks in "less than half the steps" of previous generations, the bottleneck for autonomous coding is no longer the raw capability of the models, but the willingness of engineering teams to trust them with unmonitored execution.[5]

What to know

  • Anthropic released Claude Opus 5.5, cutting input token costs by 20 percent and cache reads by 60 percent.
  • OpenAI immediately countered with GPT-6 Sol and Luna, halving the API prices of their GPT-5.6 equivalents.
  • Both companies are heavily marketing the models' capabilities in agentic coding and autonomous terminal use.
  • Independent testing suggests GPT-6 Sol wins on raw token price, while Opus 5.5 requires fewer retries for complex tasks.
  • The aggressive price cuts arrive as Anthropic reportedly prepares for an initial public offering this fall.

Where opinion splits

Enterprise Developers

Engineering teams prioritizing cost-efficiency and high-volume execution.

For developers running thousands of automated tasks daily, the raw API cost is the deciding factor. This camp favors GPT-6 Sol and Luna for routine data extraction, basic code generation, and high-volume text processing. They view the 50 percent price cut as the critical enabler for scaling AI agents across entire codebases without breaking budget constraints.

Quality-First Engineers

Teams focused on complex, long-horizon coding tasks where accuracy outweighs token price.

Engineers tackling difficult repository migrations or multi-step debugging prefer Claude Opus 5.5. They argue that a model's sticker price is deceptive if it requires multiple retries to produce working code. By completing tasks in fewer steps and communicating more clearly, this camp believes Opus 5.5 ultimately costs less in compute time and human oversight.

Sources

Source coverage

7 outlets

3 viewpoints surfaced

High-Volume Integrators 45%Efficiency and Quality Advocates 40%Market Analysts 15%
  1. [1]OpenAIHigh-Volume Integrators

    Introducing GPT-6 Sol and Luna

    Read on OpenAI
  2. [2]AnthropicEfficiency and Quality Advocates

    Introducing Claude Opus 5.5

    Read on Anthropic
  3. [3]MashableHigh-Volume Integrators

    ChatGPT-6 Sol and Luna are here: Pricing and benchmarks data

    Read on Mashable
  4. [4]BloombergMarket Analysts

    Anthropic Unveils Cheaper Claude Opus 5.5 Ahead of Expected IPO

    Read on Bloomberg
  5. [5]SiliconANGLEMarket Analysts

    Anthropic releases Claude Opus 5.5 and cuts price 20%, OpenAI answers with GPT-6 Sol and Luna

    Read on SiliconANGLE
  6. [6]APIMasterEfficiency and Quality Advocates

    GPT-6 Sol vs Claude Opus 5.5: Cost, Coding and Agent Workflows

    Read on APIMaster
  7. [7]GitHub BlogHigh-Volume Integrators

    OpenAI's GPT-6 Sol and GPT-6 Luna now available

    Read on GitHub Blog

Comments

Stay informed

Every angle. Every day.

Get Technology stories with full source coverage and perspective breakdowns delivered to your inbox.