Skip to main content
Live
Main content

OpenAI cuts GPT-5.6 Luna 80% as Anthropic undercuts its own flagship

US token prices have dropped nearly a quarter since mid-July as DoorDash and Airbnb shift workloads to Chinese models from Moonshot and DeepSeek.

Jaeden Schafer
Editor in Chief · · 5 min read
OpenAI logo

OpenAI cut the price of GPT-5.6 Luna, its self-described fastest and most affordable model, by 80 percent, and Anthropic launched Claude Opus 5 at half the price of its flagship Fable 5. Together the moves have pulled token prices from leading US labs down almost a quarter since mid-July, according to Silicon Data's token price index. Both companies are plotting IPOs at trillion-dollar valuations while investors hunt for evidence that the industry's spending can throw off returns.

OpenAI dropped GPT-5.6 Luna from $1 to $0.20 per million input tokens and from $6 to $1.20 per million output tokens. Anthropic priced Opus 5 at $5 per million input tokens and $25 per million output tokens, positioned explicitly at half the cost of Fable 5. This week Anthropic also called off a planned September price increase on Sonnet 5.

The immediate pressure is coming from China. Moonshot's Kimi K3 and DeepSeek's V4 Flash have closed the performance gap enough that cost-conscious buyers are willing to switch, and they are switching from inside Silicon Valley itself. DoorDash and Airbnb have both said they have started using Chinese-made models to rein in AI bills.

Key facts

  • 01OpenAI cut GPT-5.6 Luna input tokens from $1 to $0.20 per million and output from $6 to $1.20 per million — an 80% reduction.
  • 02Anthropic launched Claude Opus 5 at $5 per million input tokens and $25 per million output, half the price of its Fable 5 flagship.
  • 03Token prices from leading US labs have fallen almost 25% since mid-July, according to Silicon Data's index.
  • 04DoorDash and Airbnb have started using Chinese-made models to rein in AI bills.
  • 05Both OpenAI and Anthropic are plotting IPOs at trillion-dollar valuations.

That switch is happening at a moment when Anthropic and OpenAI are shifting enterprise customers away from flat subscriptions toward usage-based billing, where the meter runs on tokens consumed. Some corporate users have responded by capping AI usage internally, others by piloting cheaper alternatives outright. When enterprise finance teams see the invoice climb, the open Chinese model that a developer can download, tweak, and self-host starts looking a lot more attractive than a proprietary API.

Artificial Analysis, which benchmarks models on math, science, coding, and reasoning, found Opus 5 at medium effort delivered similar performance and cost per task to Kimi K3 at max effort. GPT-5.6 Luna at max effort performed similarly to DeepSeek V4 Flash at max, but cost just under twice as much per task. In other words, US mid-tier models are now roughly at parity on quality with Chinese open competitors, but the Americans are still charging a premium — which is exactly what the recent cuts are designed to close.

The US labs have cut the middle and are defending the top.
Mantas Lukauskas, AI tech lead at Hostinger

Mantas Lukauskas, AI tech lead at website host Hostinger, which has run large language models in production since 2020, said prices for the very best models were flat to rising even as the middle collapses. He framed the current moment as the first real test of whether Anthropic and OpenAI can protect the price of their most advanced offerings while ceding ground below.

Anthropic and OpenAI declined to comment on the competitive framing. A person close to Anthropic pushed back on the read that Opus 5's pricing was a response to Chinese rivals, saying the below-flagship price point simply reflected how the

family of models is built, so there's no connection to competitors.
Person close to Anthropic, Anthropic source

The nuance the price sheet hides is that headline token cost isn't the same as cost per task. More capable models can finish a job in fewer tokens or fewer attempts, and most modern models expose an effort setting that trades compute for accuracy. A model that looks expensive per token can be cheaper per completed job — which is the argument the US labs will lean on hardest as the pricing conversation moves to boardrooms.

Related · from this week
Anthropic Q2 revenue hits $11.5B, a 14-fold jump ahead of IPO
Jaeden Schafer · 5 min read →

Still, the trajectory is unmistakable. A nearly 25 percent drop in effective prices from the frontier US labs in about a month is the sharpest downward move the market has seen, and it happened because customers had a real alternative. Open Chinese weights change the negotiation. When DoorDash and Airbnb are on the record about switching, every other enterprise procurement team gets to point at that filing in its next contract discussion.

The strategic question for OpenAI and Anthropic on the road to trillion-dollar listings is whether the top of the stack — the truly frontier models where American labs still lead — throws off enough margin to fund the capex, or whether the middle tier they just torched was carrying more of the P&L than they'd like to admit. Cutting the middle and defending the top is a coherent strategy only if the top holds. Kimi K3 and DeepSeek V4 Flash are already close enough at mid-tier that the frontier gap is the only pricing power left, and Chinese labs have not shown any inclination to stop closing it.

ShareXLinkedInEmail
AI Box

Every AI model. One chat.

The latest models from ChatGPT, Claude, Gemini, Sora, ElevenLabs — 80+ models in a single chat. Compare answers side by side. Pick the best one every time.

  • ChatGPT, Claude, Gemini, Grok, DeepSeek — in one chat
  • Generate images & video with Sora, Veo, Ideogram
  • Compare any two models side by side
  • From $8.99/mo · 80+ models, all included
Try AI Boxaibox.ai
Trusted by 3,000+ teams
Got a tip?

Working on something we should cover, or seeing a story we missed? Send leads, documents, or feedback to hello@aichatdaily.com. For sensitive tips, see our secure tips page for Signal and PGP options.

Spotted an error? Email hello@aichatdaily.com with the URL and the issue, or read our full corrections policy.

AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at

Keep reading

More from Business

Anthropic logo
Business

Anthropic Q2 revenue hits $11.5B, a 14-fold jump ahead of IPO

Claude's maker crossed $11.5B in a single quarter and posted positive adjusted operating income as it lines up a fall listing.

Jaeden Schafer5 min read
US names six Chinese AI firms in industrial-scale distillation campaign
Security

US names six Chinese AI firms in industrial-scale distillation campaign

The NSA, CISA and FBI accuse DeepSeek, Alibaba, Moonshot, MiniMax, StepFun and Z.AI of siphoning capabilities from Claude, GPT, Gemini and Grok.

Jaeden Schafer5 min read
Microsoft logo
Business

Microsoft coaches sales team to pitch against OpenAI and Anthropic

Executives at an internal FY27 strategy meeting told salespeople to frame Copilot as faster and more secure than Claude and rival models.

Jaeden Schafer4 min read