Deepseek makes its 75 percent discount permanent, pricing output tokens at least 34x below GPT-5.5
Source published ·Modelwire updated
Original coverage: The Decoder ↗·How Modelwire adds context

The development
DeepSeek has made its 75 percent discount on V4-Pro permanent, according to The Decoder’s May 23 report. The reported rates are $0.435 per million uncached input tokens and $0.87 per million output tokens. The output rate is about 34.5 times below GPT-5.5’s listed $30 per million at that time. This unit-price gap matters for token-intensive applications, but token consumption and output quality determine whether a completed task actually costs less.
Modelwire’s AI-generated summary of coverage from The Decoder.
Modelwire analysis
Analyst takeOur AI-generated reading of the wider context and the next developments to watch.
Making the discount permanent changes the planning horizon: teams can compare a recurring published rate rather than assume a promotion will end. It still leaves the economics of their own workloads unresolved.
For an agent that calls a model repeatedly, a token-price difference can accumulate across a workflow. But cheaper tokens do not guarantee cheaper accepted results if the model needs more attempts or produces errors. Teams should compare the same tasks, quality criteria and total usage before switching.
Watch for task-level comparisons reporting token consumption, retries, latency and accepted outputs. Those measurements would show whether the advertised unit-price advantage survives actual deployment constraints.
This interpretation is generated from the summary above and available source metadata. Our methodology · Report an error
MentionsDeepseek · Deepseek V4-Pro · OpenAI · GPT-5.5 · Anthropic · Google
How this coverage is produced
Modelwire uses AI to generate summaries and context from source headlines, snippets, and selected archive coverage. Automated checks do not verify every claim, and items are not routinely reviewed by a person before publication. Zacaria Solis operates the site. Read the linked source for the full evidence and report errors through our corrections process.
Modelwire summarizes, we don’t republish. The full content lives on the-decoder.com. If you’re a publisher and want a different summarization policy for your work, see our takedown page.