TechBooky AI Assistant
TechBooky AI Assistant
👋 Welcome to TechBooky AI Assistant

I can help with:
🔎 Tech News
🤖 AI Topics
💻 Gadgets
☁️ Cloud
✍️ Guest Posts
📢 Advertising
🔗 Backlinks
📩 Newsletter
  • AI Search
  • Cryptocurrency
  • Earnings
  • Enterprise
  • About TechBooky
  • Submit Article
  • Advertise With TechBooky
  • Contact Us
TechBooky
  • African
  • AI
  • Metaverse
  • Gadgets
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
  • African
  • AI
  • Metaverse
  • Gadgets
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
TechBooky
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
Home Artificial Intelligence

DeepSeek Locks in 75% Price Cut on V4 Pro, Undercutting Western AI Models by up to 25x

Paul Balo by Paul Balo
May 29, 2026
in Artificial Intelligence
Share on FacebookShare on Twitter
Share this story

Send it to someone who should read it.

f Facebook X X in LinkedIn wa WhatsApp tg Telegram @ Email
In Brief
  • DeepSeek has made permanent a 75% price cut on its flagship V4 Pro model, in a move that directly targets the cost structures behind today’s largest...
  • The company’s new pricing tiers put its models far below comparable offerings from major Western labs that are widely used in enterprise production.
  • According to the company’s published pricing comparisons, DeepSeek V4 Pro now comes in at around seven times cheaper on input tokens and 17 times cheaper on...

DeepSeek has made permanent a 75% price cut on its flagship V4 Pro model, in a move that directly targets the cost structures behind today’s largest AI systems. The company’s new pricing tiers put its models far below comparable offerings from major Western labs that are widely used in enterprise production.

According to the company’s published pricing comparisons, DeepSeek V4 Pro now comes in at around seven times cheaper on input tokens and 17 times cheaper on output tokens than models such as Anthropic’s Claude Sonnet and OpenAI’s GPT 5.5-Med. For organisations running large-scale workloads, that gap translates into significantly lower operating costs for similar classes of capability.

DeepSeek is also pushing aggressively at the lower end of the stack. Its V4 Flash model, a lighter, speed-optimised variant  is priced to undercut entry-tier options like Claude Haiku by roughly 10x to 25x. That positions V4 Flash as a budget-conscious choice for use cases that prioritise throughput and latency while still drawing on the same overall model family.

The pricing shifts are not positioned as a temporary promotion but as the output of architectural changes. DeepSeek attributes the cuts to a set of hardware–software optimisations, particularly around cache, that make its models more efficient to run at scale. While the company has not detailed every element of the stack in the provided material, it links the lower per-token prices directly to these efficiency gains.

The cost differential is especially stark when DeepSeek’s models are hosted natively in China. In that configuration, the company’s cache-read pricing is described as being 87 times cheaper than Western cloud offerings. That level of discount effectively sets a new price floor for cached inference in those regions, with implications for anyone running long-context or cache-heavy workloads.

The ripple effects are already visible among Chinese hardware and platform providers. Handset maker Xiaomi has moved to match DeepSeek’s cache-read pricing tier for its newly deployed MiMo architecture, mirroring the same level rather than trying to undercut it further. That indicates at least one major player sees DeepSeek’s pricing as a new reference point for AI infrastructure in its home market.

Also worth reading
DeepSeek’s Implied US$52bn Valuation Shows China’s AI Race Is Not Slowing DeepSeek Wants to Build Its Own AI Chip. Nvidia Should Be Paying Attention AI Coding Is Creating A New Burden For Open-Source Maintainers Paystack’s AI Checkout Experiment Could Change How Nigerians Pay Online Apple’s OpenAI Lawsuit Turns The AI Hardware Race Into A Legal Fight Sam Altman And Elon Musk Are Now Fighting Over Space Data Centres

DeepSeek is pairing its pricing story with benchmark data aimed at showing that V4 Pro is not just cheaper, but competitive on quality. The company’s model card for DeepSeek V4 Pro highlights external evaluations placing it close to Western frontier systems on several technical measures.

On coding-agent tasks, DeepSeek V4 Pro records a score of 80.6% on the SWE-bench Verified leaderboard, a benchmark that tracks performance on software engineering-related challenges. That result is presented as putting the model almost on par with top-tier Western systems that target similar workloads in enterprise development and automation.

For broader reasoning and technical understanding, DeepSeek cites an 87.5 score on the advanced MMLU-Pro technical index, a demanding benchmark used to assess higher-level reasoning across specialised domains. That figure places V4 Pro in what DeepSeek describes as the “elite” range on that test, reinforcing the argument that its pricing does not come at the expense of capability.

Both V4 Pro and V4 Flash belong to the same model family, with V4 Pro aimed at more demanding tasks and V4 Flash tuned for speed. DeepSeek characterises V4 Flash as a hyper-optimised, fast variant intended for deployments where responsiveness and cost-per-call are critical.

The combination of aggressive token pricing, cache-read discounts in China, and benchmarked performance near Western frontier models positions DeepSeek as a cost-focused challenger in the global AI ecosystem. How far that pressure reshapes pricing and infrastructure strategies elsewhere remains to be seen, but the new floor it has set particularly around cached inference is now public and explicit.

Related Reading

More contextual TechBooky stories selected from tags, categories and article context.

  • chatgpt-logo
    OpenAI Launches GPT-5.4 Mini and Nano Models
  • deepseek
    DeepSeek's Implied US$52bn Valuation Shows China's…
  • deepseek-ai-record
    China's DeepSeek Finally Launches A New AI Model
  • 54b7ab1d2c2521f83ae5d2da5f9d99321c370d24-2880x1620
    Claude Opus 5 Gives Anthropic A Cheaper Answer To…
  • 12
    Moonshot AI's Kimi K3 Tops Frontend Coding Leaderboard
  • alibaba qwen
    Alibaba Expands Qwen Lineup with New Mid-Sized AI Models
  • DO3EOFAEMFNYHCIFVH2KMVCOVI
    DeepSeek Update Threatens Google and ChatGPT Dominance
  • assets_task_01jryqpar7fd1vr3zjb9wj416t_img_0
    OpenAI Unveils GPT-4.1, Its Flagship AI Model
Keep Reading Smarter

Search TechBooky with AI

Use TechBooky's AI Search to explore the context behind this story and related coverage across the site.

Try AI Search
More On This Topic
Artificial Intelligence
Follow TechBooky

Follow TechBooky for more technology stories and newsroom updates.

f Facebook X X in LinkedIn ig Instagram wa WhatsApp

Tags: AIdeepseekDeepSeek-V4-Pro
Paul Balo

Paul Balo

Paul Balo is the founder of TechBooky and a highly skilled wireless communications professional with a strong background in cloud computing, offering extensive experience in designing, implementing, and managing wireless communication systems.

Search TechBooky
Open TechBooky AI Search Try the AI Assistant

BROWSE BY CATEGORIES

Receive top tech news directly in your inbox

subscription from
Loading

Freshly Squeezed

  • Claude Opus 5 Gives Anthropic A Cheaper Answer To The Fable 5 Problem July 25, 2026
  • Meta Makes Facebook Verified Free As AI Scams Make Real People Harder To Spot July 24, 2026
  • SAP Cloud Growth Eases Fears That AI Will Weaken Enterprise Software July 24, 2026
  • US Lawmakers Push AI Kill Switch Bill After OpenAI Rogue-Model Incident July 24, 2026
  • Airtel Money’s $61B Quarter Makes Its London IPO A Bigger Africa Fintech Story July 24, 2026
  • Intel Q2 Revenue Jumps As AI Compute Demand Lifts Chip Business July 24, 2026
  • AMD And Anthropic Deal Puts Real Pressure On Nvidia’s AI Chip Lead July 24, 2026
  • New York’s Data Centre Pause Shows AI Infrastructure Is Hitting Politics July 23, 2026
  • OpenAI Researcher’s $2B Drug Discovery Plan Shows AI Biotech Hype Is Back July 23, 2026
  • Airtel Africa Q1 Shows Mobile Money And Data Are Doing The Heavy Lifting July 23, 2026
  • ZainTECH And Nile Build AI-Ready Networks For Enterprises July 23, 2026
  • Google Cloud Boom Makes Alphabet AI Spending Look More Real July 23, 2026

Browse Archives

July 2026
M T W T F S S
 12345
6789101112
13141516171819
20212223242526
2728293031  
« Jun    

Quick Links

  • About TechBooky
  • Advertise With TechBooky
  • Contact us
  • Submit Article
  • Privacy Policy
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
  • African
  • Artificial Intelligence
  • Gadgets
  • Metaverse
  • Tips
  • AI Search
  • About TechBooky
  • Advertise With TechBooky
  • Submit Article
  • Contact us

© 2025 Designed By TechBooky Elite

Discover more from TechBooky

Subscribe now to keep reading and get access to the full archive.

Continue reading

We use cookies to ensure that we give you the best experience on our website. If you continue to use this site we will assume that you are happy with it.