📡 Breaking news
Analyzing latest trends...
AI Text-to-Speech.

DeepSeek V4 Pro Ends Cheap API Era Rates Jump Up to 12x Starting August 16.

DeepSeek V4 Pro Ends Cheap API Era Rates Jump Up to 12x Starting August 16.
DeepSeek Implements Steep API Price Increases for DeepSeek V4 Pro, Raising Some Rates Up to 12x

Following the rollout of DeepSeek V4 Pro, leading Chinese AI lab DeepSeek has officially implemented its previously announced API price hikes. Standard usage costs have tripled across core models, with specific cached prompt storage rates jumping by as much as 12x.

Under the revised fee structure, standard input and output token pricing for both DeepSeek V4 Flash and DeepSeek V4 Pro have risen 3x and 4.6x, respectively. The most dramatic price escalation affects DeepSeek V4 Pro’s Prompt Cache storage, which surged from $0.003625 per million tokens to $0.044 per million tokens marking an approximate 12-fold increase.

Despite the substantial rate increases, DeepSeek maintains a off-peak discount structure, offering developers a 50% price reduction for API calls processed during non-peak hours. The new pricing matrix takes effect starting August 16.

Why have alert caching fees increased dramatically (12x)? Caching alert messages allows developers to reuse complex system alert messages, long codebases, or large reference documents without paying the full computing fee for every API call. Because developers rely heavily on large context windows for enterprise-level coding agents and RAG pipelines, maintaining active cache memory across GPU clusters places a massive VRAM burden on DeepSeek, forcing them to reprice caching to cover the actual hardware costs.

By offering half-price pricing during off-peak server hours, DeepSeek is actively attempting to rebalance server load across global time zones. This will encourage enterprise customers performing non-urgent batch tasks (such as overnight code rework, data processing, or model evaluation) to schedule API calls outside of peak business hours in Asia and Europe.

Even after a 3x to 4.6x price increase, DeepSeek's updated API rates remain significantly cheaper than leading Western models like the Claude 3.5 Sonnet or OpenAI's core model. Although the price increase narrows the profit margin, DeepSeek continues to position itself as a cost-effective option for developers building high-volume LLM applications.

 

 

Source: DeepSeek 

💬 AI Content Assistant

Ask me anything about this article. No data is stored for your question.

Comments

Popular posts from this blog

When Agents Overreach Australian Case Study Sparks Debate over Autonomous AI Liability.

Etsy Layoffs 220 Roles Cut in Product and Engineering Restructure to Drive Faster Execution.

AI-Driven Cyberattacks Hit U.S. Corporate Giants Supply Chains Disrupted Across Healthcare and Retail.

Sony Reports Q1 FY2026 Revenue Hits ¥2.84 Trillion as Sensors and Music Offset Gaming Flatline.

Critical Security Flaw in AI Summarizer 'tl;dv' Exposes Confidential Client Meetings Globally.

Apple Battles AI Deepfakes in iOS 27 New Apple Reference Image Feature Authenticates Real Photos.

When AI Filters Fail DeepMind Safety Unit Sets Up Direct Pipeline Over Incorrect Rejections.