DeepSeek Implements Steep API Price Increases for DeepSeek V4 Pro, Raising Some Rates Up to 12xFollowing the rollout of DeepSeek V4 Pro, leading Chinese AI lab DeepSeek has officially implemented its previously announced API price hikes. Standard usage costs have tripled across core models, with specific cached prompt storage rates jumping by as much as 12x.
Under the revised fee structure, standard input and output token pricing for both DeepSeek V4 Flash and DeepSeek V4 Pro have risen 3x and 4.6x, respectively. The most dramatic price escalation affects DeepSeek V4 Pro’s Prompt Cache storage, which surged from $0.003625 per million tokens to $0.044 per million tokens marking an approximate 12-fold increase.
Despite the substantial rate increases, DeepSeek maintains a off-peak discount structure, offering developers a 50% price reduction for API calls processed during non-peak hours. The new pricing matrix takes effect starting August 16.
Why have alert caching fees increased dramatically (12x)? Caching alert messages allows developers to reuse complex system alert messages, long codebases, or large reference documents without paying the full computing fee for every API call. Because developers rely heavily on large context windows for enterprise-level coding agents and RAG pipelines, maintaining active cache memory across GPU clusters places a massive VRAM burden on DeepSeek, forcing them to reprice caching to cover the actual hardware costs.
By offering half-price pricing during off-peak server hours, DeepSeek is actively attempting to rebalance server load across global time zones. This will encourage enterprise customers performing non-urgent batch tasks (such as overnight code rework, data processing, or model evaluation) to schedule API calls outside of peak business hours in Asia and Europe.
Even after a 3x to 4.6x price increase, DeepSeek's updated API rates remain significantly cheaper than leading Western models like the Claude 3.5 Sonnet or OpenAI's core model. Although the price increase narrows the profit margin, DeepSeek continues to position itself as a cost-effective option for developers building high-volume LLM applications.
Source: DeepSeek
DeepSeek Implements Steep API Price Increases for DeepSeek V4 Pro, Raising Some Rates Up to 12xFollowing the rollout of DeepSeek V4 Pro, leading Chinese AI lab DeepSeek has officially implemented its previously announced API price hikes. Standard usage costs have tripled across core models, with specific cached prompt storage rates jumping by as much as 12x.
Under the revised fee structure, standard input and output token pricing for both DeepSeek V4 Flash and DeepSeek V4 Pro have risen 3x and 4.6x, respectively. The most dramatic price escalation affects DeepSeek V4 Pro’s Prompt Cache storage, which surged from $0.003625 per million tokens to $0.044 per million tokens marking an approximate 12-fold increase.
Despite the substantial rate increases, DeepSeek maintains a off-peak discount structure, offering developers a 50% price reduction for API calls processed during non-peak hours. The new pricing matrix takes effect starting August 16.
Why have alert caching fees increased dramatically (12x)? Caching alert messages allows developers to reuse complex system alert messages, long codebases, or large reference documents without paying the full computing fee for every API call. Because developers rely heavily on large context windows for enterprise-level coding agents and RAG pipelines, maintaining active cache memory across GPU clusters places a massive VRAM burden on DeepSeek, forcing them to reprice caching to cover the actual hardware costs.
By offering half-price pricing during off-peak server hours, DeepSeek is actively attempting to rebalance server load across global time zones. This will encourage enterprise customers performing non-urgent batch tasks (such as overnight code rework, data processing, or model evaluation) to schedule API calls outside of peak business hours in Asia and Europe.
Even after a 3x to 4.6x price increase, DeepSeek's updated API rates remain significantly cheaper than leading Western models like the Claude 3.5 Sonnet or OpenAI's core model. Although the price increase narrows the profit margin, DeepSeek continues to position itself as a cost-effective option for developers building high-volume LLM applications.
Source: DeepSeek
Comments
Post a Comment