Skip to main content
Back to News
DeepSeek V4 to launch mid-July with peak-time API pricing, doubling off-peak rates
Product
2 min read
CN

DeepSeek V4 to launch mid-July with peak-time API pricing, doubling off-peak rates

The AMW Read

Novelty 2: dynamic peak/off-peak pricing is new for DeepSeek and rare among frontier model labs. Significance 2: segment-level impact on inference pricing norms and competitive positioning of V4 vs. Western frontier models.
NoveltySignificance
Foundation Models · Player Map
DeepSeek AI
DeepSeek AI

Foundation Models / LLMs

View Company Profile

DeepSeek V4 to launch mid-July with peak-time API pricing, doubling off-peak rates

The DeepSeek team confirmed Monday that the official release of DeepSeek V4 is scheduled for mid-July, building on the earlier preview release with enhanced performance in agent-based task execution, mathematical reasoning, and code generation. The model lineup will standardize on a 1-million-token context window. A new API pricing plan introduces peak and off-peak pricing for the first time: peak hours run 9:00-12:00 and 14:00-18:00 daily, with usage charged at twice the off-peak rate.

Why it matters: DeepSeek’s explicit time-based API pricing represents a structural shift in how frontier model labs manage inference demand. As the segment moves toward consumption-based monetization, this dynamic pricing approach mirrors cloud-compute elasticity models—a signal that the Chinese foundation-model market is maturing past flat-rate access. The 1M-token context window also matches or exceeds Western frontier models, keeping DeepSeek a legitimate competitor in the global foundation-model race, particularly for enterprise workloads requiring long-context reasoning.

Capability improvements in agentic task execution and code generation position the V4 release as a direct challenge to GPT-4.5 and Claude Opus 4.5-era models, especially in cost-sensitive Asian markets. Peak pricing may compress enterprise usage into off-peak windows, effectively reshaping developer behavior and inference load patterns. The naming “V4” rather than a point release continues DeepSeek’s habit of discrete generation jumps, a tactic that preserves marketing buzz but leaves the market wondering about incremental benchmarks between preview and final versions.

#DeepSeek #FoundationModels #APIPricing #InferenceEconomics #LongContext #AgenticAI

#DeepSeek V4#Foundation Models#API Pricing#Long Context#Agentic AI#China AI

How This Connects

Based on Foundation Models · Player Map

  1. 23h agoAlibaba to Raise US$10.2 Billion in New Shares to Fund Full-Stack AI PushAlibaba
  2. 1d agoAnthropic reportedly plans an October IPO at up to a $2 trillion valuation, even as ARR growth shows signs of deceleration.Anthropic
  3. 1d agoAnthropic hires Google TPU veteran Amir Salek to accelerate in-house AI chip developmentAnthropic
  4. 1w agoAlibaba's Qwen team has open-sourced Qwen3.8-27B, a 27-billion-parameter multimodal model designed f...Qwen
  5. 1w agoAlibaba has released Qwen 3.8 27B, an Apache 2.0-licensed open-weight dense model with 27 billion pa...Alibaba Qwen 3.8 27B launch
  6. 1mo agoDeepSeek V4 to launch mid-July with peak-time API pricing, doubling off-peak rates · THIS ARTICLE

Related News

More news from DeepSeek AI

Stay updated with the latest news and announcements from DeepSeek AI.

View all DeepSeek AI news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard