Skip to main content
Back to News
Technology
2 min read
CN

DeepSeek Launches V4.1-Flash at $0.003 per Million Off-Peak Cached Tokens, Claims Benchmark Edge Over GPT-5.6 and Claude Opus 5

The AMW Read

Extends DeepSeek's known pricing-and-benchmark arc rather than resolving it, but self-reported wins over top-tier rivals carry segment-wide pricing-competition significance.
NoveltySignificance
Foundation Models · Case Studies
DeepSeek AI
DeepSeek AI

Foundation Models / LLMs

View Company Profile

DeepSeek Launches V4.1-Flash at $0.003 per Million Off-Peak Cached Tokens, Claims Benchmark Edge Over GPT-5.6 and Claude Opus 5

DeepSeek has released V4.1-Flash, a new model priced at $0.003 per million tokens for off-peak cached input. The company says its benchmark results surpass OpenAI's GPT-5.6 and Anthropic's Claude Opus 5, positioning V4.1-Flash as the latest entrant in the price-performance race among frontier model labs.

The launch complicates DeepSeek's own pricing narrative. In mid-August, the company raised API prices as much as 12x and added peak-time pricing, citing compute constraints as it shifted from aggressive discounting toward sustainable revenue. An off-peak cached rate this low reads as demand segmentation, not a reversal — pushing latency-tolerant workloads into cheap windows while protecting margin at peak hours. It also arrives with a caution flag from DeepSeek's own record: independent testing of the prior V4 Flash found its agentic gains traced to post-training, and the model failed 46.2% of real-world agent tasks despite topping leaderboards. Self-reported wins over GPT-5.6 and Claude Opus 5 deserve the same scrutiny.

For builders, test the off-peak tier on production tasks rather than trusting leaderboard rank, given the benchmark-to-reality gap DeepSeek's own V4 Flash exposed. For investors, the timing matters: per the AI Market Watch index, DeepSeek remains privately funded by parent High-Flyer Quant with no outside VC capital as of July 2025 and a first external round in negotiation as of May 2026 (index coverage, not a full census) — a cheap, headline model launch supports a valuation story more than it proves a margin one.

#DeepSeek #LLMPricing #FoundationModels #AICompetition #FrontierModels

#DeepSeek#V4.1-Flash#LLM API pricing#GPT-5.6#Claude Opus 5

How This Connects

Based on Foundation Models · Case Studies

  1. 5h agoDeepSeek Launches V4.1-Flash at $0.003 per Million Off-Peak Cached Tokens, Claims Benchmark Edge Over GPT-5.6 and Claude Opus 5 · THIS ARTICLE
  2. 3d agoMistral Raises €3B Series D at €21B+ Valuation, Led by Samsung ElectronicsMistral
  3. 4d agoAnthropic assembles $517 billion in compute commitments over 11 monthsAnthropic
  4. 6d agoNvidia is reportedly weighing a $2.5 billion investment in Thinking Machines Lab that would value the startup at roughly $40 billion.Thinking Machines Lab
  5. 1w agoOpenAI launches Astra, its most capable model, as opaque-reasoning and AGI claims fuel a fresh safety debate.OpenAI
  6. 1w agoAnthropic Signs Reported $35B Lambda Cloud Deal for Texas AI ComputeAnthropic

Related News

More news from DeepSeek AI

Stay updated with the latest news and announcements from DeepSeek AI.

View all DeepSeek AI news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard