AI Briefing
KO
Pick

DeepSeek Makes V4 Pro Price Discount Permanent

·2026.05.23 21:51

Key point

DeepSeek has permanently cut DeepSeek-V4-Pro API prices to 1/4 of the previous level and significantly lowered cache hit prices as well.

Details

Even after the 75% discount promotion ends, DeepSeek-V4-Pro API prices will remain permanently at 1/4 of the original price.

The supported models are DeepSeek-V4-Flash and DeepSeek-V4-Pro, both of which support Thinking Mode and non-thinking mode (the default is Thinking Mode). The context length is 1M, and the maximum output is 384K tokens.

The key changes and features are as follows:

  • Cache hit price reduction: The input cache hit price for all models has been lowered to 1/10 of the launch price.
  • Per-model pricing (per 1M tokens):
    • DeepSeek-V4-Flash: Input (cache miss) $0.14 / Output $0.28
    • DeepSeek-V4-Pro: Input (cache miss) $0.435 / Output $0.87
  • Concurrency limits: The Flash model has a limit of 2500, and the Pro model has a limit of 500.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.