deepseek-v4-pro, deepseek-v4-pro-0813, deepseek-v4-flash and deepseek-v4-flash-0731 run on DeepSeek’s own official relay and are now on the new peak-tier rates; deepseek-v4-flash-260425, deepseek-v4-pro-260425 and deepseek-v4-flash-ga-260731 are relayed from BytePlus (ByteDance’s overseas Volcano Engine).
Current rates per 1M tokens (prompt / completion / cache read) — pick whichever fits:
deepseek-v4-flash,deepseek-v4-flash-0731: $0.44 / $1.32 / $0.0141 —— DeepSeek official relaydeepseek-v4-pro,deepseek-v4-pro-0813: $1.32 / $3.96 / $0.044 —— DeepSeek official relaydeepseek-v4-flash-ga-260731: $0.44 / $1.32 / $0.0136 —— BytePlus relaydeepseek-v4-flash-260425: $0.14 / $0.28 / $0.028 —— BytePlus relay, April snapshotdeepseek-v4-pro-260425: $1.74 / $3.48 / $0.15 —— BytePlus relay, April snapshot
deepseek-v4-flash-260425 costs roughly a third of the current release on prompt tokens and a fifth on completion tokens, a clear win for batch work with mostly fresh context and little prefix reuse. But its cache read at $0.028 is about twice the $0.0141 of the current release, so long conversations and long system prompts — anything with a high cache-hit rate — end up more expensive. Same shape for deepseek-v4-pro-260425: cheaper completions, pricier prompts and cache reads.
Recharge bonuses stack on all of the versions above. For the rate change itself and the peak-window conversion, see the deep dive.
← Back to Live Updates · 📚 Monthly archive · 🔗 Promotions