tencent cloud

このページは現在英語のみで、表示されている言語を準備しています。
关闭
【LLM Service TokenHub】【Agent Development Platform】DeepSeek-V4-Flash 0731 GA 、DeepSeek-V4-Flash-Vision-Exp (All Vendor Direct) Price Reduction Announcement
2026-09-09 18:33:44
LLM Service Platform TokenHub will reduce the prices of DeepSeek-V4-Flash-0731 GA (Vendor Direct) and DeepSeek-V4-Flash-Vision-Exp (Vendor Direct) starting from 12:00 (UTC+8) on September 10, 2026 (Beijing time). The inference input price will be reduced by 31.82%, the inference output price by 9.09%, and the cache hit price by 57.14%, allowing you to enjoy the same high-quality model services at a lower cost.

Adjustment Details
Model Name
Peak/Off-Peak Billing
Input Tokens
(USD / million tokens)
Output Tokens
(USD / million tokens)
Cache
(USD / million tokens)
DeepSeek-V4-Flash 0731 GA (Vendor Direct)
OFF-PEAK
Reduced from 0.22 to 0.15
Reduced from 0.66 to 0.6
Reduced from 0.007 to 0.003
PEAK
Reduced from 0.44 to 0.3
Reduced from 1.32 to 1.2
Reduced from 0.014 to 0.006
DeepSeek-V4-Flash-Vision-Exp (Vendor Direct)
OFF-PEAK
Reduced from 0.22 to 0.15
Reduced from 0.66 to 0.6
Reduced from 0.007 to 0.003
PEAK
Reduced from 0.44 to 0.3
Reduced from 1.32 to 1.2
Reduced from 0.014 to 0.006
For pricing of all models, please refer toModel Pricing

Adjustment Details
Model Name
Peak/Off-Peak Billing
Input Tokens
(credits / million tokens)
Output Tokens
(credits / million tokens)
Cache
(credits / million tokens)
DeepSeek-V4-Flash 0731 GA (Vendor Direct)
OFF-PEAK
Reduced from 150 to 100
Reduced from 450 to 400
Reduced from 5 to 2
PEAK
Reduced from 300 to 200
Reduced from 900 to 800
Reduced from 10 to 4
DeepSeek-V4-Flash-Vision-Exp (Vendor Direct)
OFF-PEAK
Reduced from 150 to 100
Reduced from 450 to 400
Reduced from 5 to 2
PEAK
Reduced from 300 to 200
Reduced from 900 to 800
Reduced from 10 to 4

img