DeepSeek V4 Flash 0731 API price comparison
DeepSeek V4 Flash is a state-of-the-art large language model combining efficient sparse attention, strong reasoning, and integrated agent capabilities for robust long-context understanding and versatile AI applications.
Cheapest overall
AIHubMix
0.42 USD/M tokens combined per 1M tokens
Lowest input
AIHubMix
0.14 USD/M tokens input per 1M tokens
Lowest output
AIHubMix
0.28 USD/M tokens output per 1M tokens
Last refreshed
Sep 6, 2026
5 listed platforms
Best listed option
DeepSeek V4 Flash 0731 is currently cheapest overall on AIHubMix, with input at 0.14 USD/M tokens and output at 0.28 USD/M tokens per million tokens.
Pricing insight
Across 5 listed routes, DeepSeek V4 Flash 0731 is cheapest overall on AIHubMix at 0.42 USD/M tokens. The median combined price is 10.80 USD/M tokens, so the cheapest route is about 96% below the median listed route.
-
Lowest total cost
AIHubMix
Best first check for balanced coding-agent sessions: 0.42 USD/M tokens.
-
Prompt-heavy work
AIHubMix
Useful for large repositories and long context reads: 0.14 USD/M tokens input.
-
Response-heavy work
AIHubMix
Useful for long answers, edits, and agent loops: 0.28 USD/M tokens output.
Median combined price
10.80 USD/M tokens
Baseline for judging whether the cheapest route is unusually low.
Cache price coverage
80%
4 of 5 routes publish cache read or write prices.
Fresh rows
4/5
Routes crawled in the last 72 hours.
Use case fit
DeepSeek V4 Flash 0731 is often compared for cost-conscious coding and chat workloads. It is worth checking the cheapest input and output platforms separately for batch jobs.
Platform price comparison
| Platform | Ratio | Input | Output | Total |
|---|---|---|---|---|
| AIHubMix | 0.02x | 0.14 USD/M tokens | 0.28 USD/M tokens | 0.42 USD/M tokens |
| Atlas Cloud | 0.08x | 0.44 USD/M tokens | 1.32 USD/M tokens | 1.76 USD/M tokens |
| ClaudeCN | 0.48x | 2.70 USD/M tokens | 8.10 USD/M tokens | 10.80 USD/M tokens |
| Yunwu | 1.61x | 9.00 USD/M tokens | 27.00 USD/M tokens | 36.00 USD/M tokens |
| ZetaTechs | 5.25x | 58.61 USD/M tokens | 58.61 USD/M tokens | 117.22 USD/M tokens |
FAQ
Where is DeepSeek V4 Flash 0731 cheapest?
AIHubMix has the lowest listed combined price at 0.42 USD/M tokens per 1M tokens.
What do input and output prices mean?
Input is charged for prompt tokens you send. Output is charged for tokens generated by the model.
How often are prices refreshed?
Prices are crawler-backed and this page was last refreshed Sep 6, 2026. Always confirm before large purchases.