DeepSeek: DeepSeek V4 Flash 0731 API Pricing
Live token cost for DeepSeek: DeepSeek V4 Flash 0731 from DeepSeek. Use the figures below for budgeting, then tune your exact request mix in the interactive calculator. Prices refresh every 24 hours from OpenRouter.
Capabilities
DeepSeek: DeepSeek V4 Flash 0731 cost at scale
Estimated monthly cost across common production volumes. Assumes 30-day months and the request shapes shown.
| Tier | Requests / day | In / out tokens | $ / month |
|---|---|---|---|
| Hobby | 1,000 | 500 / 200 | $3.78 |
| Startup | 10,000 | 1,500 / 500 | $105.00 |
| Growth | 100,000 | 3,000 / 800 | $1,932.00 |
| Enterprise | 1,000,000 | 8,000 / 2,000 | $50,400.00 |
Compare DeepSeek: DeepSeek V4 Flash 0731 vs.
Frequently asked questions
How much does DeepSeek: DeepSeek V4 Flash 0731 cost?
DeepSeek: DeepSeek V4 Flash 0731 costs $0.14 per 1M input tokens and $0.28 per 1M output tokens, with cached input at $0.00 per 1M tokens. A typical 1,500-token in / 500-token out request costs $0.00035.
Does DeepSeek: DeepSeek V4 Flash 0731 support cached input?
Yes. DeepSeek: DeepSeek V4 Flash 0731 supports prompt caching at $0.00 per 1M cached input tokens. Reuse the same system prompt or context across requests to cut input cost dramatically.
What is the DeepSeek: DeepSeek V4 Flash 0731 context window?
DeepSeek: DeepSeek V4 Flash 0731 supports a context window of 1,048,576 tokens (1049K). Max output per response is 384,000 tokens.
What is DeepSeek: DeepSeek V4 Flash 0731 good for?
DeepSeek: DeepSeek V4 Flash 0731 is a good fit for general-purpose LLM tasks via the DeepSeek API — chat, code, writing, summarization. For other use cases, run your specific input/output mix through the interactive calculator to compare against alternative models.