ZeroTwo home

DeepSeek V4 Flash

Default
Efficiency-optimized 284B-parameter MoE with 13B active and a 1M-token context window. Supports thinking and non-thinking modes. Served via Fireworks AI.
Efficiency-optimized 284B-parameter MoE with 13B active and a 1M-token context window. Supports thinking and non-thinking modes. Served via Fireworks AI.

DeepSeek V4 Flash is an AI model from DeepSeek, available on ZeroTwo. It costs $0.22 per million input tokens and $0.66 per million output tokens, has a 1,000,000-token context window, can return up to 384,000 output tokens.

Price
$0.22$0.66
Input • Output

Efficiency-optimized 284B-parameter MoE with 13B active and a 1M-token context window. Supports thinking and non-thinking modes. Served via Fireworks AI.

1,000,000 context window
384,000 max output tokens
Reasoning token support
Pricing
Pricing is based on the number of tokens used. For tool-specific models, like search and computer use, there's a fee per tool call. See details in the pricing page.

Rates are Fireworks AI serverless Standard-tier pricing for DeepSeek V4 Flash (0731), not DeepSeek's own API list price. Fireworks Priority tier is $0.275 input and $0.825 output per million tokens; cached input bills at $0.007 per million.

Text tokens
Per 1M tokens
Batch API price
Input
$0.22
Cached input
Output
$0.66
Quick comparison
Input
Cached input
Output
DeepSeek V4 Flash
$0.22

Questions

How much does DeepSeek V4 Flash cost?

DeepSeek V4 Flash is priced at $0.22 per million input tokens and $0.66 per million output tokens. On ZeroTwo it is included in your plan's credits rather than billed per token.

What is DeepSeek V4 Flash's context window?

DeepSeek V4 Flash accepts up to 1,000,000 tokens of context and can return up to 384,000 output tokens.

Can I use DeepSeek V4 Flash for free?

You can try DeepSeek V4 Flash on ZeroTwo's free plan, which includes a monthly credit allowance across every model in the catalogue. Heavier use moves to a paid plan rather than a per-token bill.