All models
deepseek

DeepSeek: DeepSeek V4 Flash

Vendor: deepseek|deepseek-v4-flash
Playground
Modalities
TextText
In / out price97% off97% cheaper than the provider's official price per 1M tokens.
$0.0075 / $0.0075 per 1M
Context
1M
Released
Apr 23, 2026

Overview

DeepSeek's cost-efficient MoE model: 284B parameters, 13B active, hybrid attention for cheap long-context work up to 1M tokens. Supports high and xhigh reasoning effort — a fit for fast assistants and high-volume agents.

Context
1M
Max output
Released
Apr 23, 2026
Step-by-step reasoning
Yes
Cheap

Price per 1M tokens

Input and output cost the same, billed per actual usage — no subscriptions, no minimums.

Input
$0.0075/M
$0.220/M97% off97% cheaper than the provider's official price per 1M tokens.
Output
$0.0075/M
$0.660/M99% off99% cheaper than the provider's official price per 1M tokens.

Official list price: $0.220 / $0.660 per 1M

Performance

Gateway metrics for the last 24 hours.

Loading metrics…

FAQ

Are there any rate limits?

The default limit is request rate per key. Each key can be restricted by models, spend and IP in your dashboard.

How is billing calculated?

Pay as you go: one price per 1M tokens, input and output cost the same, charged proportionally to real usage.

Is streaming supported?

Yes, stream: true behaves exactly like in the OpenAI API — tokens arrive as they are generated.