fireworks/deepseek-v4-flash

Common Name: DeepSeek V4 Flash

Fireworks
SupportedTool InvocationSupportedReasoning
CompareTry in Chat

DeepSeek's cost-efficient hybrid-thinking model in the V4 family.

Specifications

Context
1000K
Maximum Output
384K
Inputtext
Outputtext

Performance (7-day Average)

Collecting…
Collecting…
Collecting…

Pricing

Input$0.14/MTokens
Cached Input$0.028/MTokens
Output$0.28/MTokens

Availability Trend (24h)

Performance Metrics (24h)

Similar Models

$0.14/$0.28/M
ctx1.0Mmax131Kavailtps
InOutCap

DeepSeek's official V4 Flash release for cost-efficient reasoning and agentic workloads.

$0.15/$0.60/M
ctx131Kmax33Kavailtps
InOutCap

OpenAI's open-weight 120B model for production, agentic tasks, and high-reasoning use cases.

$0.07/$0.30/M
ctx131Kmax33Kavailtps
InOutCap

OpenAI's open-weight 20B model for lower-latency reasoning and specialized use cases.

$3.00/$15.00/M
ctx1.0Mmax131Kavailtps
InOutCap

Kimi's flagship multimodal reasoning model for long-context knowledge work and agentic workflows.