Model information
Released
2026-07-31
Context window
1M
Max output
66K
Model parameters
304B
Weights
Open
Modalities
Input
Text
Output
Text
Identity
- Version
- deepseek-v4-flash-0731
- License
- MIT
Resources
Capabilities
Open weightReasoningStructuredTemperatureTool call
Providers
| Provider | Provider model ID | Protocol | Context | Max output | Input | Output | Capabilities | Notes | Pricing | Streaming |
|---|---|---|---|---|---|---|---|---|---|---|
| OpenRouter | deepseek/deepseek-v4-flash-0731 | openai_chat | 1M | 66K | ReasoningTool callStructuredTemperature | Reasoning levels: max, high, low, Default reasoning: high, Hugging Face: deepseek-ai/DeepSeek-V4-Flash-0731 | input $0.09/M · output $0.18/M cache read $0.018/M | Yes | ||
| NanoGPT | deepseek/deepseek-v4-flash-0731 | openai_chat | 1M | 66K | ReasoningTool callStructuredTemperature | - | input $0.14/M · output $0.28/M cache read $0.014/M | Yes | ||
| TensorX | deepseek/deepseek-v4-flash-0731 | openai_chat | 1M | 66K | ReasoningTool callStructuredTemperature | - | input $0.25/M · output $0.3/M cache read $0.06/M | Yes | ||
| Kilo Gateway | deepseek/deepseek-v4-flash-0731 | openai_chat | 1M | 66K | ReasoningTool callStructuredTemperature | - | input $0.14/M · output $0.28/M cache read $0.028/M | Yes | ||
| Vercel AI Gateway | deepseek/deepseek-v4-flash-0731 | openai_chat | 1M | 66K | ReasoningTool callStructuredTemperature | - | input $0.13/M · output $0.26/M cache read $0.028/M | Yes | ||
| Ambient | deepseek/deepseek-v4-flash-0731 | openai_chat | 1M | 66K | ReasoningTool callStructuredTemperature | - | input $0.14/M · output $0.28/M cache read $0.028/M · cache write $0/M | Yes | ||
| DeepSeek | deepseek-v4-flash-0731 | custom | 1M | 66K | - | - | - | Yes |