Catalog capabilities
- Tool calling
- Streaming
- Structured output
- Reasoning
RouterLab publishes this DeepSeek model under the model ID deepseek-v4-flash. The facts below come from the active RouterLab catalog.
Model ID
deepseek-v4-flash
Context window
1,000,000 tokens
Maximum output
384,000 tokens
Input / 1M tokens
$0.09
Output / 1M tokens
$0.18
Cache read / 1M tokens
$0.018
Cache write / 1M tokens
$0.09
Cache read reuses tokens already stored in the provider cache. Cache write creates or stores new cache tokens.
curl https://api.routerlab.ch/v1/chat/completions \
-H "Authorization: Bearer $ROUTERLAB_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"deepseek-v4-flash","messages":[{"role":"user","content":"Hello RouterLab"}]}'Related models