Catalog capabilities
- Vision input
- Tool calling
- Streaming
- Structured output
- Reasoning
RouterLab publishes this DeepSeek model under the model ID deepseek-v4.1-flash. The facts below come from the active RouterLab catalog.
Model ID
deepseek-v4.1-flash
Context window
1,000,000 tokens
Maximum output
384,000 tokens
Input / 1M tokens
$0.25
Output / 1M tokens
$1.15
Cache read / 1M tokens
$0.006
Cache write / 1M tokens
$0.03
Cache read reuses tokens already stored in the provider cache. Cache write creates or stores new cache tokens.
curl https://api.routerlab.ch/v1/chat/completions \
-H "Authorization: Bearer $ROUTERLAB_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"deepseek-v4.1-flash","messages":[{"role":"user","content":"Hello RouterLab"}]}'Related models