Back to the AI model catalog
DeepSeekAvailableHigh availability

DeepSeek V4.1 Flash API

RouterLab publishes this DeepSeek model under the model ID deepseek-v4.1-flash. The facts below come from the active RouterLab catalog.

Model ID

deepseek-v4.1-flash

OpenAI-compatible API

Context window

1,000,000 tokens

Maximum output

384,000 tokens

Input / 1M tokens

$0.25

Output / 1M tokens

$1.15

Cache read / 1M tokens

$0.006

Cache write / 1M tokens

$0.03

Cache read reuses tokens already stored in the provider cache. Cache write creates or stores new cache tokens.

Catalog capabilities

  • Vision input
  • Tool calling
  • Streaming
  • Structured output
  • Reasoning

Minimal API request

curl https://api.routerlab.ch/v1/chat/completions \
  -H "Authorization: Bearer $ROUTERLAB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"deepseek-v4.1-flash","messages":[{"role":"user","content":"Hello RouterLab"}]}'