deepseek-v4-flashModel origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$0.09
Output / 1M
$0.18
Cache read / 1M
$0.018
Cache write / 1M
$0.09
Model catalog
Explore available models, compare API formats, context windows and pricing. Use model identifiers directly in your integrations.
45
available models
Active public catalog
7
catalog categories
Categories derived from the catalog
3
API formats
Routes published in the catalog
6
embedding models
Also visible in this catalog
Quick access
45 of 45 models
deepseek-v4-flashModel origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$0.09
Output / 1M
$0.18
Cache read / 1M
$0.018
Cache write / 1M
$0.09
deepseek-v4-flash-0731Model origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$0.09
Output / 1M
$0.18
Cache read / 1M
$0.018
Cache write / 1M
$0.09
deepseek-v4-proModel origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$1.25
Output / 1M
$2.55
Cache read / 1M
$0.10
Cache write / 1M
$0.10
deepseek-v4-pro-0813Model origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$1.25
Output / 1M
$2.55
Cache read / 1M
$0.10
Cache write / 1M
$0.10
deepseek-v4.1-flashModel origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$0.25
Output / 1M
$1.15
Cache read / 1M
$0.006
Cache write / 1M
$0.03
glm-5.2Model origin
GLM
Context
1.0M
Format
OpenAI-compatible
Max output
131.1K
Input / 1M
$0.95
Output / 1M
$3.00
Cache read / 1M
$0.18
Cache write / 1M
$0.95
Model origin
GLM
Context
1.3M
Format
OpenAI-compatible
Max output
262.1K
Input / 1M
$1.40
Output / 1M
$4.40
Cache read / 1M
$0.26
Cache write / 1M
$1.40
Model origin
GLM
Context
1.0M
Format
OpenAI-compatible
Max output
131.1K
Input / 1M
$0.15
Output / 1M
$0.50
Cache read / 1M
$0.03
Cache write / 1M
$0.15
Model origin
OpenAI
Context
960K
Format
OpenAI-compatible
Max output
64K
Input / 1M
$0.834
Output / 1M
$2.501
Cache read / 1M
$0.042
Cache write / 1M
Same as input
kimi-k2.7-codeModel origin
Kimi
Context
262.1K
Format
OpenAI-compatible
Max output
32.8K
Input / 1M
$0.70
Output / 1M
$3.45
Cache read / 1M
$0.15
Cache write / 1M
$0.70
Model origin
Kimi
Context
1.0M
Format
OpenAI-compatible
Max output
131.1K
Input / 1M
$3.00
Output / 1M
$15.00
Cache read / 1M
$0.30
Cache write / 1M
$3.00
minimax-m2.7Model origin
MiniMax
Context
204.8K
Format
OpenAI-compatible
Max output
204.8K
Input / 1M
$0.29
Output / 1M
$1.19
Cache read / 1M
$0.29
Cache write / 1M
$0.29
minimax-m3Model origin
MiniMax
Context
204.8K
Format
OpenAI-compatible
Max output
204.8K
Input / 1M
$0.28
Output / 1M
$1.15
Cache read / 1M
$0.06
Cache write / 1M
$0.28
Model origin
Qwen
Context
991.8K
Format
OpenAI-compatible
Max output
131.1K
Input / 1M
$2.00
Output / 1M
$6.00
Cache read / 1M
$0.25
Cache write / 1M
$2.50
Model origin
AWS Claude
Context
200K
Format
Claude Messages
Max output
64K
Input / 1M
$0.60
Output / 1M
$3.00
Cache read / 1M
$0.06
Cache write / 1M
$0.75
Model origin
AWS Claude
Context
200K
Format
Claude Messages
Max output
64K
Input / 1M
$3.00
Output / 1M
$15.00
Cache read / 1M
$0.30
Cache write / 1M
$3.75
Model origin
AWS Claude
Context
200K
Format
Claude Messages
Max output
64K
Input / 1M
$1.20
Output / 1M
$6.00
Cache read / 1M
$0.12
Cache write / 1M
$1.50
Model origin
Open source
Context
8.2K
Format
Embeddings
Max output
—
Input / 1M
$0.01
Output / 1M
$0.00
Cache read / 1M
Not applicable
Cache write / 1M
Not applicable
Model origin
Claude
Context
1M
Format
Claude Messages
Max output
128K
Input / 1M
$10.00
Output / 1M
$50.00
Cache read / 1M
$1.00
Cache write / 1M
$12.50
Model origin
Claude
Context
200K
Format
Claude Messages
Max output
64K
Input / 1M
$1.00
Output / 1M
$5.00
Cache read / 1M
$0.10
Cache write / 1M
$1.25
Model origin
Claude
Context
1M
Format
Claude Messages
Max output
128K
Input / 1M
$5.00
Output / 1M
$25.00
Cache read / 1M
$0.50
Cache write / 1M
$6.25
Model origin
Claude
Context
1M
Format
Claude Messages
Max output
128K
Input / 1M
$2.00
Output / 1M
$10.00
Cache read / 1M
$0.20
Cache write / 1M
$2.50
Model origin
Context
1.0M
Format
OpenAI-compatible
Max output
65.5K
Input / 1M
$4.00
Output / 1M
$18.00
Cache read / 1M
$0.40
Cache write / 1M
$0.40
Model origin
Context
1.0M
Format
OpenAI-compatible
Max output
65.5K
Input / 1M
$1.50
Output / 1M
$9.00
Cache read / 1M
$0.15
Cache write / 1M
$0.15
Model origin
Context
1.0M
Format
OpenAI-compatible
Max output
65.5K
Input / 1M
$0.75
Output / 1M
$3.75
Cache read / 1M
$0.075
Cache write / 1M
$0.075
Model origin
Context
1.0M
Format
OpenAI-compatible
Max output
65.5K
Input / 1M
$0.75
Output / 1M
$3.75
Cache read / 1M
$0.075
Cache write / 1M
$0.075
Model origin
Context
1.0M
Format
OpenAI-compatible
Max output
65.5K
Input / 1M
$0.75
Output / 1M
$3.75
Cache read / 1M
$0.075
Cache write / 1M
$0.075
Model origin
OpenAI
Context
922K
Format
OpenAI-compatible
Max output
128K
Input / 1M
$0.20
Output / 1M
$1.20
Cache read / 1M
$0.02
Cache write / 1M
$0.25
Model origin
OpenAI
Context
922K
Format
OpenAI-compatible
Max output
128K
Input / 1M
$5.00
Output / 1M
$30.00
Cache read / 1M
$0.50
Cache write / 1M
$6.25
Model origin
OpenAI
Context
922K
Format
OpenAI-compatible
Max output
128K
Input / 1M
$2.00
Output / 1M
$12.00
Cache read / 1M
$0.20
Cache write / 1M
$2.50
Model origin
OpenAI
Context
922K
Format
OpenAI-compatible
Max output
128K
Input / 1M
$10.00
Output / 1M
$50.00
Cache read / 1M
$1.00
Cache write / 1M
$12.50
Model origin
xAI (SpaceX)
Context
500K
Format
OpenAI-compatible
Max output
—
Input / 1M
$2.00
Output / 1M
$6.00
Cache read / 1M
$0.50
Cache write / 1M
Same as input
Model origin
Qwen
Context
32.8K
Format
Embeddings
Max output
—
Input / 1M
$0.012
Output / 1M
$0.00
Cache read / 1M
Not applicable
Cache write / 1M
Not applicable
Model origin
Qwen
Context
32.8K
Format
Embeddings
Max output
—
Input / 1M
$0.02
Output / 1M
$0.00
Cache read / 1M
Not applicable
Cache write / 1M
Not applicable
Model origin
Qwen
Context
32.8K
Format
Embeddings
Max output
—
Input / 1M
$0.09
Output / 1M
$0.00
Cache read / 1M
Not applicable
Cache write / 1M
Not applicable
Model origin
OpenAI
Context
8.2K
Format
Embeddings
Max output
—
Input / 1M
$0.13
Output / 1M
$0.00
Cache read / 1M
Not applicable
Cache write / 1M
Not applicable
Model origin
OpenAI
Context
—
Format
OpenAI-compatible
Max output
—
Input / 1M
$0.00
Output / 1M
$0.00
Cache read / 1M
Same as input
Cache write / 1M
Same as input
Model origin
Open source
Context
32.8K
Format
Embeddings
Max output
—
Input / 1M
$0.05
Output / 1M
$0.00
Cache read / 1M
Not applicable
Cache write / 1M
Not applicable
deepseek-v4-flash-0731-trialModel origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
deepseek-v4-pro-0813-trialModel origin
DeepSeek
Context
1M
Format
OpenAI-compatible
Max output
384K
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
gemini-3.7-flash-trialModel origin
Context
1.0M
Format
OpenAI-compatible
Max output
65.5K
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
Model origin
GLM
Context
1.0M
Format
OpenAI-compatible
Max output
131.1K
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
Model origin
xAI (SpaceX)
Context
500K
Format
OpenAI-compatible
Max output
—
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
Model origin
MiniMax
Context
204.8K
Format
OpenAI-compatible
Max output
204.8K
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
Model origin
Qwen
Context
991.8K
Format
OpenAI-compatible
Max output
131.1K
Input / 1M
$0.01
Output / 1M
$0.01
Cache read / 1M
$0.001
Cache write / 1M
$0.001
Need a dedicated embeddings view?
Compare dimensions, context and cost on the specialized page.
Next step
See how to integrate a model into your application in minutes.