infery
← All models

DeepSeek V4.1 Flash

deepseek-v4.1-flash

Chat & Textby DeepSeek

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the cost-efficient tier of the V4.1 family. 1,048,576 token context window, maximum output of 384,000 tokens. Higher uptime with 4 providers.

chatstreamingvisiontoolsjson mode

Details

Accepts
text
Context window
1M tokens

Pricing

Input
18.75 cr / 1M tokens
Output
75 cr / 1M tokens
Cached input
0.375 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).