deepseek-v4.1-flash
Text Generationby DeepSeek
DeepSeek V4.1 Flash is an open-weight sparse mixture-of-experts model and the first built on DeepSeek's Causal Encoder-Decoder (CED) architecture, activating 8B parameters on input and 16B on output. It keeps the one-million-token context window of the V4 Flash line, which suits large codebases, long documents, extended conversations and multi-step agent tasks, and supports function calling alongside both fast non-thinking replies and higher-effort reasoning.
Specifications
Parameters
Not disclosed
Context length
1,048,576 tokens
Task
Text Generation
Provider
DeepSeek
Billing
Per token
API
OpenAI-compatible
Pricing
Input
$0.30
per 1M tokens
Output
$1.20
per 1M tokens