New Model | DeepSeek V4.1 Flash
DeepSeek V4.1 Flash adds native image understanding for coding, document and screenshot workflows. See the model page for current pricing before rollout.
DeepSeek V4.1 Flash
- Model ID:
deepseek-v4.1-flashon EvoLink; the page URL is/deepseek-v4-1-flash - Context window: 1,000,000 tokens, with a maximum output of 384,000 tokens; choose an output budget for the task
- Billing: input, cache hit and output have separate rate entries. Check the model pricing section and final account charges; per-token rates alone do not determine task cost
- Image input: PNG, JPEG, WebP and GIF. Measure representative image usage before scaling
- Protocols: Chat Completions, Messages and Responses take the same model ID; image fields differ:
image_url, animageblock andinput_image - Thinking and caching: Thinking is enabled by default and prefix caching is automatic; record the thinking setting you send, check the protocol documentation, and read cached tokens from usage
- Legacy routes: on EvoLink,
deepseek-v4-flashanddeepseek-v4-proare not affected and continue to serve V4 Flash and V4 Pro;deepseek-v4-flash-vision-expnow redirects to V4.1 Flash. See the migration guide