DeepSeek V4.1 Flash

deepseek-v4.1-flash · DeepSeek

DeepSeek-V4.1-Flash model official release. This is the smallest model in DeepSeek’s new model-architecture series, featuring native multimodal visual understanding capabilities. The new architecture is designed to deliver a higher capability ceiling, faster inference speeds, greater throughput, and scalability to larger-parameter models.

API Pricing

Input$0.1408 / 1M tokens
Output$0.5632 / 1M tokens
Cache read$0.0028 / 1M tokens

Specifications

Context1M tokens
Modalitiestext, image
CapabilitiesThinking, Streaming, Tool calling, Web search, Structured outputs, Prompt caching
Endpointschat_completions, claude_api

Frequently asked questions

What is DeepSeek V4.1 Flash?

DeepSeek-V4.1-Flash model official release. This is the smallest model in DeepSeek’s new model-architecture series, featuring native multimodal visual understanding capabilities. The new architecture is designed to deliver a higher capability ceiling, faster inference speeds, greater throughput, and scalability to larger-parameter models.

What is the context length of DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash has a 1,000,000 token context window.

How much does DeepSeek V4.1 Flash cost?

On AIHubMix, DeepSeek V4.1 Flash costs $0.1408 per million input tokens and $0.5632 per million output tokens. Cached input reads are billed at $0.0028 per million tokens.

What modalities does DeepSeek V4.1 Flash support?

DeepSeek V4.1 Flash accepts text and image input.

What capabilities does DeepSeek V4.1 Flash support?

DeepSeek V4.1 Flash supports tool calling, function calling, structured outputs and thinking. Per-protocol parameter support is listed in the capability table on this page.

How do I call DeepSeek V4.1 Flash via API?

DeepSeek V4.1 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-v4.1-flash — no other code changes needed.

Who created DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was DeepSeek V4.1 Flash released?

DeepSeek V4.1 Flash was released on September 8, 2026 by DeepSeek.

More models from DeepSeek

See all DeepSeek models →

DeepSeek V4 Flash 0731

by DeepSeek

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model…

$0.142/1M in · $0.284/1M out
1,000,000 tokens context

DeepSeek V4 Flash Vision Exp

by DeepSeek

DeepSeek’s officially released new multimodal visual-understanding model…

$0.155/1M in · $0.62/1M out
1,000,000 tokens context

DeepSeek V4 Pro 0813

by DeepSeek

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent…

$0.6918/1M in · $2.0754/1M out
1,000,000 tokens context

DeepSeek V4 Flash 0731 Fast

by DeepSeek

DeepSeek V4 Flash 0731 Fast is a high-speed deployment of DeepSeek’s agentic model…

$0.28/1M in · $1.4/1M out
1,000,000 tokens context

DeepSeek V4 Flash

by DeepSeek

(This model currently points to the older 0423 version; if you need to request the latest…

$0.142/1M in · $0.284/1M out
1,000,000 tokens context

DeepSeek V4 Pro

by DeepSeek

(This model currently points to the older 0423 version; if you need to request the latest…

$1.69/1M in · $3.38/1M out
1,000,000 tokens context

Use DeepSeek V4.1 Flash via the AIHubMix unified API — one interface for every major LLM.