deepseek-v4.1-flash · DeepSeek
DeepSeek-V4.1-Flash model official release. This is the smallest model in DeepSeek’s new model-architecture series, featuring native multimodal visual understanding capabilities. The new architecture is designed to deliver a higher capability ceiling, faster inference speeds, greater throughput, and scalability to larger-parameter models.
DeepSeek-V4.1-Flash model official release. This is the smallest model in DeepSeek’s new model-architecture series, featuring native multimodal visual understanding capabilities. The new architecture is designed to deliver a higher capability ceiling, faster inference speeds, greater throughput, and scalability to larger-parameter models.
DeepSeek V4.1 Flash has a 1,000,000 token context window.
On AIHubMix, DeepSeek V4.1 Flash costs $0.1408 per million input tokens and $0.5632 per million output tokens. Cached input reads are billed at $0.0028 per million tokens.
DeepSeek V4.1 Flash accepts text and image input.
DeepSeek V4.1 Flash supports tool calling, function calling, structured outputs and thinking. Per-protocol parameter support is listed in the capability table on this page.
DeepSeek V4.1 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-v4.1-flash — no other code changes needed.
DeepSeek V4.1 Flash is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek V4.1 Flash was released on September 8, 2026 by DeepSeek.
DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model…
DeepSeek’s officially released new multimodal visual-understanding model…
DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent…
DeepSeek V4 Flash 0731 Fast is a high-speed deployment of DeepSeek’s agentic model…
(This model currently points to the older 0423 version; if you need to request the latest…
(This model currently points to the older 0423 version; if you need to request the latest…
Use DeepSeek V4.1 Flash via the AIHubMix unified API — one interface for every major LLM.