Deepseek logo

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

Input Price

$0.40/1M tokens

Output Price

$1.40/1M tokens

Intelligence

39.5

Coding

N/A

Specifications

Technical details and pricing.

ProviderDeepseek
Context Window1,040,000 tokens
Release DateSep 10, 2026
ModalitiesText, Image → Text
CapabilitiesFunction Calling, Structured Outputs, JSON Mode, Vision
AvailabilityChat, Inference API

EU-hosted inference API

Power your AI projects with open-source models.

Drop-in OpenAI-compatible API. No data leaves Europe.

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.3 Flash

$0.20 / $0.60

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4.1 Flash

$0.40 / $1.40

per M tokens

Frequently Asked Questions

What is DeepSeek V4.1 Flash good for?

Use DeepSeek V4.1 Flash for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does DeepSeek V4.1 Flash cost?

Pricing is based on usage. Current rates are $0.40/1M tokens for input and $1.40/1M tokens for output.

Can I use DeepSeek V4.1 Flash in LLMBase Chat?

Yes. Use the chat action on this page to open DeepSeek V4.1 Flash in LLMBase Chat. Access still follows the current plan and model tier.

Does DeepSeek V4.1 Flash support images or audio?

DeepSeek V4.1 Flash can understand images.

Where does DeepSeek V4.1 Flash run?

Processing terms depend on the selected model. DeepSeek V4.1 Flash is not marked as EU-hosted, so it is processed by the company that operates it. LLMBase shows this for every model so you can choose per use case.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.