Ling 3.1 Flash – Benchmarks, Pricing & Intelligence Analysis
Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
Not available on LLMBase
Ling 3.1 Flash is not currently available in LLMBase Chat or the LLMBase Inference API. Benchmarks, pricing, and model details remain available for research and comparison.
Input Price
$0.07/1M tokens
Output Price
$0.22/1M tokens
Intelligence
20.1
Coding
50.6
Specifications
Technical details and pricing.
EU-hosted inference API
Power your AI projects with open-source models.
Drop-in OpenAI-compatible API. No data leaves Europe.
MiniMax
MiniMax M3
$0.40 / $1.40
per M tokens
Z.ai
GLM 5.3 Flash
$0.20 / $0.60
per M tokens
MoonshotAI
Kimi K3
$4.00 / $18.00
per M tokens
DeepSeek
DeepSeek V4.1 Flash
$0.40 / $1.40
per M tokens
Frequently Asked Questions
What is Ling 3.1 Flash good for?
Use Ling 3.1 Flash for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Ling 3.1 Flash cost?
Pricing is based on usage. Current rates are $0.07/1M tokens for input and $0.22/1M tokens for output.
Is Ling 3.1 Flash available in LLMBase Chat yet?
Not yet on this page. We will add chat and inference actions automatically once an exact catalog entry exists for Ling 3.1 Flash.
Does Ling 3.1 Flash support images or audio?
Ling 3.1 Flash focuses on text-based tasks.
Similar Models
Other models you might want to explore.
Ling 3.0 Flash VL
Inclusionai
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
DetailsLing 3.0 Flash Fin
Inclusionai
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
DetailsDeepSeek V4.1 Flash
Deepseek
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
DetailsBenchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.