Ling-3.0-flash – Benchmarks, Pricing & Intelligence Analysis
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Not available on LLMBase
Ling-3.0-flash is not currently available in LLMBase Chat or the LLMBase Inference API. Benchmarks, pricing, and model details remain available for research and comparison.
Specifications
Technical details and pricing.
Benchmarks
4 benchmark scores from Artificial Analysis.
Composite Indices
Higher is better; speed and price are normalized
Standard Benchmarks
Only benchmarks with data are shown
EU-Hosted Inference API
Power your AI projects with the best open-source models.
Drop-in OpenAI-compatible API. No data leaves Europe.
Explore Inference APIMiniMax
MiniMax M3
$0.40 / $1.40
per M tokens
Z.ai
GLM 5.2
$1.40 / $4.40
per M tokens
MoonshotAI
Kimi K3
$4.00 / $18.00
per M tokens
DeepSeek
DeepSeek V4 Pro
$1.80 / $3.60
per M tokens
Frequently Asked Questions
What is Ling-3.0-flash good for?
Use Ling-3.0-flash for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Ling-3.0-flash cost?
Pricing is based on usage. Current rates are $0.07/1M tokens for input and $0.22/1M tokens for output.
Is Ling-3.0-flash available in LLMBase Chat yet?
Not yet on this page. We will add chat and inference actions automatically once an exact catalog entry exists for Ling-3.0-flash.
Does Ling-3.0-flash support images or audio?
Ling-3.0-flash focuses on text-based tasks.
Similar Models
Other models you might want to explore.
Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.