Qwen3.8 2.4T A95B
95BQwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Input Price
$2.00/1M tokens
Output Price
$6.00/1M tokens
Intelligence
39.9
Coding
71.9
Specifications
Technical details and pricing.
EU-hosted inference API
Power your AI projects with open-source models.
Drop-in OpenAI-compatible API. No data leaves Europe.
MiniMax
MiniMax M3
$0.40 / $1.40
per M tokens
Z.ai
GLM 5.3 Flash
$0.20 / $0.60
per M tokens
MoonshotAI
Kimi K3
$4.00 / $18.00
per M tokens
DeepSeek
DeepSeek V4.1 Flash
$0.40 / $1.40
per M tokens
Frequently Asked Questions
What is Qwen3.8 2.4T A95B good for?
Use Qwen3.8 2.4T A95B for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Qwen3.8 2.4T A95B cost?
Pricing is based on usage. Current rates are $2.00/1M tokens for input and $6.00/1M tokens for output.
Can I use Qwen3.8 2.4T A95B in LLMBase Chat?
Yes. Use the chat action on this page to open Qwen3.8 2.4T A95B in LLMBase Chat. Access still follows the current plan and model tier.
Does Qwen3.8 2.4T A95B support images or audio?
Qwen3.8 2.4T A95B focuses on text-based tasks.
Where does Qwen3.8 2.4T A95B run?
Processing terms depend on the selected model. Qwen3.8 2.4T A95B is not marked as EU-hosted, so it is processed by the company that operates it. LLMBase shows this for every model so you can choose per use case.
Similar Models
Other models you might want to explore.
Qwen3.8 Max Prime
Qwen
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
DetailsQwen3.8 Omni Flash
Qwen
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
DetailsQwen3.8 Max (0902)
Qwen
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
DetailsBenchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.