Qwen logo

Qwen3.8 2.4T A95B

95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Input Price

$2.00/1M tokens

Output Price

$6.00/1M tokens

Intelligence

39.9

Coding

71.9

Specifications

Technical details and pricing.

ProviderQwen
Context Window1,048,576 tokens
Release DateAug 12, 2026
ModalitiesText
CapabilitiesFunction Calling
AvailabilityChat only

EU-hosted inference API

Power your AI projects with open-source models.

Drop-in OpenAI-compatible API. No data leaves Europe.

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.3 Flash

$0.20 / $0.60

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4.1 Flash

$0.40 / $1.40

per M tokens

Frequently Asked Questions

What is Qwen3.8 2.4T A95B good for?

Use Qwen3.8 2.4T A95B for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Qwen3.8 2.4T A95B cost?

Pricing is based on usage. Current rates are $2.00/1M tokens for input and $6.00/1M tokens for output.

Can I use Qwen3.8 2.4T A95B in LLMBase Chat?

Yes. Use the chat action on this page to open Qwen3.8 2.4T A95B in LLMBase Chat. Access still follows the current plan and model tier.

Does Qwen3.8 2.4T A95B support images or audio?

Qwen3.8 2.4T A95B focuses on text-based tasks.

Where does Qwen3.8 2.4T A95B run?

Processing terms depend on the selected model. Qwen3.8 2.4T A95B is not marked as EU-hosted, so it is processed by the company that operates it. LLMBase shows this for every model so you can choose per use case.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.