Qwen logo

Qwen3.8 2.4T A95B

95B

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Input Price$2.00/1M tokens
Output Price$6.00/1M tokens
Intelligence58.1
Coding71.8

Specifications

Technical details and pricing.

ProviderQwen
Context Window262,144 tokens
Release DateAug 3, 2026
ModalitiesText
CapabilitiesFunction Calling
AvailabilityChat only

Benchmarks

4 benchmark scores from Artificial Analysis.

GPQA92.7%
HLE43.0%
SciCode52.9%
LCR74.3%

Composite Indices

Higher is better; speed and price are normalized

Standard Benchmarks

Only benchmarks with data are shown

EU-Hosted Inference API

Power your AI projects with the best open-source models.

Drop-in OpenAI-compatible API. No data leaves Europe.

Explore Inference API

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.2

$1.40 / $4.40

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4 Pro

$1.80 / $3.60

per M tokens

Frequently Asked Questions

What is Qwen3.8 2.4T A95B good for?

Use Qwen3.8 2.4T A95B for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Qwen3.8 2.4T A95B cost?

Pricing is based on usage. Current rates are $2.00/1M tokens for input and $6.00/1M tokens for output.

Can I use Qwen3.8 2.4T A95B in LLMBase Chat?

Yes. Use the chat action on this page to open Qwen3.8 2.4T A95B in LLMBase Chat. Access still follows the current plan and model tier.

Does Qwen3.8 2.4T A95B support images or audio?

Qwen3.8 2.4T A95B focuses on text-based tasks.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.