Qwen3.8 Max Prime
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
Input Price
$2.00/1M tokens
Output Price
$6.00/1M tokens
Intelligence
45.4
Coding
76.2
Specifications
Technical details and pricing.
EU-hosted inference API
Power your AI projects with open-source models.
Drop-in OpenAI-compatible API. No data leaves Europe.
MiniMax
MiniMax M3
$0.40 / $1.40
per M tokens
Z.ai
GLM 5.3 Flash
$0.20 / $0.60
per M tokens
MoonshotAI
Kimi K3
$4.00 / $18.00
per M tokens
DeepSeek
DeepSeek V4.1 Flash
$0.40 / $1.40
per M tokens
Frequently Asked Questions
What is Qwen3.8 Max Prime good for?
Use Qwen3.8 Max Prime for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Qwen3.8 Max Prime cost?
Pricing is based on usage. Current rates are $2.00/1M tokens for input and $6.00/1M tokens for output.
Can I use Qwen3.8 Max Prime in LLMBase Chat?
Yes. Use the chat action on this page to open Qwen3.8 Max Prime in LLMBase Chat. Access still follows the current plan and model tier.
Does Qwen3.8 Max Prime support images or audio?
Qwen3.8 Max Prime can understand images.
Where does Qwen3.8 Max Prime run?
Processing terms depend on the selected model. Qwen3.8 Max Prime is not marked as EU-hosted, so it is processed by the company that operates it. LLMBase shows this for every model so you can choose per use case.
Similar Models
Other models you might want to explore.
Qwen3.8 Max (0902)
Qwen
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
DetailsQwen3.8 Omni Flash
Qwen
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
DetailsQwen3.8 Flash
Qwen
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
DetailsBenchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.