Qwen3.8 Flash
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Specifications
Technical details and pricing.
EU-hosted inference API
Power your AI projects with open-source models.
Drop-in OpenAI-compatible API. No data leaves Europe.
MiniMax
MiniMax M3
$0.40 / $1.40
per M tokens
Z.ai
GLM 5.2
$1.40 / $4.40
per M tokens
MoonshotAI
Kimi K3
$4.00 / $18.00
per M tokens
DeepSeek
DeepSeek V4 Pro
$1.80 / $3.60
per M tokens
Frequently Asked Questions
What is Qwen3.8 Flash good for?
Use Qwen3.8 Flash for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Qwen3.8 Flash cost?
Pricing is based on usage. Current rates are $0.15/1M tokens for input and $0.47/1M tokens for output.
Can I use Qwen3.8 Flash in LLMBase Chat?
Yes. Use the chat action on this page to open Qwen3.8 Flash in LLMBase Chat. Access still follows the current plan and model tier.
Does Qwen3.8 Flash support images or audio?
Qwen3.8 Flash can understand images.
Similar Models
Other models you might want to explore.
Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.