Qwen logo

Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

Input Price

$0.10/1M tokens

Output Price

$0.80/1M tokens

Intelligence

12.5

Coding

N/A

Specifications

Technical details and pricing.

ProviderQwen
Context Window1,000,000 tokens
Release DateMar 30, 2026
ModalitiesText, Image → Text
CapabilitiesFunction Calling, Vision
AvailabilityChat only

EU-hosted inference API

Power your AI projects with open-source models.

Drop-in OpenAI-compatible API. No data leaves Europe.

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.3 Flash

$0.20 / $0.60

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4.1 Flash

$0.40 / $1.40

per M tokens

Frequently Asked Questions

What is Qwen3.8 Omni Flash good for?

Use Qwen3.8 Omni Flash for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Qwen3.8 Omni Flash cost?

Pricing is based on usage. Current rates are $0.10/1M tokens for input and $0.80/1M tokens for output.

Can I use Qwen3.8 Omni Flash in LLMBase Chat?

Yes. Use the chat action on this page to open Qwen3.8 Omni Flash in LLMBase Chat. Access still follows the current plan and model tier.

Does Qwen3.8 Omni Flash support images or audio?

Qwen3.8 Omni Flash can understand images.

Where does Qwen3.8 Omni Flash run?

Processing terms depend on the selected model. Qwen3.8 Omni Flash is not marked as EU-hosted, so it is processed by the company that operates it. LLMBase shows this for every model so you can choose per use case.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.