Gemma 4 26B A4B
26BGemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Specifications
Technical details and pricing.
Inference API
Power your AI projects with open-source models.
OpenAI-compatible API. Processing terms vary by selected model; review the current Privacy Policy and DPA before using sensitive data.
MiniMax
MiniMax M3
$0.40 / $1.40
per M tokens
Z.ai
GLM 5.3 Flash
$0.20 / $0.60
per M tokens
MoonshotAI
Kimi K3
$4.00 / $18.00
per M tokens
DeepSeek
DeepSeek V4 Pro
$1.80 / $3.60
per M tokens
Frequently Asked Questions
What is Gemma 4 26B A4B good for?
Use Gemma 4 26B A4B for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Gemma 4 26B A4B cost?
Pricing is based on usage. Current rates are $0.12/1M tokens for input and $0.37/1M tokens for output.
Can I use Gemma 4 26B A4B in LLMBase Chat?
Yes. Use the chat action on this page to open Gemma 4 26B A4B in LLMBase Chat. Access still follows the current plan and model tier.
Does Gemma 4 26B A4B support images or audio?
Gemma 4 26B A4B can understand images.
Similar Models
Other models you might want to explore.
Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.