Google logo

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Input Price$0.30/1M tokens
Output Price$2.50/1M tokens
Intelligence36.5
Coding49.3

Specifications

Technical details and pricing.

ProviderGoogle
Context Window1,048,576 tokens
Release DateJul 21, 2026
ModalitiesText, Image โ†’ Text
CapabilitiesFunction Calling, Vision
AvailabilityChat only

Benchmarks

4 benchmark scores from Artificial Analysis.

GPQA83.8%
HLE17.5%
SciCode40.9%
LCR62.0%

Composite Indices

Higher is better; speed and price are normalized

Standard Benchmarks

Only benchmarks with data are shown

EU-Hosted Inference API

Power your AI projects with the best open-source models.

Drop-in OpenAI-compatible API. No data leaves Europe.

Explore Inference API

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.2

$1.40 / $4.40

per M tokens

MoonshotAI

Kimi K2.7 Code

$0.95 / $4.00

per M tokens

DeepSeek

DeepSeek V4 Pro

$1.80 / $3.60

per M tokens

Frequently Asked Questions

What is Gemini 3.5 Flash-Lite good for?

Use Gemini 3.5 Flash-Lite for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Gemini 3.5 Flash-Lite cost?

Pricing is based on usage. Current rates are $0.30/1M tokens for input and $2.50/1M tokens for output.

Can I use Gemini 3.5 Flash-Lite in LLMBase Chat?

Yes. Use the chat action on this page to open Gemini 3.5 Flash-Lite in LLMBase Chat. Access still follows the current plan and model tier.

Does Gemini 3.5 Flash-Lite support images or audio?

Gemini 3.5 Flash-Lite can understand images.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.