Google logo

Gemini 3.5 Flash Lite

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Input Price$0.30/1M tokens
Output Price$2.50/1M tokens
Intelligence27.6
Coding49.3

Specifications

Technical details and pricing.

ProviderGoogle
Context Window1,048,576 tokens
Release DateJul 21, 2026
ModalitiesText, Image → Text
CapabilitiesFunction Calling, Vision
AvailabilityChat only

Inference API

Power your AI projects with open-source models.

OpenAI-compatible API. Processing terms vary by selected model; review the current Privacy Policy and DPA before using sensitive data.

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.3 Flash

$0.20 / $0.60

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4 Pro

$1.80 / $3.60

per M tokens

Frequently Asked Questions

What is Gemini 3.5 Flash Lite good for?

Use Gemini 3.5 Flash Lite for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Gemini 3.5 Flash Lite cost?

Pricing is based on usage. Current rates are $0.30/1M tokens for input and $2.50/1M tokens for output.

Can I use Gemini 3.5 Flash Lite in LLMBase Chat?

Yes. Use the chat action on this page to open Gemini 3.5 Flash Lite in LLMBase Chat. Access still follows the current plan and model tier.

Does Gemini 3.5 Flash Lite support images or audio?

Gemini 3.5 Flash Lite can understand images.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.