Nvidia logo

Nemotron 3.5 Lightning – Benchmarks, Pricing & Intelligence Analysis

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Not available on LLMBase

Nemotron 3.5 Lightning is not currently available in LLMBase Chat or the LLMBase Inference API. Benchmarks, pricing, and model details remain available for research and comparison.

Input PriceFree
Output PriceFree
Intelligence23.6
Coding26.8

Specifications

Technical details and pricing.

ProviderNvidia
Context Window28,672 tokens
Release DateAug 11, 2026
ModalitiesText
AvailabilityNot available

Benchmarks

4 benchmark scores from Artificial Analysis.

GPQA74.3%
HLE10.6%
SciCode31.6%
LCR55.3%

Composite Indices

Higher is better; speed and price are normalized

Standard Benchmarks

Only benchmarks with data are shown

EU-Hosted Inference API

Power your AI projects with the best open-source models.

Drop-in OpenAI-compatible API. No data leaves Europe.

Explore Inference API

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.2

$1.40 / $4.40

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4 Pro

$1.80 / $3.60

per M tokens

Frequently Asked Questions

What is Nemotron 3.5 Lightning good for?

Use Nemotron 3.5 Lightning for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Nemotron 3.5 Lightning cost?

Pricing is based on usage. Current rates are Free for input and Free for output.

Is Nemotron 3.5 Lightning available in LLMBase Chat yet?

Not yet on this page. We will add chat and inference actions automatically once an exact catalog entry exists for Nemotron 3.5 Lightning.

Does Nemotron 3.5 Lightning support images or audio?

Nemotron 3.5 Lightning focuses on text-based tasks.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.