Thinkingmachines logo

Thinking Machines: Inkling – Benchmarks, Pricing & Intelligence Analysis

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

Not available on LLMBase

Thinking Machines: Inkling is not currently available in LLMBase Chat or the LLMBase Inference API. Benchmarks, pricing, and model details remain available for research and comparison.

Input Price$1.00/1M tokens
Output Price$4.05/1M tokens
Intelligence32.2
Coding52.1

Specifications

Technical details and pricing.

ProviderThinkingmachines
Context Window524,288 tokens
Release DateJul 15, 2026
ModalitiesText
CapabilitiesFunction Calling
AvailabilityNot available

Inference API

Power your AI projects with open-source models.

OpenAI-compatible API. Processing terms vary by selected model; review the current Privacy Policy and DPA before using sensitive data.

MiniMax

MiniMax M3

$0.40 / $1.40

per M tokens

Z.ai

GLM 5.3 Flash

$0.20 / $0.60

per M tokens

MoonshotAI

Kimi K3

$4.00 / $18.00

per M tokens

DeepSeek

DeepSeek V4 Pro

$1.80 / $3.60

per M tokens

Frequently Asked Questions

What is Thinking Machines: Inkling good for?

Use Thinking Machines: Inkling for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.

How much does Thinking Machines: Inkling cost?

Pricing is based on usage. Current rates are $1.00/1M tokens for input and $4.05/1M tokens for output.

Is Thinking Machines: Inkling available in LLMBase Chat yet?

Not yet on this page. We will add chat and inference actions automatically once an exact catalog entry exists for Thinking Machines: Inkling.

Does Thinking Machines: Inkling support images or audio?

Thinking Machines: Inkling focuses on text-based tasks.

Benchmarks and pricing use Artificial Analysis where available. Catalog specs are used as a fallback.