Provider logo

Inkling Thinking

thinkingmachines/inkling:thinking
Provider logo

Inkling Thinking

thinkingmachines/inkling:thinking

The thinking version of Thinking Machines' 975B-parameter open-weights Mixture-of-Experts generalist with 41B active parameters. It reasons natively over text, images, and audio, and is built for agentic coding, tool use, detailed instruction following, long-context work, and controllable thinking effort.

Added Jul 15, 2026

Model weights

Context Window

1.0M

Max Output

32.8K

Input Price (Auto)

$1.00/1M

Output Price (Auto)

$4.05/1M

Cache Read (Auto)

$0.17/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

42.3

Better than 88% of models compared

Coding Index

52.1

Better than 67% of models compared

Agentic Index

34.1

Better than 76% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

87.2%

Better than 86% of models compared

HLE

Humanity's Last Exam

31.9%

Better than 87% of models compared

AA-LCR

Long context reasoning evaluation

73.3%

Better than 87% of models compared

GDPval-AA

Economically valuable tasks

36.8%

CritPt

Research-level physics reasoning

5.4%

Coding

SciCode

Python programming for scientific computing

46.1%

Better than 85% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

41.5%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

67.7%

Last updated Aug 16, 2026

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…

Compare Inkling Thinking with similar models from the same provider or model family.

Inkling

thinkingmachines/inkling

The non-thinking version of Thinking Machines' 975B-parameter open-weights Mixture-of-Experts generalist with 41B active parameters. It gives faster direct answers across text, images, and audio, and is built for agentic coding, tool use, detailed instruction following, and long-context work.

Inkling Small

thinkingmachines/Inkling-Small

The direct-answer version of Thinking Machines' 276B-parameter open-weights multimodal Mixture-of-Experts model with 12B active parameters. It accepts text and images, and is designed for coding, tool use, instruction following, and general conversational work.

Inkling Small Thinking

thinkingmachines/Inkling-Small:thinking

The reasoning version of Thinking Machines' 276B-parameter open-weights multimodal Mixture-of-Experts model with 12B active parameters. It reasons over text and images, and is designed for agentic coding, tool use, instruction following, and long workflows with controllable thinking effort.

Abliterated Model Large V2

abliteration-ai/abliterated-model-large-v2

Abliteration.ai's default unrestricted large text reasoning model is derived from GLM-5.3 and weight-modified to reduce refusals compared with the base model. It is intended for harder reasoning and evaluation workloads, with automatic prompt caching and a one-million-token context window.

Granite 4.2 8B

ibm-granite/granite-4.2-8b

IBM Granite 4.2 8B is an Apache 2.0-licensed dense model with native step-by-step reasoning and specialized training for agentic work. It can plan before acting, sequence tools, navigate codebases, work in terminals, and verify results across coding, search, mathematics, science, and complex instruction-following tasks.

GLM 5.3 TEE

TEE/glm-5.3

GLM-5.3 is Z.AI's open-weight reasoning model for complex software engineering, autonomous agents, vulnerability research, and long-horizon tasks. This text-only deployment runs inside a Phala Trusted Execution Environment with Redpill attestation and signed completion receipts.