Thinking Machines
·Released on Jul 17, 2026

Inkling API Benchmarks, Pricing & Provider Data

Compare Inkling with another model

Choose a model to open its comparison page.

Share on X
LLM

Inkling benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.00000001/request. The page also shows measured API speed and first-token latency.

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic an...

Quality
#48of 112
60.0
LMSpeed score
Speed
#52of 81
25char/s
0.79 s
Cost
#113of 186
$0.00000001/ 1M · 8:1 in:out
$0.0000000016 in · $0.0000000084 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
7 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents48.6Coding47.9Reasoning55.2Knowledge51.3Math65.4Multilingual-Multimodal45.7Instruction following56.4
#1Math65.4Provisional1/4 Measured dimensions
80% interval: 49.081.7
foundational mathPrior only
competition math65.4
aime · aime2026
proof frontierPrior only
applied tool mathPrior only
#2Instruction following56.4Provisional1/4 Measured dimensions
80% interval: 40.472.4
constraint followingPrior only
structured outputPrior only
novel instruction generalization56.4
ifbench · if_bench
long multiturn instructionPrior only
#3Reasoning55.2Estimated2/4 Measured dimensions
80% interval: 44.466.0
abstract logicPrior only
scientific causal55.9
critpt · critpt / gpqa · gpqa / hle · hle / hle · hle_no_tools
multistep constraintsPrior only
evidence verification54.5
lcr · lcr
#4Knowledge51.3Provisional1/4 Measured dimensions
80% interval: 37.365.3
broad knowledge51.3
aa_omniscience · aa_omniscience_index / benchlm_category_knowledge · benchlm_category_knowledge
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
#5Agents48.6RatedGlobal rank #423/4 Measured dimensions
80% interval: 40.356.8
planningPrior only
tool use50.3
mcp_atlas · mcp_atlas / tau · aa_tau3_banking
environment execution50
browsecomp · browse_comp / enterprise_ops_gym · aa_enterprise_ops_gym / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability45.4
briefcase · aa_briefcase_elo
#6Coding47.9RatedGlobal rank #294/4 Measured dimensions
80% interval: 41.054.7
code generation54
scicode · scicode
repository engineering45.6
swe_pro · swe_pro
debugging testing51.3
swe_verified · swe_verified
tooling quality40.6
terminalbench · aa_terminal_bench21 / terminalbench · benchlm_coding_terminal_bench2
#7Multimodal45.7Estimated2/4 Measured dimensions
80% interval: 33.857.6
perception ocrPrior only
document spatial49.4
charxiv · charxiv
visual reasoning42
mmmu_pro · mmmu_pro
video actionPrior only
No data:Multilingual

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
1.0Mtokens
1.3K pages of text
OUTPUT
471.9Ktokens
8K128K1M4M
1.0M

Features

Technical Details

Input
Output
Released
Jul 2026
Tokenizer
Other
Architecture
text+image+audio->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 12, 2026

Overall

undefined metric} other undefined metrics}}

Overall score60.0#48 / 112
Overall score
LMSpeed rank#48
Score60.0
UpdatedSep 12, 2026
Confidence4

Speed & latency

undefined metric} other undefined metrics}}

Output speed83.3 tok/s#52 / 81Time to first token2.63 s#56 / 81
Output speed
LMSpeed rank#52
Score83.3 tok/s
UpdatedSep 12, 2026
Confidence4
Time to first token
LMSpeed rank#56
Score2.63 s
UpdatedSep 12, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$1.00/M#113 / 186Output price$4.05/M#114 / 186
Input price
LMSpeed rank#113
Score$1.00/M
UpdatedSep 12, 2026
Confidence4
Output price
LMSpeed rank#114
Score$4.05/M
UpdatedSep 12, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score48.6#4280% interval40.3–56.83/4 Measured dimensionsAgentic score57.6#46 / 77Terminal-Bench 2.063.8#27 / 53
Dimensions and evidence
Planning & decompositionPrior only
Tool use50.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mcp_atlas · mcp_atlas · z 0.07 · q 1.00
tau · aa_tau3_banking · z -0.99 · q 1.00
Environment & long-horizon execution50
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -0.38 · q 1.00
enterprise_ops_gym · aa_enterprise_ops_gym · z -0.92 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z 0.00 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z 0.00 · q 1.00
Recovery & completion reliability45.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
briefcase · aa_briefcase_elo · z -1.27 · q 1.00
Agentic score
LMSpeed rank#46
Score57.6
UpdatedSep 12, 2026
Confidence4
Terminal-Bench 2.0V3 evidence
LMSpeed rank#27
Score63.8
UpdatedSep 12, 2026
Confidence4
BrowseCompV3 evidence
LMSpeed rank#20
Score77.1
UpdatedSep 12, 2026
Confidence4
MCP AtlasV3 evidence
LMSpeed rank#15
Score74.1
UpdatedSep 12, 2026
Confidence4
Design Arena Agentic Web Dev
LMSpeed rank#1
Score1257.0
UpdatedSep 12, 2026
Confidence4
AA Agentic Index
LMSpeed rank#38
Score24.3
UpdatedSep 12, 2026
Confidence4
GDPval-AA
LMSpeed rank#35
Score33.3
UpdatedSep 12, 2026
Confidence4
GDPval-AAV3 evidence
LMSpeed rank#36
Score1165.0
UpdatedSep 12, 2026
Confidence4
AA EnterpriseOps-GymV3 evidence
LMSpeed rank#15
Score38.0
UpdatedSep 12, 2026
Confidence4
Terminal-Bench 2.1 (Vals)
LMSpeed rank#38
Score47.6
UpdatedSep 12, 2026
Confidence4
AA-AnalystAgent
LMSpeed rank#13
Score23.8
UpdatedSep 12, 2026
Confidence4
AA BriefcaseV3 evidence
LMSpeed rank#18
Score844.0
UpdatedSep 3, 2026
Confidence3
AA Tau3 BankingV3 evidence
LMSpeed rank#19
Score29.1
UpdatedSep 3, 2026
Confidence3

Coding

V3.0

undefined metric} other undefined metrics}} · Rated

Score47.9#2980% interval41.0–54.74/4 Measured dimensionsSciCode47.0%#41 / 89Coding score49.0#58 / 87
Dimensions and evidence
Code generation54
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.37 · q 1.00
Repository engineering45.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_pro · swe_pro · z -0.49 · q 1.00
Debugging & testing51.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_verified · swe_verified · z 0.16 · q 1.00
Tool-assisted development & quality40.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · aa_terminal_bench21 · z -2.74 · q 1.00
terminalbench · benchlm_coding_terminal_bench2 · z -0.24 · q 1.00
SciCodeV3 evidence
LMSpeed rank#41
Score47.0%
UpdatedSep 12, 2026
Confidence4
Coding score
LMSpeed rank#58
Score49.0
UpdatedSep 12, 2026
Confidence4
SWE-bench VerifiedV3 evidence
LMSpeed rank#23
Score77.6
UpdatedSep 12, 2026
Confidence4
SWE-bench ProV3 evidence
LMSpeed rank#38
Score54.3
UpdatedSep 12, 2026
Confidence4
Terminal-Bench 2.0V3 evidence
LMSpeed rank#23
Score63.8
UpdatedSep 12, 2026
Confidence4
AA Coding Index
LMSpeed rank#44
Score52.1
UpdatedSep 12, 2026
Confidence4
AA-SciCodeV3 evidence
LMSpeed rank#47
Score47.0
UpdatedSep 12, 2026
Confidence4
FrontierSWE v2
LMSpeed rank#11
Score4.1
UpdatedSep 12, 2026
Confidence4
LiveCodeBench (Vals)
LMSpeed rank#22
Score85.5
UpdatedSep 12, 2026
Confidence4
SWE-bench (Vals)
LMSpeed rank#25
Score77.6
UpdatedSep 12, 2026
Confidence4
AA Terminal-Bench 2.1V3 evidence
LMSpeed rank#20
Score55.1
UpdatedSep 3, 2026
Confidence3

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score55.280% interval44.4–66.02/4 Measured dimensionsGPQA87.2%#53 / 218HLE31.9%#52 / 216
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning55.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z 0.22 · q 1.00
gpqa · gpqa · z 0.62 · q 1.00
hle · hle · z 0.64 · q 1.00
hle · hle_no_tools · z -1.50 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification54.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z 0.25 · q 1.00
GPQAV3 evidence
LMSpeed rank#53
Score87.2%
UpdatedSep 12, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#52
Score31.9%
UpdatedSep 12, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#41
Score77.3
UpdatedSep 12, 2026
Confidence4
CritPtV3 evidence
LMSpeed rank#45
Score5.4
UpdatedSep 12, 2026
Confidence4
Reasoning score
LMSpeed rank#26
Score74.1
UpdatedSep 12, 2026
Confidence4

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score51.380% interval37.3–65.31/4 Measured dimensionsKnowledge score65.1#52 / 83GPQA-D87.9#24 / 32
Dimensions and evidence
Broad knowledge51.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z 0.19 · q 1.00
benchlm_category_knowledge · benchlm_category_knowledge · z -0.27 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Knowledge scoreV3 evidence
LMSpeed rank#52
Score65.1
UpdatedSep 12, 2026
Confidence4
GPQA-D
LMSpeed rank#24
Score87.9
UpdatedSep 12, 2026
Confidence4
HLE w/o tools
LMSpeed rank#20
Score30.0
UpdatedSep 12, 2026
Confidence4
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#63
Score25.5
UpdatedSep 12, 2026
Confidence4
AA-GPQA Diamond
LMSpeed rank#49
Score87.2
UpdatedSep 12, 2026
Confidence4
AA-HLE
LMSpeed rank#45
Score31.9
UpdatedSep 12, 2026
Confidence4
AA-Omniscience IndexV3 evidence
LMSpeed rank#37
Score2.0
UpdatedSep 12, 2026
Confidence4
AA-Omniscience Accuracy
LMSpeed rank#30
Score41.6
UpdatedSep 12, 2026
Confidence4
AA-Omniscience Hallucination Rate
LMSpeed rank#58
Score67.7
UpdatedSep 12, 2026
Confidence4
GPQA Diamond (Vals)
LMSpeed rank#27
Score87.1
UpdatedSep 12, 2026
Confidence4
MMLU-Pro (Vals)
LMSpeed rank#28
Score86.3
UpdatedSep 12, 2026
Confidence4
AA Openness Index
LMSpeed rank#4
Score38.9
UpdatedSep 4, 2026
Confidence3

Math

V3.0

undefined metric} other undefined metrics}} · Provisional

Score65.480% interval49.0–81.71/4 Measured dimensionsMath score78.9#10 / 63AIME2697.1#2 / 14
Dimensions and evidence
Foundational mathPrior only
Competition math65.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aime · aime2026 · z 2.05 · q 1.00
Advanced proofsPrior only
Applied & tool-assisted mathPrior only
Math score
LMSpeed rank#10
Score78.9
UpdatedSep 12, 2026
Confidence4
AIME26V3 evidence
LMSpeed rank#2
Score97.1
UpdatedSep 12, 2026
Confidence4

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Estimated

Score45.780% interval33.8–57.62/4 Measured dimensionsMultimodal Grounded score49.2#49 / 56MMMU-Pro73.5#28 / 31
Dimensions and evidence
Perception & OCRPrior only
Document & spatial understanding49.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
charxiv · charxiv · z 0.07 · q 1.00
Visual reasoning42
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
mmmu_pro · mmmu_pro · z -1.13 · q 1.00
Video & grounded actionPrior only
Multimodal Grounded score
LMSpeed rank#49
Score49.2
UpdatedSep 12, 2026
Confidence4
MMMU-ProV3 evidence
LMSpeed rank#28
Score73.5
UpdatedSep 12, 2026
Confidence4
CharXivV3 evidence
LMSpeed rank#14
Score82.0
UpdatedSep 12, 2026
Confidence4
CharXiv w/o tools
LMSpeed rank#8
Score78.1
UpdatedSep 12, 2026
Confidence4
AA-MMMU-Pro
LMSpeed rank#41
Score73.5
UpdatedSep 12, 2026
Confidence4
Design Arena Website
LMSpeed rank#44
Score1226.0
UpdatedSep 12, 2026
Confidence4

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score56.480% interval40.4–72.41/4 Measured dimensionsInstruction Following score85.2#32 / 52IFBench79.8#4 / 14
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization56.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · if_bench · z 0.57 · q 1.00
Multi-turn & long instructionsPrior only
Instruction Following score
LMSpeed rank#32
Score85.2
UpdatedSep 12, 2026
Confidence4
IFBenchV3 evidence
LMSpeed rank#4
Score79.8
UpdatedSep 12, 2026
Confidence4

OpenRouter endpoints

3 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
BaseTen
baseten/fp8
$1/M$4.05/M99.8%undefined tokens / undefined tokens
Together
together
$1/M$4.05/M98.5%undefined tokens / undefined tokens
DeepInfra
deepinfra/fp8
$0.950/M$4.05/M93.3%undefined tokens / undefined tokens

Pricing Comparison

Compare Inkling API pricing across 20 providers. Prices range from $0.00000001/request to $10273.97/M. Future Hub offers the lowest rate at $0.00000001/request.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
thinkingmachines/inkling:free
default
$75.00/M
$75.00/M
L1
99%
thinkingmachines/inkling:free
free
$0.00005/request
-
L1
100%
thinkingmachines/inkling
default
$75.00/M
$75.00/M
L1
99%
thinkingmachines/inkling:free
default
$547.50/M
$547.50/M
L1
70%
thinkingmachines/inkling
default
$75.00/M
$75.00/M
L1
47%
accounts/fireworks/models/inkling
测试专用
$10273.97/M
$10273.97/M
L1
100%
thinkingmachines/inkling:free
default
$75.00/M
$75.00/M
L1
100%
laohuang/thinkingmachines/inkling
default
$0.020/request
-
L1
0%
inkling
openrouter
$0.00000001/request
-
L1
0%
L2
100%
inkling
price
-100%$0.0010/M
Cache read$0.0002/MCache write$0.0010/MCache write 1h$0.0016/M
-100%$0.0040/M
L1
0%
inkling
无限制
$7.30/M
Cache read$1.24/M
$29.57/M
L1
0%
inkling
model
$7.30/M
Cache read$1.24/M
$29.57/M

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Inkling include?
LMSpeed shows Inkling benchmark context, API price, output speed, first-token latency, and provider data across 20 providers when those signals are available.
What is the Inkling API price?
Inkling has pricing from undefined provider} other undefined providers}}, ranging from $0.00000001/request to $10273.97/M. Future Hub has the lowest listed price.
What does the Inkling API pricing table include?
The Inkling API pricing table compares 20 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Inkling API pricing?
Future Hub currently has the lowest listed Inkling price at $0.00000001/request across undefined provider} other undefined providers}}.
Can I compare Inkling API price and speed together?
Yes. LMSpeed shows Inkling API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Inkling API free?
Inkling does not currently have a free API tier on LMSpeed. All 20 providers charge per token.

Also known as

accounts/fireworks/models/inklinginklinginkling-freelaohuang/thinkingmachines/inklingthinkingmachines/inkling

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation