Inkling
Inkling is capable in long context, instruction following, and factuality; and behind the leaders in reasoning, agentic tasks, multimodal tasks, coding, and math. Too few results yet to rate safety or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety or Multilingual.
Price
$1.00input$4.05outputper million tokens
From Artificial Analysis · 4 providers tracked · All prices
Evidence
80results on53benchmarks
- 17 independently verified
- 16 aggregator
- 47 vendor-reported
From 16 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingReasoning
As listed by OpenRouter
Research
2 papers reference InklingInkling benchmark results
80 results on 53 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
16.0% behind the leader1 of 3 ranked benchmarks measured
- 77.33Oct 8, 2026
18.3% behind the leader2 of 3 ranked benchmarks measured
- 79.80Oct 7, 2026
- LiveBench · Instruction Following70.10Oct 8, 2026livebench_instruction_following@2026-06-25
Show 1 more instruction following resultHide 1 instruction following result
- 83.40Jul 16, 2026
22.1% behind the leader3 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy41.55Oct 8, 2026omniscienceAccuracy
- 43.90Oct 7, 2026
- AA-Omniscience · Non-hallucination32.34Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- 40.30Sep 21, 2026
- 20.90Jul 16, 2026
27.5% behind the leader6 of 6 ranked benchmarks measured
- GPQA Diamond87.17Oct 8, 2026gpqa
- LiveBench · Reasoning78.35Oct 8, 2026livebench_reasoning@2026-06-25
- 50.00Jul 23, 2026
- Humanity's Last Exam31.88Oct 8, 2026aa_hle
- 36.53Jul 18, 2026
- 5.43Oct 8, 2026
Show 5 more reasoning resultsHide 5 reasoning results
- 5.40Jul 30, 2026
- 87.20Jul 30, 2026
- 88.30Jul 16, 2026
- Humanity's Last Exam29.70Jul 30, 2026HLE (text only)
- Humanity's Last Exam29.60Jul 16, 2026HLE text only
29.3% behind the leader5 of 7 ranked benchmarks measured
- 76.00Oct 7, 2026
- τ-Bench V3 · Banking29.07Oct 8, 2026tauBanking
- 28.78Oct 8, 2026
- Terminal-Bench 4.01.01Oct 8, 2026
30.8% behind the leader2 of 6 ranked benchmarks measured
- CharXiv (reasoning)78.10Oct 7, 2026CharXiv-R
- MMMU-Pro73.47Oct 8, 2026aa_mmmu_pro
Show 1 more multimodal resultHide 1 multimodal result
- 73.50Oct 7, 2026
36.9% behind the leader6 of 10 ranked benchmarks measured
- 77.60Oct 7, 2026
- LiveBench · Coding71.02Oct 8, 2026livebench_coding@2026-06-25
- SciCode46.99Oct 8, 2026aa_scicode
- LiveBench · Agentic Coding49.39Oct 8, 2026livebench_agentic_coding@2026-06-25
- Terminal-Bench 2.155.06Oct 8, 2026terminalbenchV21
- 1412.18Aug 29, 2026
Show 3 more coding resultsHide 3 coding results
- 46.10Jul 30, 2026
- 77.40Jul 16, 2026
- 63.80Oct 7, 2026
52.2% behind the leader3 of 5 ranked benchmarks measured
- LiveBench · Mathematics88.36Oct 8, 2026livebench_math@2026-06-25
- 33.33Sep 21, 2026
- 4.88Sep 21, 2026
Show 3 more math resultsHide 3 math results
- 97.10Oct 7, 2026
- 95.10Jul 16, 2026
- 86.30Jul 30, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_data_analysis72.78Oct 8, 2026livebench_data_analysis@2026-06-25
- livebench_language73.46Oct 8, 2026livebench_language@2026-06-25
- AA Intelligence25.00Sep 21, 2026Artificial Analysis Intelligence Index
- 1445Jul 23, 2026
- 79.50Jul 18, 2026
- AA Intelligence24.98Oct 8, 2026aa_intelligence_index
Show 29 more resultsHide 29 results
- 24.32Sep 9, 2026
- 2.00Oct 8, 2026
- 79.50Jul 30, 2026
- Artificial Analysis Coding Index52.06Sep 9, 2026aa_coding_index
- BrowseComp (context management)77.10Jul 30, 2026BrowseComp (with context management)
- 77.10Jul 16, 2026
- 77.10Jul 16, 2026
- 83.40Jul 16, 2026
- charxiv_rq_with_python82.00Jul 16, 2026Charxiv RQ (with python)
- 61.10Jul 21, 2026
- 1238.00Jul 30, 2026
- 1233.00Jul 16, 2026
- 88.70Oct 7, 2026
- 86.80Jul 16, 2026
- 46.00Jul 30, 2026
- HLE (with tools)46.60Jul 16, 2026HLE with tools
- 77.20Oct 7, 2026
- 77.50Jul 16, 2026
- 73.10Jul 16, 2026
- 0.16Jul 16, 2026
- 98.80Jul 16, 2026
- 98.60Jul 16, 2026
- SWEBench Pro Public54.30Jul 30, 2026SWEBench Pro (public)
- 53.20Jul 16, 2026
- 45.50Jul 30, 2026
- 90.00Jul 16, 2026
- 91.40Jul 16, 2026
- τ³-Bench Banking23.70Oct 7, 2026Tau3 Banking
- τ³-Bench Banking13.60Jul 16, 2026Tau 3 Banking
Inkling: common questions
Who makes Inkling?
Inkling is made by Thinking Machines.
When was Inkling released?
Inkling was released on Jul 15, 2026, according to Thinking Machines's own announcement.
What is Inkling good at?
Inkling is capable in long context, instruction following, and factuality; and behind the leaders in reasoning, agentic tasks, multimodal tasks, coding, and math. Too few results yet to rate safety or multilingual tasks.
How much does Inkling cost?
Inkling costs $1.00 per million input tokens and $4.05 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 69% of the 331 priced models we track.
How many benchmarks has Inkling been tested on?
We track 80 results for Inkling on 53 benchmarks from 16 sources, 17 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Inkling support?
OpenRouter lists tool calling and reasoning for Inkling.
About this record
Where Inkling's numbers come from, and every name it appears under.
- Tracked since
- Jul 15, 2026
- Newest source mention
- Sep 16, 2026
Where the results come from
Verification: 80 scores · 17 independently verified · 16 aggregator-attributed · 47 vendor-reported. How these tiers are assigned
From 16 sources on 11 sites. thinkingmachines.ai supplies 19 of them; the 17 independently verified results come from 8 sites. Bars are coloured by trust tier.
- thinkingmachines.ai19
- artificialanalysis.ai17
- huggingface.co17
- api.llm-stats.com11
- livebench.ai7
- epoch.ai3
- arcprize.org2
- datasets-server.huggingface.co1
- labs.scale.com1
- lmarena.ai1
- simple-bench.com1