Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GPT-5.6 Luna vs Inkling-Small

OpenAI · released

Wins 7 of 7 areas

Coding · Agents · Reasoning · Facts · Images and charts · Math · Long documents

Thinking Machines · released

Wins 0 of 7 areas

—

GPT-5.6 Luna is the stronger all-rounder.

Scores updated · 37 tests both models report · How we compare

What it costs

Prices per million tokens, roughly 750,000 words. The bars show the cost of a million tokens read plus a million written.

GPT-5.6 Luna$0.20 to read · $1.20 to write$1.40Price from Artificial Analysis
Inkling-Small$0.30 to read · $1.20 to write$1.50Price from Artificial Analysis

GPT-5.6 Luna costs 7% less for the same work.

The biggest differences

The tests each model wins by the widest margin, up to three each. Scores are out of 100.

Where GPT-5.6 Luna pulls ahead

  • Extremely hard research-level math problemsFrontierMath Tier 4+41.4points aheadGPT-5.6 Luna58.5Inkling-Small17.1
  • Unpublished advanced math problemsFrontierMath Tiers 1-3 (v2)+35.8points aheadGPT-5.6 Luna82.1Inkling-Small46.3
  • Command-line tasks in a real terminalTerminal-Bench 2.1+25.8points aheadGPT-5.6 Luna80.9Inkling-Small55.1

Where Inkling-Small pulls ahead

Every test, side by side

All 37 tests both models report. The winning score is in its model's colour; marks a score checked independently.

CodingGPT-5.6 Luna

Full coding ranking

AgentsGPT-5.6 Luna

Full agents ranking

ReasoningGPT-5.6 Luna

Full reasoning ranking

FactsGPT-5.6 Luna

Full facts ranking

Images and chartsGPT-5.6 Luna

Full images and charts ranking

MathGPT-5.6 Luna

Full math ranking

Long documentsGPT-5.6 Luna
  • AA-LCRGPT-5.6 Luna by 883.775.7+8

Full long documents ranking

Other results13 tests, not counted

Tests outside the eight areas. They are not counted above: several are summary scores built from other tests, or the same test under another name.

Questions people ask

Which is better, GPT-5.6 Luna or Inkling-Small?

GPT-5.6 Luna wins all seven areas where both have results: coding, agents, reasoning, facts, images and charts, math and long documents. Inkling-Small wins none.

Which is better for coding?

GPT-5.6 Luna. It wins 5 of the 5 coding tests both models report; Inkling-Small wins none.

Which is cheaper?

GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output tokens; Inkling-Small costs $0.30 and $1.20. That makes GPT-5.6 Luna about 7% cheaper for the same work.

How do you compare the two?

We use the 37 benchmark tests both models have published scores on. The verdict counts the 24 tests in the eight capability areas, and a gap under one point (ten on rating-style scales) counts as a tie. The other 13 are listed but not counted, because several are summary scores or repeat a test. Each score is the one shown on the model's own page.