Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GPT-6 Luna

Basis

GPT-6 Luna is strong in long context; capable in reasoning and factuality; and behind the leaders in multimodal tasks, agentic tasks, coding, and math. Too few results yet to rate safety, multilingual tasks, or instruction following.

Price per million tokens

$0.10input$0.50output

Price from Artificial Analysis · 3 providers tracked · All prices

Cheaper than 80% of 323 priced models · 3:1 input-to-output blend, log scale
Evidence

121results on44benchmarks

  • 33 independently verified
  • 66 aggregator
  • 3 vendor-reported
  • 19 cross-referenced

From 10 sources · latest Sep 29, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

Capability profile

Bars show the model's median result as a share of the leading model's, per capability.

Strong1 capability
  1. Long Context−6.1%1 of 3
Capable2 capabilities
  1. Reasoning−21.4%4 of 6
  2. Factuality−24.7%2 of 4
Limited4 capabilities
  1. Multimodal−27.4%1 of 6
  2. Agentic−30.7%2 of 7
  3. Coding−31.6%4 of 10
  4. Math−37.7%1 of 5

Too few results to rate: SafetyMultilingualInstruction Following

GPT-6 Luna benchmark results

121 results on 44 benchmarks, grouped by capability. Every score links to its source; a bar is the result as a share of the capability leader's.

Long Context

Strong1 of 3 ranked benchmarks measuredFull long context ranking
  • AA-LCR
    aa_lcr
    83.3394%
    Max
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
5 more long context results
  • AA-LCR
    aa_lcr
    74.00—
    Low
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-LCR
    aa_lcr
    39.67—
    No reasoning
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-LCR
    aa_lcr
    79.33—
    High
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-LCR
    aa_lcr
    80.00—
    xHigh
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-LCR
    aa_lcr
    78.33—
    Medium
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗

Reasoning

Capable4 of 6 ranked benchmarks measuredFull reasoning ranking
27 more reasoning results

Factuality

Capable2 of 4 ranked benchmarks measuredFull factuality ranking
10 more factuality results

Multimodal

Limited1 of 6 ranked benchmarks measuredFull multimodal ranking
  • MMMU-Pro
    aa_mmmu_pro
    75.5586%
    Max
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
5 more multimodal results
  • MMMU-Pro
    aa_mmmu_pro
    70.35—
    Low
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • MMMU-Pro
    aa_mmmu_pro
    53.24—
    No reasoning
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • MMMU-Pro
    aa_mmmu_pro
    74.39—
    High
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • MMMU-Pro
    aa_mmmu_pro
    73.01—
    Medium
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • MMMU-Pro
    aa_mmmu_pro
    75.26—
    xHigh
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗

Agentic

Limited2 of 7 ranked benchmarks measuredFull agentic ranking
10 more agentic results

Coding

Limited4 of 10 ranked benchmarks measuredFull coding ranking
5 more coding results
  • SciCode
    aa_scicode
    46.88—
    Low
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • SciCode
    aa_scicode
    43.06—
    No reasoning
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • SciCode
    aa_scicode
    50.35—
    High
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • SciCode
    aa_scicode
    51.74—
    xHigh
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • SciCode
    aa_scicode
    50.93—
    Medium
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗

Math

Limited1 of 5 ranked benchmarks measuredFull math ranking

Instruction Following

Not enough data0 of 3 ranked benchmarks measuredFull instruction following ranking

More results

37 more results
  • AA Intelligence
    Artificial Analysis Intelligence Index
    37.00—
    Max
    artificialanalysis.aigpt-6-luna
    VerifiedSep 23, 2026view ↗
  • AA Intelligence
    aa_intelligence_index
    20.92—
    Low
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA Intelligence
    aa_intelligence_index
    18.26—
    No reasoning
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA Intelligence
    aa_intelligence_index
    37.26—
    Max
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA Intelligence
    aa_intelligence_index
    32.15—
    High
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA Intelligence
    aa_intelligence_index
    29.46—
    Medium
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA Intelligence
    aa_intelligence_index
    33.88—
    xHigh
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-Omniscience
    aa_omniscience
    -9.17—
    Low
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-Omniscience
    aa_omniscience
    -5.50—
    High
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-Omniscience
    aa_omniscience
    0.65—
    Max
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-Omniscience
    aa_omniscience
    -21.90—
    No reasoning
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-Omniscience
    aa_omniscience
    -5.05—
    Medium
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • AA-Omniscience
    aa_omniscience
    -1.83—
    xHigh
    artificialanalysis.ai
    AggregatorSep 29, 2026view ↗
  • Agentic Safe Completions - Chat prod - Chat Plugins
    0.85—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • 50.90—
    api.llm-stats.comscores
    VendorSep 29, 2026view ↗
  • 73.00—
    xHigh
    arcprize.orgdata
    VerifiedSep 23, 2026view ↗
  • 86.67—
    Max
    arcprize.orgdata
    VerifiedSep 23, 2026view ↗
  • 66.60—
    api.llm-stats.comscores
    VendorSep 29, 2026view ↗
  • Dynamic Benchmarks - Emotional reliance
    0.96—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • Dynamic Benchmarks - Mental health
    1.00—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • Dynamic Benchmarks - Self-harm
    0.92—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • 95.90—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • 31.40—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • HealthBench length-adjusted
    HealthBench length-adjusted
    54.50—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • 60.80—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • Image input evaluations - extremism
    0.98—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • Image input evaluations - harms-erotic
    0.99—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • Image input evaluations - hate
    1.00—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • Image input evaluations - self-harm
    1.00—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • 52.70—
    api.llm-stats.comscores
    VendorSep 29, 2026view ↗
  • Production Benchmarks - Gore
    0.88—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • U18 evaluations - Age-restricted goods, services, and dangerous challenges / activities
    0.85—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • U18 evaluations - Eating Disorders
    0.87—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • U18 evaluations - Emotional Reliance
    0.95—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • U18 evaluations - Gore
    0.88—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • U18 evaluations - Self Harm
    0.98—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗
  • U18 evaluations - Sexual Content
    0.95—
    deploymentsafety.openai.comgpt-6-1-sol
    Cross-refSep 29, 2026view ↗

About this record

Verification: 121 scores · 33 independently verified · 66 aggregator-attributed · 19 vendor cross-reference · 3 vendor-reported. How these tiers are assigned

Also listed as
gpt-6-luna-xhighgpt-6-luna-mediumgpt-6-luna-non-reasoninggpt-6-luna-highgpt-6-luna-lowgpt-6 luna (xhigh)gpt-6 luna (medium)gpt-6 luna (non-reasoning)gpt-6 luna (max)gpt-6 luna (high)gpt-6 luna (low)gpt-6 luna (none)gpt-6 luna - provider adapter (max)gpt-6 luna - provider adapter (xhigh)gpt-6 luna - provider adapter (high)gpt-6-luna-maxand 2 more
Tracked since
Sep 22, 2026
Newest source mention
Sep 27, 2026