← All models

GPT-6 Sol (max)

★ featured
OpenAI · released 2026-09-22 · 3 offers

Context 872K tokens

Top 1 cheapest providers (Adjusted $/task)

Within the active global provider, residency and confidentiality filters.

#ProviderAdjusted $/task
1Amazon Bedrock

Composite

1 of 7 inputs6 radar axes: DesignArena's two boards share one

91.7

dominance-adjusted from 95.0: a better-measured model that is at least as good on each of these inputs ranks above it

AA CodingCoding Agent v1.4AA IntelligenceAA AgenticEpoch ECISoftware ECIDesignArenaAA Intelligence: percentile 97

AA Coding Coding Agent v1.4 AA Intelligence 47.5AA Agentic Epoch ECI Software ECI DesignArena —/—

Radar: percentile among all models measured on each input; a gap means not measured.

Benchmark sheet

2 of 265 registered benchmark versions · bars show the percentile among all models measured on each benchmark.

2 of 3 values are OpenAI's own claims (†), not independent measurements; a matching independent result replaces a claim as soon as one exists.

Compare this model →

Agentic

  • Agents' Last Exam V1developer's claim56.4% self-reported by the developer

    Agents' Last Exam V1 v1 · Published board — Agents' Last Exam V1 score OpenAI reports for its own GPT-6 models in the 2026-09-22 GPT-6 Sol and Luna launch post.

Coding

  • DeepSWE v1.1developer's claim68.8% self-reported by the developer

    DeepSWE v1.1 v1.1 · Published board — DeepSWE v1.1 score OpenAI reports for its own GPT-6 models in the 2026-09-22 GPT-6 Sol and Luna launch post.

Reasoning

Missing coverage · 267 benchmark versions

No result does not mean a zero, or that the model was never tested. Collection failures and disputed versions retain their distinct status.

Variants / reasoning settings

Artificial Analysis snapshot 2026-09-22 · Data: Artificial Analysis · Data: Epoch AI (CC BY)

VariantAA CodingAA Intelligence
GPT-6 Sol (max)47.5
GPT-6 Sol (xhigh)44.1
GPT-6 Sol (high)42.8
GPT-6 Sol (medium)39.8
GPT-6 Sol (low)33.9
GPT-6 Sol (Non-reasoning)28.1
Token offers by platform · 1 offers (Adjusted $/task)

Click any underlined price to see how it is estimated and where each input comes from. How we calculate adjusted cost.

1 offers within the active global filters; “—” means the catalog is active but no public token price is available.

OpenRouter (1)

Amazon Bedrock
GPT-6 Sol (max) — benchmarks & cost | Benchmark Heaven