← All models

GPT-6 Sol (xhigh)

★ featured
OpenAI · released 2026-09-22 · 3 offers

Context 872K tokens

Top 1 cheapest providers (Adjusted $/task)

Within the active global provider, residency and confidentiality filters.

#ProviderAdjusted $/task
1Amazon Bedrock

Composite

1 of 7 inputs6 radar axes: DesignArena's two boards share one

84.4

dominance-adjusted from 91.1: a better-measured model that is at least as good on each of these inputs ranks above it

AA CodingCoding Agent v1.4AA IntelligenceAA AgenticEpoch ECISoftware ECIDesignArenaAA Intelligence: percentile 95

AA Coding Coding Agent v1.4 AA Intelligence 44.1AA Agentic Epoch ECI Software ECI DesignArena —/—

Radar: percentile among all models measured on each input; a gap means not measured.

Benchmark sheet

3 of 265 registered benchmark versions · bars show the percentile among all models measured on each benchmark.

3 of 4 values are OpenAI's own claims (†), not independent measurements; a matching independent result replaces a claim as soon as one exists.

Compare this model →

Agentic

  • AutomationBench 1.0.6 v1.0.6developer's claim33.2% self-reported by the developer

    AutomationBench 1.0.6 v1.0.6 · Published board — AutomationBench 1.0.6 score OpenAI reports for its own GPT-6 models in the 2026-09-22 GPT-6 Sol and Luna launch post.

  • OSWorld 2.0 offline (v2026.08.08 release)developer's claim60.5% self-reported by the developer

    OSWorld 2.0 offline (v2026.08.08 release) v2026.08.08 · Published board — OSWorld 2.0 offline partial reward OpenAI reports for its own GPT-6 models in the 2026-09-22 GPT-6 Sol and Luna launch post.

Efficiency

  • AutomationBench 1.0.6 cost per task v1.0.6developer's claim$0.27 self-reported by the developer

    AutomationBench 1.0.6 cost per task v1.0.6 · Published board — USD cost per AutomationBench 1.0.6 task OpenAI reports for its own GPT-6 models in the 2026-09-22 GPT-6 Sol and Luna launch post.

Reasoning

Missing coverage · 266 benchmark versions

No result does not mean a zero, or that the model was never tested. Collection failures and disputed versions retain their distinct status.

Variants / reasoning settings

Artificial Analysis snapshot 2026-09-22 · Data: Artificial Analysis · Data: Epoch AI (CC BY)

VariantAA CodingAA Intelligence
GPT-6 Sol (max)47.5
GPT-6 Sol (xhigh)44.1
GPT-6 Sol (high)42.8
GPT-6 Sol (medium)39.8
GPT-6 Sol (low)33.9
GPT-6 Sol (Non-reasoning)28.1
Token offers by platform · 1 offers (Adjusted $/task)

Click any underlined price to see how it is estimated and where each input comes from. How we calculate adjusted cost.

1 offers within the active global filters; “—” means the catalog is active but no public token price is available.

OpenRouter (1)

Amazon Bedrock
GPT-6 Sol (xhigh) — benchmarks & cost | Benchmark Heaven