← All models

Grok 4.3 (medium)

deprecated by benchmark source
xAI · released 2026-04-30 · 5 offers

Output 133 tokens/sFirst token 25 sContext 1M tokens

Top 3 cheapest providers (Adjusted $/task)

The same list price can give a different adjusted $/task (caching, token efficiency) — click a price for its inputs.

Within the active global provider, residency and confidentiality filters.

#ProviderAdjusted $/task
1AWS Bedrock
2Azure AI Foundry
3Google Vertex AI

Composite

3 of 7 inputs · 2 from the model family6 radar axes: DesignArena's two boards share one

46.6

includes −8.1 for its Benchmaxxing signal (from 54.8; why, switch off in Options)

dominance-adjusted from 58.0: a better-measured model that is at least as good on each of these inputs ranks above it

Benchmaxxing signal +8.1, medium Benchmaxxing tag, uncertain: The 80 % interval reaches below zero — treat this tag as uncertain · medium report →The 80 % interval reaches below zero — treat this tag as uncertain.

AA CodingCoding Agent v1.4AA IntelligenceAA AgenticEpoch ECISoftware ECIDesignArenaAA Intelligence: percentile 80DesignArena: percentile 16

AA Coding Coding Agent v1.4 AA Intelligence 24.8AA Agentic Epoch ECI Software ECI DesignArena 1144/1017

Radar: percentile among all models measured on each input; a gap means not measured.

Benchmark sheet

9 of 140 registered benchmark versions · bars show the percentile among all models measured on each benchmark.

Composite attachments (used in the score, not counted as exact benchmarks):
  • DesignArena Web Apps (agentic) · attached
  • DesignArena Full-Stack · attached
Compare this model →

Agentic

Instruction-following

Knowledge

Long-context

  • AA-LCR v1.17875.0%

    AA-LCR v1.1 v1.1 · Published board — Tests reasoning across multiple long documents with corrected answer keys and grading.

Reasoning

Science

Tool-use

Vision

Missing coverage · 135 benchmark versions

No result does not mean a zero, or that the model was never tested. Collection failures and disputed versions retain their distinct status.

Variants / reasoning settings

Artificial Analysis snapshot 2026-09-19 · Data: Artificial Analysis · Data: Epoch AI (CC BY)

VariantAA CodingAA Intelligence
Grok 4.3 (high)42.225.4
Grok 4.3 (Non-reasoning)35.214.5
Grok 4.3 (medium)24.8
Grok 4.3 (low)24.3
Token offers by platform · 3 offers (Adjusted $/task)

Click any underlined price to see how it is estimated and where each input comes from. How we calculate adjusted cost.

3 offers within the active global filters; “—” means the catalog is active but no public token price is available.

AWS Bedrock (1)

AWS Bedrock

Azure AI Foundry (1)

Azure AI Foundry

Google Vertex AI (1)

Google Vertex AI
Grok 4.3 (medium) — benchmarks & cost | Benchmark Heaven