← All models

Llama 3.3 70B

open weights
Meta · 4 offers

Top 4 cheapest providers (Adjusted $/task)

The same list price can give a different adjusted $/task (caching, token efficiency) — click a price for its inputs.

Within the active global provider, residency and confidentiality filters.

#ProviderAdjusted $/task
1Azure AI Foundry
2IONOS
3AWS Bedrock
4Google Vertex AI

Composite

1 of 7 inputs · 1 from the model family6 radar axes: DesignArena's two boards share one

5.3

AA CodingCoding Agent v1.4AA IntelligenceAA AgenticEpoch ECISoftware ECIDesignArenaEpoch ECI: percentile 7

AA Coding Coding Agent v1.4 AA Intelligence AA Agentic Epoch ECI 127.3Software ECI DesignArena —/—

Radar: percentile among all models measured on each input; a gap means not measured.

Benchmark sheet

2 of 140 registered benchmark versions · bars show the percentile among all models measured on each benchmark.

Composite attachments (used in the score, not counted as exact benchmarks):
  • Epoch ECI · attached
Compare this model →

Agentic

  • Elimination Game Benchmark504

    Elimination Game Benchmark published 2026-09-10 · Published board — Multi-player tournament where 8 LLM players converse, form alliances, and vote to eliminate each other, with final rankings scored via TrueSkill.

Other

Missing coverage · 143 benchmark versions

No result does not mean a zero, or that the model was never tested. Collection failures and disputed versions retain their distinct status.

Token offers by platform · 4 offers (Adjusted $/task)

Click any underlined price to see how it is estimated and where each input comes from. How we calculate adjusted cost.

4 offers within the active global filters; “—” means the catalog is active but no public token price is available.

Azure AI Foundry (1)

Azure AI Foundry

IONOS (1)

IONOS

AWS Bedrock (1)

AWS Bedrock

Google Vertex AI (1)

Google Vertex AI
Llama 3.3 70B — benchmarks & cost | Benchmark Heaven