JevBench by Benchmark Heaven · released v1.4.1

Jev alternatives, compared on the published board

The best Jev alternative depends on your use case. Compare Jev-class decision models using the same released benchmark: Intelligence, Calibration, Speed and Cost, with each row’s evidence and openness notes.

JevBench Scores in the current top five

The Jev row is the reference; the four rows below it are current alternatives. The overall score is a composite, so check the separate axes for your use case.

  1. 1Jev 1.13.0API63.3I 53 · C 76 · S 83 · K 52 · $0.040
  2. 2JevK5 v0.2.062.0I 49 · C 75 · S 91 · K 60 · ~$0.022 est.
  3. 3Hopper59.4I 48 · C 79 · S 87 · K 59 · ~$0.024 est.
  4. 4Winnow-12B Q855.6I 48 · C 65 · S 82 · K 53 · ~$0.037 est.
  5. 5reflex 4B54.0I 47 · C 70 · S 68 · K 60 · ~$0.022 est.
All top-five values as a table
RankSystemScoreIntelligenceCalibrationSpeedCostCost basis
1Jev 1.13.0 (TypeSafe AI)
Reference system
63.353.176.383.352.0measured
2JevK5 v0.2.062.048.974.591.159.5estimated
3Hopper59.448.079.186.858.7estimated
4Winnow-12B Q855.648.364.882.352.9estimated
5reflex 4B (kshetrajna12)54.047.570.468.059.7estimated

Cost evidence is labeled by basis so estimated and announced values are not presented as measured charges.

Looking for an open source Jev alternative?

These rows have explicit openness, license and repository fields in the published artifact. “Open” is the board’s status; read the linked source and exact terms before deploying a system.

For a broader list, use the board’s row disclosures. Rows without explicit openness, license and repository evidence are not classified here.

See the use-case chooser, including accuracy, speed, cost and self-hosting evidence.

Frequently asked questions

What does JevBench compare?
JevBench compares published Jev-class decision systems across Intelligence, Calibration, Speed and Cost. This page uses the released v1.4.1 aggregate.
How should I choose between Jev alternatives?
There is no single best fit for every use. Compare the published axes and their evidence, then choose for your use case. Intelligence, Calibration, Speed and Cost — equal-weight harmonic mean, with generalization and Jev-class gates.
Which Jev alternatives have public self-hosting evidence?
The list below includes only rows marked as open or open weights in the published artifact that also include a license note and public repository link. Check the linked source and its terms before deployment; missing evidence remains unknown.
Are cost and speed values directly measured?
The board labels cost evidence as measured, estimated or announced and shows the conditions behind speed measurements where available. The Cost and Speed axes are benchmark scores, not a universal bill or wall-clock guarantee.