decider-12b v1, stock Gemma-4-12B-it (Mapika)

Data: JevBench v1.6.1, published 2026-10-06

Capability 72.2 · rank #10 of 92 on the open-weights board

Last measured: 2026-10-04 · measurement release: v1.6.1

apache-2.0

Base model: google/gemma-4-12B-itsource

intelligence

60.6

calibration

83.8

speed

90.5

cost

57.8

Composite (secondary)

70.4 · Composite rank #5

Cost and measurement conditions

$0.026 per 1,000 decisions (estimated)

add-requests run 46 (v1.5 basis, unpublished) documented hosted-model estimate; no exact base-model floor applies

p50 latency: 0.3 s

x2 + 0.15 s (assumption, not measured)

Current frozen 1500-item pool measured 2026-10-04 by coordinator plan P83 on one evaluator-owned Lium RTX PRO6000. Offline contributor image, serial loopback-only requests, no batching or retry; standard selfhosted x2+0.15 s latency adjustment. Current stage receipt verifies the same model pin used by the prior add-request run. Cost is the carried documented estimate recorded in current registry.

Published source