decider-12b v2 (Mapika)

Data: JevBench v1.6.1, published 2026-10-06

Capability 72.5 · rank #8 of 92 on the open-weights board

Last measured: 2026-10-04 · measurement release: v1.6.1

apache-2.0

Base model: google/gemma-4-12B-itsource

intelligence

63.0

calibration

82.0

speed

90.5

cost

57.8

Composite (secondary)

70.9 · Composite rank #4

Cost and measurement conditions

$0.026 per 1,000 decisions (estimated)

add-requests run 46 (v1.5 basis, unpublished) documented hosted-model estimate; no exact base-model floor applies

p50 latency: 0.3 s

x2 + 0.15 s (assumption, not measured)

Current frozen 1500-item pool measured 2026-10-04 by coordinator plan P83 on one evaluator-owned Lium RTX PRO6000. Offline contributor image, serial loopback-only requests, no batching or retry; standard selfhosted x2+0.15 s latency adjustment. Current stage receipt verifies the same model pin used by the prior add-request run. Cost is the carried documented estimate recorded in current registry.

Published source