SIDE BY SIDE
Compare selected AI models side by side.
Compare uses, prices, speed evidence, and version details for up to three models. Language, video, decision models and historical evaluations are shown separately.
Decision & classification models
| Decision criteria | Kev-4B ↗Jared Palmer |
|---|---|
| Version reviewed | Kev 1.0 · Kev-4B |
| Judgment types | Choose an option · Yes / no probability · Rubric score |
| Access & hardware | Self-host on an NVIDIA GPU or Apple Silicon via MLX. Training and serving code are available; CPU deployment is not verified here. |
| API price · USD | Hosted price unverifiedNo hosted rate verified for this entry. Self-hosting needs compute, capacity and maintenance; downloadable weights do not make inference free. |
| Languages | English; the model card places other languages outside its intended scope. |
| Context conditions | Validated context: 8,192 tokens. The server accepts longer states, but that is not equivalent to validated accuracy. |
| Probability interpretation | The release applies a fitted temperature. Probability thresholds need rechecking after model, domain or calibration changes. |
| Our accuracy / calibration | Not measured |
| Our p50 / p95 latency | Not measured |
| Our cost / 1,000 decisions | Not measured |
| Limits | API compatibility does not establish equal accuracy or transferable thresholds. Keep the checkpoint and precision with any benchmark; this catalog does not combine its different evaluation suites. |
| Sources | Jared Palmer · Kev-4B model card ↗Checked 2026-10-05 |
Source-reviewed capabilities, not a leaderboard. Hosted prices and self-hosting costs use different bases. Unknown measurements are not zero. Measure the complete task →