JevMade hello@JevMade.com
← Back to experiments

Benchmarks & research

jev-benchmarks

A reproducible evaluation suite measures calibration, selective risk and latency in probabilistic decision models.

Source screenshot of jev-benchmarks
SOURCE SCREENSHOT · source ↗ · captured 2026-09-21Full screenshot ↗
Primitives
choice, score
Platform
Python
Added
Project created
GitHub stars
8 · snapshot 2026-09-18T22:58:45Z

Source checked 2026-09-19 — opened the GitHub repository directly.