JevMade hello@JevMade.com
← Back to experiments

Benchmarks & research

jeval

Evaluates whether classifier confidence is calibrated and derives human-review thresholds from the cost of mistakes.

Source screenshot of jeval
SOURCE SCREENSHOT · source ↗ · captured 2026-09-25Full screenshot ↗