JevMade hello@JevMade.com
← Back to experiments

Benchmarks & research

system-one-gemma

system-one-gemma adds a scoring head to Gemma 3 270M for calibrated decisions without text generation.

Source screenshot of system-one-gemma
SOURCE SCREENSHOT · source ↗ · captured 2026-09-21Full screenshot ↗

What it does

It was trained on 12,913 questions spanning six tasks; accuracy is modest, but reported confidence closely tracks observed correctness.

Maker-reported (not independently measured by JevMade): Trained on 12,913 questions across 6 tasks, evaluated on held-out data: · The accuracy is modest (it's a 270M model!) but the calibration is excellent — when it says 80% confidence, it's right ~80% of the time. That's the real value: you can trust the probabilities

Primitives
choice, score, noul
Platform
Python
Added
Project created
GitHub stars
0 · snapshot 2026-09-18T22:58:45Z

Source checked 2026-09-19 — opened the GitHub repository directly.