JevMade Sign in
← Back to experiments

Benchmarks & research

jev-certify

Turns Jev's probabilities into a routing threshold with a provable bound on how many queries get silently misrouted.

Bookmark: jev-certify Keep this in your collection.
Leave a noteWhat would you try with this? : jev-certify

Only you can see your notes.

Source screenshot of jev-certify
SOURCE SCREENSHOTFull screenshot ↗

What it does

It also shows where the bound stops working: Jev returns exactly 1.0 on more than half its answers, and some of those are wrong.

How you can use it

A developer can use this code to sort incoming messages. The app connects to an outside AI service called Jev using an access key. Math rules limit how often the AI sends a message to the wrong place.

You can test the tool by grading about forty AI answers yourself. The system cannot catch every mistake. The AI sometimes claims it is completely sure when it is actually wrong. Sudden changes in what people ask will also break these rules.

Maker-reported (not independently measured by JevMade): 2,412 journalled Jev decisions costing $0.23 in total · at a 5% risk target Jev settles 84.75% of traffic with a 2.65% error rate among settled queries · 40 labels beat what 254 labels achieve

Primitives
choice, noul
Platform
Python
Added
Project created

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.