JevMade Sign in
← Back to experiments

Benchmarks & research

VIGILIA Jev evaluation

Compare Jev with VIGILIA’s own labels on a sealed set of risky and harmless agent actions.

Source screenshot of VIGILIA Jev evaluation
SOURCE SCREENSHOTFull screenshot ↗

What it does

The maker reports agreement, refusals, and limitations; this is not independent proof of correctness.

How you can use it

Collect a private list of past software actions that you already checked yourself. Mark each example as either a real danger or a false alarm. Next, have a developer write code to send those exact examples to TypeSafe and ask the model to choose between your labels.

Remember that high agreement between two automated checks is not proof that an action is truly safe. Keep human review for sensitive steps. The model may also refuse to review certain activities, such as software accessing password files before contacting the internet.

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.