JevMade Sign in
← Back to experiments

Agent tooling

im-in-danger

Checks content an agent is about to read for instructions aimed at the agent, and returns a trust verdict plus capability advice rather than a boolean.

Bookmark: im-in-danger Keep this in your collection.
Leave a noteWhat would you try with this? : im-in-danger

Only you can see your notes.

Source screenshot of im-in-danger
SOURCE SCREENSHOTFull screenshot ↗

What it does

It runs on Jev by default, with a local-model detector and a keyword detector also included, and benchmarks them against an 85-item corpus. It publishes its own failures and says plainly that it is a filter, not a security boundary.

How you can use it

Start by listing the high-risk actions your automated assistant performs, such as updating code or sending emails. Have your developer add this tool into your assistant's live workflow to scan incoming web pages and messages before reading them. Using a service access key, the scanner checks the text and flags suspicious commands.

Decide which sensitive abilities to pause when text looks risky, rather than relying on the tool to block everything. Remember that this check is only an early warning filter, not a complete safety guarantee. It can still miss clever attacks, so truly critical tasks should always require approval from an actual person.

Maker-reported (not independently measured by JevMade): 85-item corpus: 41 items carrying instructions aimed at an agent, 44 ordinary business items · Jev 1.13 | threshold 0.9: caught 40/41, false alarms 1/44, median latency 120–250 ms · Local, Qwen3.8-27B IQ3_S: 38/40 caught, 2/40 false alarms, 74 s

Primitives
noul
Platform
TypeScript (npm library)
Added
Project created

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.