JevMade Sign in
← Back to guides

JevMade field notes / Architecture and workflow guide

Adding safety checks to an AI assistant using a decision tool

This guide explains how to check an AI assistant's planned actions before they happen. It uses a tool to spot dangerous commands, hide private information, and handle errors without breaking the workflow.

Original by y0usafGuardrails

Listen to this guide

JevMade’s plain-English explanation

0:00 /

Our summary

This guide explores an add-on for an AI assistant that writes code. It uses Jev, an AI tool that chooses from options rather than writing answers. People use this to review the assistant's planned actions and catch dangerous commands before anything runs.

The software asks specific questions about a planned task, like whether it deletes files or shares data. It shortens long text before sending it for review. After a task finishes, it checks the results to see if any API keys, which are private access codes, leaked.

This setup is useful for developers who want to monitor their AI assistant. If the safety check loses connection, the assistant continues working anyway. The author only tested these settings on a small number of examples, so the limits might not catch every problem.

Key takeaways

  1. Cut down long text and remove sensitive details before sending data for an outside safety check.
  2. Group different types of errors into categories to give the AI assistant clear instructions on fixing them.
  3. Let the assistant keep working if the safety check fails, but limit how often errors are reported.

The repository author reported these behaviors and test results.

Repository README and documentation

Read the original guide Opens the author’s site in a new tab.

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.