Our summary
This guide explores an add-on for an AI assistant that writes code. It uses Jev, an AI tool that chooses from options rather than writing answers. People use this to review the assistant's planned actions and catch dangerous commands before anything runs.
The software asks specific questions about a planned task, like whether it deletes files or shares data. It shortens long text before sending it for review. After a task finishes, it checks the results to see if any API keys, which are private access codes, leaked.
This setup is useful for developers who want to monitor their AI assistant. If the safety check loses connection, the assistant continues working anyway. The author only tested these settings on a small number of examples, so the limits might not catch every problem.
Key takeaways
- Cut down long text and remove sensitive details before sending data for an outside safety check.
- Group different types of errors into categories to give the AI assistant clear instructions on fixing them.
- Let the assistant keep working if the safety check fails, but limit how often errors are reported.