JevMade Sign in
← Back to guides

JevMade field notes / Toolkit lessons

Give an AI assistant smaller questions to judge

Vexjoy describes an assistant toolkit that uses code for exact checks, Jev for bounded judgments and larger language models when new writing is needed.

Original by VexjoyAgent workflows

Listen to this guide

JevMade’s plain-English explanation

0:00 /

Our summary

An AI assistant may need to choose a helper or decide whether a proposed change fits the request. Vexjoy describes replacing some written AI judgments with Jev's set answers and probabilities. This makes the decisions easier for software to use, without proving that the judgments are correct.

The toolkit first uses code to search, count and enforce fixed rules. Jev then answers smaller questions about supplied evidence, such as whether a change adds unrequested behavior. In routing, one pass narrows the available helpers and another checks the smaller list. A writing model produces new text when the task needs it.

Vexjoy has not established that the probabilities deserve trust on these tasks. Reported routing comparisons used successive example sets, and shortening evidence can hide important details. The inspected code sends judgment material to hosted providers; secret-removal checks do not guarantee privacy. Model judgments should not replace permissions or firm safety rules.

Key takeaways

  1. Split a broad question into specific behaviors, and include examples of what should and should not count.
  2. Use code for exact checks and enforce actions separately from the model's judgment. A recommendation is not the repair itself.
  3. Check probabilities against labeled outcomes from your own work. Removing an untested cutoff helped this router, but does not show that every cutoff should disappear.

The page and metadata say September 17, 2026, but the body includes a usage window ending September 19 and later analysis. The displayed publication date is retained; the revision date is unknown. Toolkit figures are author reports, not reproduced benchmarks. Its draft check followed an editor's critique and was not independent review. The repository's MIT license does not establish permission to narrate the blog article.

Vexjoy · Original published

Read the original guide Opens the author’s site in a new tab.

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.