JevMade Sign in
← Back to guides

JevMade field notes / Worked Java example

Choose an answering model without hiding the first judgment

Seif Ibrahim builds a Java example where Jev judges a request's difficulty, code selects an answering model, and a separate check compares an existing answer with reference material.

Original by Seif IbrahimIntegrations

Listen to this guide

JevMade’s plain-English explanation

0:00 /

Our summary

Not every question needs the same answering model, but choosing a cheaper one can produce an inadequate answer. Seif Ibrahim shows a Java application that asks Jev to judge a request's difficulty. The application then uses its own rules to choose which configured OpenAI model should write the reply.

When Jev's confidence falls below 0.70, code chooses the higher difficulty category of the prediction and a configured backup. It keeps the original prediction and probabilities visible. A separate service checks whether an existing answer is supported by supplied reference material and includes the details needed to answer the question.

These cutoffs are example policies, not proven measures of answer quality. A failed Jev call stops the request rather than using the low-confidence backup. Requests go to hosted Jev, and generating replies also sends them to OpenAI. The separate answer check reports a result; it does not arrange human review or repair.

Key takeaways

  1. Let AI judge difficulty, but keep the rule that selects a provider model in ordinary code.
  2. Preserve the prediction even when a backup changes the selected route. Exactly 0.70 meets this example's confidence cutoff.
  3. Test new prompts and count the cost of acceptable answers. Tuning examples and hypothetical savings do not establish production savings or equal answer quality.

The article and pinned implementation were inspected, not executed. Recorded September 24 runs and four new prompts are small author-reported probes. The answer check uses grounding of at least 0.85, completeness of at least 1.5 out of 2 and confidence of at least 0.70; none is established as calibrated for this task. Support in a reference does not prove the reference true. No reliable article revision date was found.

Seif Ibrahim's Blog · Original published

Read the original guide Opens the author’s site in a new tab.

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.