JevMade Sign in
← Back to guides

JevMade field notes / Routing design walkthrough

Send customer requests to code, AI or a person

Akash Mohapatra explains a customer-service router that uses Jev to identify a request and its difficulty, then lets code choose a lookup, a writing model or human help.

Original by Akash MohapatraAgent workflows

Listen to this guide

JevMade’s plain-English explanation

0:00 /

Our summary

A customer asking where an order is may only need a database lookup, while a difficult complaint may need a person. Akash Mohapatra explains how Jev can sort incoming messages before a more expensive writing model is called. The aim is to avoid paying for writing when no new explanation is needed.

The example asks two separate questions in one call: what the customer wants and how difficult the request seems. Code sends order checks to a lookup and product or return questions to specialist writing models. Complaints go to a person when difficulty is high or its estimate is unclear.

The suggested cutoffs are examples, not tested promises of correct routing. A confident answer can still be wrong, and a failed request to the service needs its own backup route. Savings depend on how many requests avoid expensive handlers, because every message still pays to be sorted. Simple rules may make this sorting model unnecessary.

Key takeaways

  1. Keep the customer's goal and the difficulty of helping them as separate judgments. A clear complaint can still be too difficult to automate.
  2. Plan separately for an uncertain answer, a difficult case and a call that fails. A confidence check cannot handle an answer that never arrives.
  3. Measure mistakes and the share of requests that avoid expensive handlers before promising savings. Keep exact lookups and arithmetic in code.

The complete article and its question example were read, not run. It explains how requests are sent to handlers, but does not supply a complete running program. Its opening attaches confidence to every answer type; later it correctly says Noul, the yes-or-no probability, has no separate confidence value. Tokens are the units used to count and bill text. The article's limits of 250,000 tokens per second and 1,200 requests per minute differ from official limits checked October 5, 2026: 100,000 tokens per second and 80 requests per second, both subject to change. Quoted cost comparisons are not independent tests of this router.

Creuto · Original published

Read the original guide Opens the author’s site in a new tab.

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.