What it does
It was trained on 12,913 questions spanning six tasks; accuracy is modest, but reported confidence closely tracks observed correctness.
Benchmarks & research
system-one-gemma adds a scoring head to Gemma 3 270M for calibrated decisions without text generation.
Only you can see your notes.
Screenshot unavailable. Open the experiment ↗
It was trained on 12,913 questions spanning six tasks; accuracy is modest, but reported confidence closely tracks observed correctness.
Your developer can use this tool to sort incoming customer messages. You give the app a message and a list of departments. The tool does not write a reply. Instead, it assigns a percentage score to each department. This helps the app decide if a message belongs in technical support or billing.
A developer can start by testing the included code. You must first agree to the maker's rules on their website. Then, the developer uses an access key to download the files. The ready-made version restricts commercial work. Your developer must train the tool on your own data to remove this limit.