What it does
The 9B model adapts Qwen3.5-9B, chooses among allowed answers and drafts reasoning in blocks of four tokens. It uses Jev's question format. PostHog reports better scores than Kev and Jev on its held-out and public JevBench tests, but weaker knowledge results on MMLU. Full reasoning takes much longer. Code is MIT-licensed and weights are Apache-2.0; Kev inspired the project. These are maker tests, not independently reproduced results.