Before you press play
What you’ll find in the video
- Theo separates a guaranteed output format from the probabilistic—and potentially wrong—decision inside it.
- The classification example uses confidence thresholds to filter chat threads rather than reading generated explanations.
- Theo argues against using the model for difficult judging or context compaction; this is his critical assessment, not a comprehensive capability benchmark.
Worth knowing
Gemini-assisted video/transcript review. Schema adherence ensures valid output shapes, not correct reasoning; Jev lacks visual inputs, has a 32k context limit, and its speed and cost figures are vendor-reported rather than independently benchmarked.