Before you dive in
What you’ll find in the original
- A finite DOM action set fits a Jev loop, but long-context browser quality and reasoning still need benchmarks.
- Many tool calls cannot be replaced because their arguments—such as exact edit lines—must be generated rather than selected from a fixed menu.
- Turn-based games, robotics, and autonomous systems offer bounded decision trees where latency may matter, though an LLM can still perform some of these tasks.
Worth knowing
These are one day of exploratory notes. The author had not seen benchmark evidence and describes the poker and Pokémon systems as demos still in progress.