Running agents in a sandbox or VM is the wrong pattern
funky.dev
funky.dev
My current understanding of an agent session is that the entire session (modulo compaction, reordering tools, etc) is sent to the stateless inference engine every time inference happens. The quoted section doesn't seem workable unless "read the tail of the log" actually means "read the entire session log".
There are ways to make that work but workers being completely stateless makes it hard.
You are right that my writing is misleading. To clarify that, I mean the worker only needs the tail of the log to decide what to do next: inference or execute tools.
There are many companies working on personal agents, or in a broader term, we can call them the interactive agents.
However, just like the internet, not all computational work is interactive. Batch processing is also very important. This is the motivation for me to write this blog and build this project.
For example, one of our customers builds a swarm of agents that simulate customer behavior to test the products before launch. Such a task can be run in batch and doesn't require a human to sit with the agents. (Therefore, non-interactive)