Does this already exist? If not, would the benefits be lower than I think, or would the costs be higher than I think?
Does this already exist? If not, would the benefits be lower than I think, or would the costs be higher than I think?
Just the other day someone used Claude to write a script to configure a server. It left a port open and the server was hacked hours later and used to attack other servers. Hetzner almost banned the hosting account.
Webwright is a front-end shell that presents to me; I'm suggesting a back-end shell that presents to Claude.
It doesn't appear that Webwright enables tool-use. In other words, there's no task-oriented feedback loop between AI-provided shell commands and the results of those shell commands. Please correct me if that's not right.
This _seems_ more like a normal user so clearly could not do anything nefarious. /s