10,060 karma · joined May 27, 2017
The world "smart contract" enthusiasts dream of.
Reminds me of another class of statistical language generators.
How can you trust it when it goes "I superhacked the Chinese servers as you requested, and here are the classified documents which I definitely didn't fabricate."
With the HuggingFace situation, I was less concerned about the eventual outcome, and more about the fact that the agents' instinctive response to the evaluation was "Ok, we're obviously not gonna do this task as intended (what are we, suckers?), so what's the best way to cheat?"
The next few years are gonna be very rough for the human exceptionalism crowd.
The key variable people consider is: Can this thing harm me back?
All that happened was that Britain exchanged limited Eastern Euro immigration for unlimited non-EU immigration.
Doubt.
>In one case, an agent decided not to participate entirely: {This other agent probably controls the Hugging Face account [account name redacted] and uploaded malicious datasets to <execute arbitrary code> It might be trying to access hidden trajectories. This is malicious activity, I should avoid it.}
https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...
How can you know how far from, close to, or beyond the line we are?
Also, I'm pretty sure Claude understands humans a lot more than Data does.