Specialists require nuanced language when building up a body of research, in order to map out the topic and better communicate with one another.
Specialists require nuanced language when building up a body of research, in order to map out the topic and better communicate with one another.
However, currently these attacks are all some variation on "ignore previous instructions", and taking the language of fields where the level of sophistication is much higher, looks a bit pretentious
In traditional application security there are security bugs that can be mitigated. That's what makes LLM security so infuriatingly difficult: we don't know how to fix these problems!
We're trying to build systems on top of a fundamental flaw - a system that combines instructions with untrusted input and is increasingly being given tools that allow it to take actions on the input it has been exposed to.