HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by mritchie712 | Hacker News Reader
Parent
Full thread
mritchie712
·
in 2024, yes.
what AI are you using where this still works?
View on HN
wat10000
·
I haven’t tried it in a while, but LLMs inherently don’t distinguish between authorized and unauthorized instructions. I’m sure it can be improved but I’m skeptical of any claim that it’s not a problem at all.
Reply on news.ycombinator.com