OpenAI promised to make its AI safe. Employees say it 'failed' its first test
washingtonpost.com
washingtonpost.com
Now that the 4o model have been out in the wild for 2 months, have there been any claims of serious safety failures? The article doesn't seem to imply any such thing.
Meanwhile in the real world, regular politicians have reacted surprisingly quick to updating laws for new relevant loop holes: i.e. non-consented deep fake porn generation.
Current AI is nowhere near anything resembling an all-consuming AGI monster. The reality is so far from this that it's laughable and the uses of current AI are (except in terms of possible scale at which visual and text sludge can be produced) not much different from the kind of human-created spam and visual sludge made until recently mostly by humans, many of them minimally paid third world content mill writers.
I'd love to read a specifically enumerated list of other real dangers.
A16z and YC are vehemently against the CA and DC regulation right now.
Out of LeCun, Bengio, and Hinton, the one with $10m on the line says not to worry and the professor and resigned-Google-to-warn-you are saying hmm this is very serious…
Regulations can be good, but it would first be nice if they were applied with at least a modicum of rationality and fairness.
I've never heard of anything like this from AI ethicists though, the language they use tends to be more in alignment with HR people or priests in a moral panic than engineers, and is usually either completely fantastical doomsdaying, or all or nothing thinking, often implying the only way to mitigate the chatbot saying something naughty is to lobotomize it entirely instead of some external control measure, such as limiting the audience or presenting a disclaimer based on sentiment analysis for example. It's not surprising to me that they have lost credibility here, I certainly struggle to take them seriously.
I guess with AI taking over a reasonable number of dumb jobs, the next iteration is going to be AI ethicists.
Neither "alignment" fine-tuning nor output filters are likely to be 100% effective, and a single failure can be disastrous.
I don't know how true this is, but the idea that a commercial entity in the modern era would prioritize public safety over commercial interests is pretty laughable. (thinking about Boeing and Waymo most recently.)