OpenAI Announces Goal to Solve Alignment of Superintelligence Within 4 Years
twitter.com
twitter.com
And how would you quantify alignedness? Rating outputs for a given input falls prey to the first problem. Analyzing activations as they trickle through the model is intractable analytically, and training a "polygraph" model on the activations of your network raises more alignment issues (how can you be sure the polygraph isn't lying to you?)
I'm ready to eat my words, but I think perfect alignment is infeasible. The best we can hope to do is curate training data and hope the caged bird won't sing.
it's when it no longer says problematic words like bad ones related to race or gender or hierarchical power relations or different abledness
Personally I'd short them on this. It's doomed.
(I am not an AGI believer, never have been. This is scale/connectionist nonsense dreaming about "moar power" delivering the outcome)
OK so when the AI knocks on your door and plucks you up from the chair it will be the "scale/connectionist nonsense" which has done this and not something that you call as "AGI". That's cool, I guess no one can tell you how to name it.
When the time travellers with wisdom pills knock on YOUR door and force you to eat the "wise up" pill don't say I didn't warn you.
Simply adding more branches and data to a model will not make intelligence.
You're not helping your own line of reasoning arguing that more than one line of reasoning in AI risk is somehow contradictory, because they're not. They are disjoint.
If AI exists it would be dangerous
Even without existing belief in AI, LLM product without oversight is dangerous.
Even being dangerous, Bigger LLM do not automatically make AGI happen.
Yes, social policy which is unregulated and unconstrained based purely on belief of their capabilities would be bad. We've seen this before in expert systems codifying systemic bias (for instance)
The "rather" and "not because" imply some exclusivity of debate which I don't think applies. I won't magically change my mind here because of some rhetorical flourish in a line of argument
(Sorry.. I am just a bit over adversarial debate online and suspicious of direct questions like this.)