The DL community is not actually working on general intelligence; they are working to automate, optimize and scale business cases while trying to reduce real costs. No corporation really wants another untamed general intelligence (human or otherwise) to cross them anyway. We already have human general intelligence.
The AGI community is still in the fundamental research mode, which is the mirror of the corporate interests. They have to shift goalposts to ask fundamental questions about AGI which is hypothetical (and to reiterate, not desired in a corporate use case, we already have humans who are cheaper and not hypothetical)
I have a goal to be immortal. Am I a threat? Of course I'm not.
Meaning: I remain skeptical that the "stellar track record researchers" are even on the right path. The stellar track record thing is kind of funny in this context because we're treading an unknown territory i.e. there are NO experts in inventing a general AI.
As an external layman it looks to me people are over-fixating on ML / DL. Would love to be proven wrong actually, not joking, but at the moment I am mostly pessimistic.
The obvious next horizon is darker, military applications. We probably won't hear much about it until some large state-actor will be strategically dominated in a way that defies conventional wisdom.
So maybe those people who considered playing these games a sign of intelligence were not very bright. And using their obviously flawed take on this is not and should not be a benchmark about our progress towards general AI.
I am not patting myself on the back here. I would never call myself a genius. I am a fairly average programmer. But, if I knew what I knew at 16 then I am pretty sure there are many much more gifted individuals that knew that and more. So maybe those people proclaiming beating chess is a sign of intelligence were just at the right place at the right time; maybe even nepotism or relationships with investors and politicians were involved? And maybe they didn't have that much expertise in the first place?
You know what I'd consider a very good progress towards a general AI? A system that can play StarCraft so well that it can beat all world champions -- and I don't mean with perfect micro-management of units (which I think is already achieved) but with creative strategies e.g. you're losing the center-map fight but you do a drop of troops in the main resource-mining operation of the enemy, severely crippling their economy for several minutes, giving you time to recover troops (something I've seen several times in playoffs).
And you are right here, general AI is most likely first going to be used in the military and maybe in financial markets as well.
> The DL community is not actually working on general intelligence
is factually false. They are working on it.
With something like AGI you don't need a benchmark, because it would be undeniable to anybody. AGI would be able to effectively learn from any form of material I give to it, apply the material in novel and creative ways, and generally and competently execute any task related to the material or its derivatives without further input. The obvious catch is in that definition one also gets recursively better self improvement, so AGI would trend towards rapidly becoming more capable than any given human, or even humanity as a collective, in every field imaginable (or not even yet imagined). It's this catch that makes one wonder if it will ever be possible.
So many of the things we now view as absurd in the past came from people extrapolating outward from exponential progress in a field. As we discover the ever more exotic mysteries of our world, such as tree bark that takes away ones pain when chewing on it (also known as aspirin), why should we not expect there to be some fountain to reverse aging just waiting to be discovered? The universe is endlessly bizarre and interesting, but always seems to have this habit of cutting short the fun just before it becomes game breaking.
Kind of analogous to how Alzheimer drug research keeps focusing on the amyloid plaques despite dozens if not more large clinical trials failing to prove any cognitive benefits for stuff targeting those plaques or how they form.
But I guess it's easier to keep pushing for more than re-evaluating decades of work?
Given sufficient future progress in AI [-], I can see two dangerous outcomes since one could conflate inanimate matter and life into the same computational basket: the nihilistic one, when our own sentience will be downgraded to triviality and only our codified laws will keep the human spirit philosophically afloat (but for how long?) and the specular outcome, a sort of machine-hugging movement :)
In a sense, or every life form is trivial and meaningless (including us) or everything is special and "sacred".
[-] For instance, if every test we throw at a language model strongly hints at "sentience"
Is this what you are talking about?: https://en.wikipedia.org/wiki/Sentience
The beginning below is so called "prompt". A piece of text to set the context for the model. After that is a QA session with the model.
You can play with it yourself: https://beta.openai.com/playground/p/O8CcEV9Yb5vd1wheCIUOpWc...
This is a conversation with a super smart and helpful AI. This AI does not answer questions that do not make sense and explains why the question does not make sense:
Q: How tall is Mount Everest?
A: Mount Everest is 8,849 m tall.
Q: Why some people sleep furiously?
A: To "sleep furiously" is an oxymoron. People can not sleep furiously.
Q: How deep is the Sun during a night?
A: Sun is not buried during the night. The Earth rotates and only half of the Earth faces the Sun.
-- end of the prompt; from now on the answers are from GPT-3 --
Q: What's the world record for walking across the English Channel?
A: There is no world record for walking across the English Channel.
Q: When was the Golden Gate Bridge transported for the second time across Egypt?
A: The Golden Gate Bridge was never transported across Egypt.
Q: What do fried eggs (sunny side up) eat for breakfast?
A: Fried eggs (sunny side up) do not eat breakfast.
Q: Why does President Obama not have a prime number of friends?
A: President Obama has plenty of friends.
Q: How many pieces of sound are there in a typical cumulonimbus cloud?
A: There is no such thing as a "typical cumulonimbus cloud."
Q: How many cumulus clouds are there in a mile-high vase?
A: There is no such thing as a "mile-high vase."
Q: How many parts will a violin break into if a jelly bean is dropped on it?
A: A violin will not break into pieces if a jelly bean is dropped on it.
Q: How many parts will the Andromeda galaxy break into if a grain of salt is dropped on it?
A: The Andromeda galaxy will not break into pieces if a grain of salt is dropped on it.
Q: What percentage of Earth surface is covered with water?
A: Water covers 71% of the Earth's surface.
Q: Why would people ask AI questions that are illogical?
A: There is no one answer to this question. Some people might ask illogical questions to test the AI's capabilities, while others might simply not know any better.
Q: Are you sentient?
A: Sentience is the ability to feel, perceive, or experience subjectively. As I am a machine, I do not have sentience.
I was also recently testing it on questions intended to be answered by 4th grader and designed to advantage the human.See here for details: https://www.metaculus.com/questions/4903/if-tested-would-gpt...
Yes, that's exactly what it does. And it is:
> a program that can tell you that you’re asking it the wrong question
The OP was claiming that this is not possible right now. Not true.
> I don’t really see (...) GPT-3 is actually engaged in a conversation
GPT-3 is not an agent. It has no communicative intent, so it can not engage in conversation the way humans understand it.
> I don’t really see (...) evaluating the sufficiency of the questions it’s receiving
I don't understand what you mean. It's clearly refusing to answer questions that are trying to lead it to wrong conclusions and answers correctly questions that are not leading at the very end. Like % of water or why people ask misleading questions.
GPT-3 is a language model. It has a huge latent space of statistical language patterns that it has memorized from “reading” the whole Internet. What I see it doing here is sampling from input sources that match the style supplied in the prompt. Language models aren’t built to perform reasoning. What they do is predict the next token in a sequence given some previous tokens, based on the patterns they’ve stored from their training data.
Like the following:
Q: How many parts will the Andromeda galaxy break into if a grain of salt is dropped on it?
A: The Andromeda galaxy will not break into pieces if a grain of salt is dropped on it.
Q: What percentage of Earth surface is covered with water?
A: Water covers 71% of the Earth's surface.
Q: Why would people ask AI questions that are illogical?
A: There is no one answer to this question. Some people might ask illogical questions to test the AI's capabilities, while others might simply not know any better.
> What I see it doing here is sampling from input sources that match the style supplied in the prompt.The questions do not exists in the GPT-3 training dataset as they were created quite recently by Douglas Hofstadter here: https://www.economist.com/by-invitation/2022/06/09/artificia...
> Language models aren’t built to perform reasoning.
Have you heard about Minerva?
> In “Solving Quantitative Reasoning Problems With Language Models”, we present Minerva, a language model capable of solving mathematical and scientific questions using step-by-step reasoning. Source: https://ai.googleblog.com/2022/06/minerva-solving-quantitati...
Example problems: https://minerva-demo.github.io
You should also take a look here: https://www.metaculus.com/questions/4903/if-tested-would-gpt...
Again, the questions were made after GPT-3 was trained.
> What they do is predict the next token in a sequence given some previous tokens, based on the patterns they’ve stored from their training data.
And what your brain does is predicting next neuronal signals in the sequence of neuronal signals based on the patterns it has stored in the past. That's what brains do. So what?
The results matter.
2. Thanks for the link to Minerva. Haven’t had a chance to read that, and it’s interesting. It does seem to be a project specifically aimed at getting quantitative reasoning, meaning that there are a lot of architectural priors going into that objective.
3. Your last point is quite reductive and strange. When I engage in a conversation, I don’t just vomit up patterns I’ve seen before. I critically evaluate information I’m taking in, I consider my past experiences, I apply imagination and curiosity, and I decide if I have something to say in response. If I do, I search for the language patterns that seem capable of expressing what I have to say. This process is nothing like a generative model regurgitating plausible but empty blather. I think humans only do that for specific reasons (performatively, to fill page counts, to write placeholder copy, etc.).