Rethinking artificial intelligence
web.mit.edu
web.mit.edu
The article mentions revisiting fundamental assumptions, but doesn't mention a single specific thing that the team will do differently. I've studied AI on my own for some time, and the biggest problem (as far as I can tell) is that researchers can't agree on what "intelligence" means. AI research for the past fifty years hasn't been about building an artificial intelligence machine, it's been about precisely defining what "intelligence" means. Every time there was a breakthrough, after a bit of hype people realized that the program is actually pretty dumb, is a testament to the intelligence of the programmer, not the machine, and that the bar for "intelligence" simply shifts a bit higher up.
So, what exactly is this team doing differently? How do they define "intelligence", and what do they intend to build?
Without a scope definition they can't budget their funds, their time, and their human resources. Projects like these usually result in a waste of money with nothing to show for it. Of course if they try to define what they're building, it will be intimately linked to the definition of "intelligence" (assuming they claim they're building an intelligent machine). Then someone will come along and propose a counterexample that demonstrates how the machine likely isn't intelligent at all, and cannot perform well on some problem where humans do spectacularly, thereby shifting the team's scope and definition. And so, they'll be back to square one.
http://funnylogic.com/times/txt/2009-11-infinite-curiosity-l...
And for what it's worth, the definition of intelligence has been on my mind a lot.
[1] http://evans-experientialism.freewebspace.com/putnam05.htm
Refining this definition would involve researching all relevant objective tests that can be measured accurately, then forming some sort of matrix of comparison.
Initially I would pursue the testing on a single interface - using a black box approach with a text interface, then later extending it to other fields (such as movement / navigation, visual, aural etc)
If the AI can function on par with the average control, then it could be said that it is averagely intelligent based on the test matrix.
NOTE: This is just a top-of-the-head idea, I know it is a lot more complex than I make out (how do you define learning?) but it seems a logical starting point to me. Use current tests and results - just be careful not to feed the AI the original data.
1. You have a problem domain P to which the No Free Lunch Theorem applies, at least approximately. Thus, there are hard bounds on how well one algorithm can do compared to any other according to some metric.
2. Information is produced when an agent can perform significantly better than is algorithmically possible. One such metric is the compressibility of its search history.
Why is this my definition? Well, it is linked to our intuitive notion of learning and intelligence. As informally described by Hofstadter, it is an inherent ability to "step outside of the system." This means, at some time t I am behaving according to some rule set r, but at time t+c I understand the rule set and can reason about r instead of just being subject to r.
One specific result of being able to reason about a rule set is that I can take some well formed sentence, realize it can't be generated by the rule set, and use a simpler rule set to generate it. When framed in terms of Kolmogrov complexity, I'm exhibiting a general compression capability, which implies a general (though not total) capability to solve the halting problem.
Since a problem domain with structure can be compressed, this relates to my first example in that if an agent has a generally much more compressible history than mathematically expected in a (almost) No Free Lunch domain, it is exhibiting the ability to step outside of its environment's rules, reason about them, and thus compress them.
So, you can see that my definition of intelligence as the ability to create information specifies an unambiguous and measurable capability, which also happens to specify something algorithms are mathematically incapable of doing. Thus, I've have both defined intelligence and disproven the logical possibility of such an AI in one fell swoop.
BTW, this is not an original thought of mine. It is a direct result of intelligent design theory, the progeny of the absolutely brilliant William Dembski.
Some of the assumptions might be:
- That thought occurs inside the head. To what extent is thought a process distributed across multiple individuals?
- That you can make a sharp delineation between reasoning and perception. Many AI systems assume that tasks like object recognition can be completely separated from the rest of the system as it's own module. Neuroscience, on the other hand, suggests that perception, memory and reasoning are all tightly integrated together.
- That the brain can be modeled as an electrochemical system. The mainstream view is that quantum effects play no significant role, and that Penrose & Hameroff are wrong.
Language as a portable method of conveying thought is a great subject of research. But, attempting to attack a human language in all it's glory is a huge pitfall. Nouns, verbs, additives, adverbs, intonation, tense, and a 29 years of knowledge and I frequently have no idea what someone is saying. But, linking the word ball, with the concept ball, with the sensory perception of ball might be possible today. Link that to some simple nouns like roll, toss, catch, etc and we can start a meaningful integration of language and actions.
"Part of this difficulty comes from the very nature of the human mind, evolved over billions of years as a complex mix of different functions and systems. “The pieces are very disparate; they’re not necessarily built in a compatible way,” Gershenfeld says. “There’s a similar pattern in AI research. There are lots of pieces that work well to solve some particular problem, and people have tried to fit everything into one of these.” Instead, he says, what’s needed are ways to “make systems made up of lots of pieces” that work together like the different elements of the mind. “Instead of searching for silver bullets, we’re looking at a range of models, trying to integrate them and aggregate them,” he says."
It makes sense to create a single intelligent entity by combining different types of technology that works well on specific topic. We have many good techniques that do things better than human being, why try to mimic ourselves in general (the neuron way) rather than utilize a better solution to each problem?
MIT is a zombie: a corpse with delusions of youthful vigor.
(This question also goes out to anyone else that has something to add.)
http://en.wikipedia.org/wiki/AI_winter
There are various explanations as to the cause of death: the end of the Cold War; the humbling of the mega-monopolies which funded "blue sky" research (mainly AT&T); a general loss of faith resulting from a decades-long lack of progress. Take your pick.
In fact, the entire field of computer science has been stagnant for a while, shiny gadgets to please people with 5-minute attention spans notwithstanding:
http://www.eng.uwaterloo.ca/~ejones/writing/systemsresearch....
Bureaucrats have replaced thinkers:
http://unqualified-reservations.blogspot.com/2007/08/whats-w...
My advice: study physics or chemistry.
Maybe things would improve if more people volunteered time toward computer science research; I can envision something like the GNU Project, but for research rather than engineering, with proper administration, goals, tasks, and resources, to help establish purpose and vision, and attract volunteers to an overarching common goal.
Or maybe even establish something like Y Combinator for CS research, a small-scale NSF if you will. Give a small group of innovative folks a few months of funding to create something new, regardless of if it has near-term business viability.
As a person moving from physics to CS, I personally find computing to be a very interesting place right now.
Where, then, is the desktop operating system not built of recycled crud? Where can I see a conceptually original system created after the 1980s?
> the field itself is alive and well
I disagree entirely. It is a zombie, maintaining the illusion of life where there is none.
Hell, UNIX still lives, and this proves that systems research is dead:
http://www.art.net/~hopkins/Don/unix-haters/handbook.html
Why is my desktop computer running software crippled by the conceptual limitations of 1970s hardware? Why are there "files" on my disk? Where is my single, orthogonally-persistent address space? Why is my data locked up in "applications"? Why must I write programs in ASCII text files, and plod through core dumps and stack traces? Why can't I repair and resume a crashed program?
With the exception of MS, most of the systems research I was referring to is not used in (or intended for) desktop operating systems.
You're willing to look past our massive advances in optimization, control systems, search technology, computer vision, etc etc, and pretend they don't exist...
... because desktop OSes still suck?
Ok, I'll bite. What advances? I'm talking about real change, not incremental bug-stomping by plodders.
I just spent three weeks (class project) implementing a new algorithm to find the minimum cut of a directed planar graph in O(nlgn) time. The algorithm is actually quite elegant:
http://www-cvpr.iai.uni-bonn.de/pub/pub/schmidt_et_al_cvpr09...
This came out of a Ph.D. thesis written in 2008, and was applied to some computer vision problems in the paper I linked above. This isn't a minor speedup or optimization... it yields asymptotically faster results.
My vision professor is fairly young, and recently did his own Ph.D. work on Shape From Shading. This is the problem of recovering 3D shape from a single image (no stereo or video). His solution used Loopy Belief Propagation and some clever probability priors to achieve solutions that were orders of magnitude better than previous work. In fact, his solution is so good that rendering the resulting 3D estimate is identical (to the naked eye) to the original (although the actual underlying shape varies, since there are multiple shapes that can all appear the same given the lightning conditions and viewing angles).
There is also a ton of interesting progress in the last two decades making functional languages practical in terms of speed (and hence useful). My advisor did his Ph.D. in this area.
The entirety of CS is not evidenced by the current state of operating systems. In fact, I'd argue that OS research at this point has less to do with computation than it does with human-computer interaction, which seems like it requires more research about humans than computers.
but no true Scotsman would do such a thing!
MS Bob.
I'm serious. And the example I offered shows why a conceptually original system is not necessarily a good thing.
There is nothing new about straightjacketing computing into an "everyday household objects" metaphor. As in, "the desktop," for instance. It is a very old idea which simply refuses to die.
And here is what the late Erik Naggum had to say about "user friendliness," the ancient disease which gave us MS BOB:
"the clumsiness of people who have to engage their brain at every step is unbearably painful to watch, at least to me, and that's what the novice-friendly software makes people do, because there's no elegance in them, it's just a mass of features to be learned by rote."
"The Novice has been the focus of an alarming amount of attention in the computer field. It is not just that the preferred user is unskilled, it is that the whole field in its application rewards novices and punishes experts. What you learn today will be useless a few years hence, so why bother to study and know /anything/ well?"
Right now I'm looking for something like a fractal. Where we get large constructs out of a small equation.
http://www.youtube.com/watch?v=HmLS2WXZQxU
Those are obvious rare exceptions though, and evolution never built something as fast as a jet etc. and plenty of other stuff that would need foresight. Not sure if there is some sort of somewhat simple engineering principle behind brains/consciousness that we just havn't figured out yet, but if we knew it, maybe we could engineer it like a jet instead of using evolution. A jet seems extremely simple relative to just about any of nature's locomotion creation though.
For the comment below, interesting fractal 'amplification' idea. I've briefly thought about positive feedback loops possibly building something interesting. But mostly just as some vague analogy since I don't have enough programming knowledge to experiment much.