A 'Brief' History of Neural Nets and Deep Learning
andreykurenkov.com
andreykurenkov.com
Perhaps some independent validation, but I was coincidentally having this conversation the other day with a relatively well known computer vision researcher, about why it seems like the idea of neural nets has floundered for decades and suddenly it's the hot topic, and we're seeing massively improved results.
His answers, summarized, are that:
1- Big data is making possible the kind of training we could never do before.
2- Having big data & big compute has made some training breakthroughs that allowed the depth to increase dramatically. The number of layers was implicitly limited until recently because anything deep couldn't be practically trained.
3- The activation function has very commonly in the past been an S-curve, and some of the newer better results are using a linear function that is clamped on the low end at zero, but not clamped on top.
All really interesting to me. This is making me want to implement and play with neural nets!
Of course, now the big question: if we have a neural net big enough, and it works, can we simulate a human brain? (Apparently, according to my AI researcher friend, we're not there yet with the foundational building blocks. He mentioned researchers have tried simulating a life-form known to have a small number of neurons, like a thousand, and they can't get it to work yet.)
I'm not suggesting otherwise, and what you said might be true, I don't know. But honest question: how can we know it's true before we can do either? How can we suggest neural nets are a good building block of super human AI, if we can't verify we can do regular human thinking, or sub-human thinking?
Because the neural net is modeled after (simplified) brain neurons in the first place, it seems reasonable to suggest that if we can't simulate a simple brain, then we can't validate that the model is correct, right? It might be very productive, and it might 'work' in some sense, but we don't know whether it can truly act as the lego brick of brain building material until we can build a functional brain.
Your observation is true, and worth considering. Machines can travel faster than humans, but that's different in part because we didn't start trying to simulate the human foot, right? The computational equivalent would be that machines can multiply a lot of numbers much faster than a human. That is super-human calculation, and it is mechanical thinking, of a sort, but most people wouldn't say it counts as "AI", perhaps the same way that most people wouldn't say a car counts as human running, even if it is faster.
Neural nets are being used to classify images, find objects, identify people's faces in a crowd or in difficult to see situations. But, they currently can't tell you to stop classifying images because you're asking the wrong question, or interrupt you to say you're looking great today.
How is this related to the neural networks we are discussing? (deep ANNs?) As far as I know, the ANN topic in machine learning, apart from its origins, is completely unrelated to the simulation of biological neural network models: https://en.wikipedia.org/wiki/Nervous_system_network_models
They are often confused because of the name and because ANNs did begin as extremely simplified models of biological neural networks, but the machine learning concept that is setting all the records in vision / speech recognition serves no purpose as a model of a biological neural network.
I think you are overstating this. There are definite links between the two, even if ANNs end up different because of the tools we have available to us.
CNNs were explicitly designed to mimic the behavior of the visual cortex.
Most of Geoff Hinton's career has been built around thinking very hard about biological computation.
One of the major criticisms of back-propegation in neural networks is that it is biologicaly implausible.
https://en.m.wikipedia.org/wiki/Convolutional_neural_network...
Yes. I think I was already fully agreeing with you, and suggesting the same but in softer language. Maybe you meant that for the parent or more generally the thread, but I think that I should add something: I may have unintentionally conflated the discussion on neural nets and the biological research my friend was talking about.
My limited understanding of what he said - and I'll go ask him for a reference - is that we (as in scientists somewhere, not me personally) have come to a more or less clear and complete understanding of the chemical and electrical processes in the neural functioning of these simple life-forms, from top to bottom. (I don't know that's true, but that's what I think I heard.) They then tried to put together a complete simulation of this simple lifeform's neurons. This simulation, as I understand it, is not a neural network per se, but something trying to be much closer to a biological simulation. And when they turn it on, apparently, it doesn't work.
That story, if true, tends to confirm what you said; we're missing something in our understanding of biological neurons. Which, I think, lets me say more confidently that we can't suggest we have the building blocks for AI yet.
And as such, I don't know, but I do currently believe that we don't have the computational foundation for AI yet. My friend, who's done more AI research than I have, and who's probably smarter than me, said he thinks we probably have everything we need for the logic part, and the only thing missing is enough computation and enough data to simulate the amount of input a human gets. I was surprised by this response and pressed him on it. "It's turtles all the way down."
There are many possible neural networks. Most of them are not human brains.
302 to be precise - https://en.wikipedia.org/wiki/OpenWorm
Thanks for linking to the old NYT article on Frank Rosenblatt's work. One can see how researchers of the time were irked by delirious press releases when the credit-assignment problem for multilayer nets had not been addressed.
(We managed to mostly address the credit-assignment problem for multilayer nets...but the delirious press release problem remains unsolved.)
Incidentally, it's "Seymour Papert", not "Paper" (appears twice).
I knew there was a fight between Minsky and Grossberg at that time. Perhaps there have been other reasons that are not so well known that led to an AI winter. Have these winters ever be quantified though?
For the article, you can find it in here: http://msrvideo.vo.msecnd.net/rmcvideos/258318/dl/258318.pdf
contains a collection of papers by nn luminaries including rumelhart, Hinton etc. Very highly recommended.