Brain tunes itself to criticality, maximizing information processing
source.wustl.edu
source.wustl.edu
https://towardsdatascience.com/the-bayesian-brain-hypothesis...
So far it seems that it explains quite a lot of data, and many mind illnesses (e.g. many diseases can be thought as the brain under-correcting or over-correcting for the prediction error).
By under-correcting, the brain is not learning enough on its mistakes, which may lead to delusions of superiority (e.g. being stuck in usual habits, or inability to change one's world-view based on new information). On the other hand, when over-correcting, the world may seem unpredictable, frightening - leading to self-doubt, anxiety and negative thoughts.
Being wrong around 15% of the time might actually be the optimal rate for learning... https://www.independent.co.uk/news/science/failing-study-suc...
I checked Friston again. He now also has this article: https://www.frontiersin.org/articles/10.3389/fncom.2012.0004...
CLE = Conditional Lyapunov Exponents.
"In short, free energy minimization will tend to produce local CLE that fluctuate at near zero values and exhibit self-organized instability or slowing."
I've to study it more what he means with self-organized instability.
My point is that saying that the brain is maximizing harmony is quite reasonable -- and much easier to understand.
Rumelhart, D. E., Smolensky, P., McClelland, J. L., & GE, H. (1986). Schemata and Sequential Thought in PDP Model. PDP, Exploration in the Microstructure of Cognition, The MIT Press, Cambridge, MA, Vol. IIº.
For example, if you run the rubber hand experiment with non-schizophrenic people, even if you don't stroke their hand and the rubber hand at the exact same time (say the timing offset is gaussian with standard deviation sigma), with enough repeated exposures to the stimuli they will recognize the rubber hand as their own. In contrast, if you repeat the same experiment with schizophrenic people, it takes a smaller standard deviation or substantially more trials to have them recognize the rubber hand as their own.
I wish I had the references lying around, but I dug into the literature for this a few years back and found this hypothesis to be surprisingly well supported.
1) Yes. Why couldn't it?
2) No, it requires a certain level of brain complexity.
So when anybody says "we showed X was critical", they actually mean "we plotted a fuzzy cloud of data points on a log-log plot and fitted a line through it", nothing more. But you can fit a line through anything. Even a normal distribution shows up as a line on a log-log plot if your data has a small enough range.
Criticality studies trade on the reputation of physics, where the idea came from, and there it works fantastically. For instance, we can measure critical exponents for liquid/gas phase transitions to three or four significant figures, and even predict those numbers from pure theory. Applications outside of physics usually have barely one significant figure, if they're even measuring power laws at all, and no predictive theory.
From "Akin's Laws of Spacecraft Design", https://spacecraft.ssl.umd.edu/akins_laws.html
John von Neumann was at a talk where the presenter had put up a slide with a cloud of points, and had optimistically drawn a line through the cloud. Von Neumann muttered, "at least they lie on a plane."
https://www.amazon.com/Adventures-Mathematician-S-M-Ulam/dp/...
Only if the sample size is small enough.
What exactly does exhibit criticality? As correctly stated, all kind of phenomena can exhibit power laws.
Avalanches on a sandpile are sized as a power law. Typical example of self-organized criticality (Per Bak).
Back in the day I played with group renormalization theory to prove SOC, but most systems break down if there is loss on a microscopic scale. Intuitively, you need conservation laws on a microscopic scale or the very large events do not happen.
This is unlikely the case in a biological system and I won't expect it to be at a critical state, but only hovering around "an interesting area".
They measure neuron activity and spread of firing actions - if I understand correctly without looking at the paper. And they found that thanks to the sophisticated inhibitor neuron network the whole network balances around this mathematical criticality.
I guess it means if there were less inhibition then entropically there would be too many firing actions leading to a feedback frenzy, which is obviously not that effective for information processing.
And if there were more inhibition then the information wouldn't be able to spread to all the special small parts of the brain, thus making them too specialized.
Maybe.
https://www.biorxiv.org/content/biorxiv/early/2018/12/20/503...
"Avalanches were analyzed in terms of size (S, the number of spikes), and duration (D, time) (Fig 1A), and power law exponents were fit to the two distributions. In critical systems, the exponents of the two distributions can be used to predict the mean avalanche size (<S>) observed at a given duration (i.e. the distributions scale together). When <S> is plotted against avalanche duration, the difference between the empirically derived best-fit exponent and the predicted exponent serves as a compact measure of the deviation from criticality (Deviation from Criticality Coefficient, “DCC”, Fig 1B)."
Recap. There are several matters that are at times conflated.
+ Define an order parameter which "significantly changes" (undergoes a phase transition).
+ Define a control parameter that drives the system through those different regimes.
+ Establish that there is a critical point (not just a "region").
+ Define properties at the critical point that are scale-free (or having "all scales").
+ Define this critical point as an attractor in a dynamical system sense. (The system gets infinitely close to it given enough time.)
They show however that the system has an attractor "near" the phase transition. Moreover, do not really establish that this is a critical phase transition. They establish "nearness" w.r.t. the critical phase transition by fitting two power laws (size vs duration) according to an expected exponent relation (a-1)/(t-1). They see this as a quantitative measure of nearness. In other words, their nearness measure depends on their definition of criticality.
A lot of these studies talk about "near-critical" states. However, that doesn't always make mathematical sense.
PS: If you think, ah, you just have to have the control parameter as an output of a system with as input a value that defines the criticality of the system, yeah, that might work. However, the real interesting systems do not use this type of macroscopic information, but have local processes leading to the same results.
Always being just bellow the activation threshold.
And there are loops that work against these to prevent breakdown, forgetting too much, slowing down too much, etc.
Of course it's not really known yet what these control system are exactly. (At least I'm not aware we have good data and theories about this aspect of the brain.)
The brain has a tightly balanced feedback loop between excitation and inhibition. Too much excitation and the brain gets a seizure (positive feedback). Too little, and you black out.
TBH, it's quite a good example of the principle of balance in the concept of Yin-Yang.
E.g.
* Friston, K. (2009). The Free-Energy Principle: A Rough Guide to the Brain? Trends in cognitive sciences, 13(7), 293-301.
https://www.fil.ion.ucl.ac.uk/~karl/The%20free-energy%20prin...
* Friston, K. (2010) The Free-Energy Principle: A Unified Brain Theory?. Nature Reviews Neuroscience. 11(2): 127.
https://www.fil.ion.ucl.ac.uk/~karl/The%20free-energy%20prin...
* Solms M (2018) "The Hard Problem of Consciousness and the Free Energy Principle." Front. Psychol. 9: 2714. DOI: 10.3389/fpsyg.2018.02714 | PMCID: PMC6363942 | PMID: 30761057
https://www.ncbi.nlm.nih.gov/pmc/articles/PMC6363942/
https://www.frontiersin.org/articles/10.3389/fpsyg.2018.0271...
Thermodynamics and signatures of criticality in a network of neurons
https://www.pnas.org/content/112/37/11508
"The activity of a brain—or even a small region of a brain devoted to a particular task—cannot be just the summed activity of many independent neurons. Here we use methods from statistical physics to describe the collective activity in the retina as it responds to complex inputs such as those encountered in the natural environment. We find that the distribution of messages that the retina sends to the brain is very special, mathematically equivalent to the behavior of a material near a critical point in its phase diagram."
In the early 90s, I used their ftp site at wuarchive.wustl.edu very frequently. It was a reliable source to download open source software like Perl, tcl, trn, gcc, and so on.
I won’t spin mental cycles guessing. I’d rather the authors were explicit.
I'm afraid this may be a scenario where some deeper knowledge is needed to fully appreciate the discussion. You would be well rewarded for putting in the effort though, it's a fascinating notion. I'd recommend starting with the Ising model [0] which is the canonical system exhibiting critical phenomenon.
[0] https://en.wikipedia.org/wiki/Ising_model
edit: if by 'this' you are referring the the wiki article on critical phenomenon, you're definitely missing the larger picture. The examples that wiki article lists aren't metaphorical, they're all essentially 'corollaries' (in a very loose sense) of the same underlying thing. Start with the Ising model.
Sadly no further explanation of this. Exponent-relation sounds like a synonym for power, too.
Well, these ideas are very likely accurate, but so general that they are disconnected from solving practical problems. Fine, the brain is a prediction machine and it optimizes over some program space by annealing doing homeostasis/regulation, and maybe tends to occupy certain kinds of states now and then. This, however, tells us very little what the learning rules for the synaptic weights should be and how we should wire things up. In fact, I believe human-relevant problems are best solved by such a special subregion of program space that one needs pretty specific architectural priors as otherwise search will take too long. These are unlikely to be derived from general concepts, but are more likely evolved, either literally by evolutionary algorithms or by people doing the trial and error. The issue being that general concepts about prediction errors and program spaces know nothing about our specific world. E.g. none of these general concepts predict the usefulness of CNNs. CNNs exploit fairly specialized priors about object translation invariance and locality in image statistics, which are specific computations occurring in our universe when parts of it are perceived by geometric projections of EM rays onto an image plane with sensors. Hinton's capsules go into the right direction exploiting some more priors about spatial reference point invariance, but we need to go deeper. The brain disassembles the world into stable episodic chunks and operates on them, and it manages to backpropagate values through such episodic memories. Currently, no neural architecture does something like this.