LLMs are built on neural networks which are encoding a kind of strategy function through their training.
The strategy in an LLM isn't necessarily that it "thinks" about the specific problem described in your prompt and develops a strategy tailored to that problem, but rather its statistical strategy for cobbing together the tokens of the answer.
From that, it can seem as if it's making a strategy to a problem also. Certainly, the rhetoric that LLMs put out can at times seem very convincing of that. You can't be sure whether that's not just something cribbed out of the terabytes of text, in which discussions of something very similar to your problem have occurred.
The anthropomorphization of llms bother me, we don't need to pretend they are alive and thinking, at best that is marketing, at worst, by training the models to output human sounding conversations we are actively taking away the true potential these models could achieve by being ok with them being "simply a tool".
But pretending that they are intelligent is what brings in the investors, so that is what we are doing. This paper is just furthering that agenda.
This is not true. The key-values of previous tokens encode computation that can be accessed by attention, as mentioned by colah3 here: https://news.ycombinator.com/item?id=43499819
You may find https://transformer-circuits.pub/2021/framework/index.html useful.
The whitepaper you linked is a great one, I was all over it a few years back when we built our first models. It should be recommended reading for anyone interested in CS.
Anthropo language has been woven into AI from the early beginnings.
AI programs were said to have goals, and to plan and hypothesize.
They were given names like "Conniver".
The word "expert system" anthropomorphizes! It's literally saying that some piece of logic programming loaded with a base of rules and facts about medical diagnosis is a medical expert.
LLMs are black boxes, and if anything, interpretability systems show us what the heck is going on inside them. Especially useful when half the world is using these already, and we have no idea how t hey work
That doesn't mean the water itself has strategies, just that you can use water in an implementation of strategy... it's fairly well known at this point that LLMs can be used as part of strategies (see e.g. "agents"), they just don't intrinsically have any.