We've realized for some time that not only are language model technologies going to dominate our future, but that they require massive compute to run. We've also been talking about Apple - their silence, but also their hardware capabilities to run LLMs on consumer devices. Every other player (OpenAI, Facebook, Microsoft, etc) is faced with a billion dollar datacenter investment that still provides a worse experience (latency & privacy).
Apple has a huge advantage because they get to pay nothing for a superior compute infrastructure that comes with a better privacy and latency story. As soon as they release something, it's going to be good, and it's going to be free, because it runs fast on consumer hardware and consumer electricity. A billion-dollar datacenter that's upgraded every N years can't compete with that. Microsoft knows this, and they're trying to get out in front of whatever Apple comes up with. That's why they're telling manufacturers to install these "NPU" things - it's a direct equivalent to the iPhone/iPad/Mac compute capabilities.
So no, Microsoft probably knew this would be the reaction. But it's better to get dragged through the mud for lack of trust than for being left behind the cutting edge. In the former case you're just a slimeball; in the latter case you're a moronic slimeball.