Mostly, I think, we don’t really understand your argument that Intel couldn’t easily replicate the parts needed only for inference.
Obviously that is only one piece of software, but its a certainly a useful one if you are using one of the many LLMs it supports.