And you’ve opened wireshark and verified the model is sending absolutely nothing? Not caching and sending later, etc?
The model consists of a bunch of data files, it does absolutely nothing by itself.
If you run inference on your own hardware, you have absolute control on how the LLM is used, not like when you use an external service provider.