It's more akin to limiting the size of each shoe, rather than the number of shoes. The client might need memory for computational purposes. I routinely have more than 7 GB worth of inference output to be displayed. I can only display a portion of it at a time, and I react to every mouse move to render the ML annotations, which rely on that underlying 7GB array.
Laptops handle it just fine. I have no problems with macbook, for as long as there's no memory caging enforcement. I'd imaging, running it on a cell phone or a tablet should be feasible too.