Bluntly, no.
The models which are small enough to run locally perform so badly it’s not worth bothering.
To run inference on the large models the perform decently you need the equivalent of two or three top end graphics cards.
If you're serious about looking into it now, consider looking at this project that lets you run a bunch of independent machines as a cluster for inference using Bloom:
https://github.com/bigscience-workshop/petals/wiki/Launch-yo...
(You'll need around 200GB of GPU memory across the machines in the swarm)