I mean, GPT3 requires some 800GB of memory to run, do we all have gazillion dollars supercomputers at home?
I think, unless there's some real breaktrough in the field or in the hw acceleration, this kind of model is going to stay locked behind a pricy API for quite some time.