Would love an answer on this too. It would be even better not just to try using this, but also be able to run it locally, something that has been impossible for GPT-3.
If you had 1 bit per parameter (not realistic), it would still take ~100 GB of RAM just to load into memory.