When stable diffusion came out, my 8gb gpu could handle 1 image at 512x512. Now i can generate 1000x800 with batch size 4. I wonder if Chatgpt has the same unused optimisation potential
I would bet that we can do better than GPT3 with less computing resources but that would need a different model architecture.
Doubtful. As an example, OpenAI released Triton, a programming language explicitly for optimizing models. For the scale they're at they'd be crazy not to throw an engineer or two at optimizing the models they're going to send to production.
Our brains outperform GPT-3 with only 30w of power. The potential for software optimization may be many orders of magnitude.
I think I have DiffusionBee installed but I haven't updated it for months