So the A100 is 75% more efficient despite being in a high power variation of a larger node. Seems Apple still haw a lot of work to do in the GPU/ML side of things.
Honestly though I'm more concerned about the other aspects of apple's software: "The OSX Window Server crashes when all GPUs are used to the maximum"
The good news (for us) is that this is all effectively -O1 today; there's still potential for 2-8x speedups over these numbers and that's not even factoring in the inaccessible HW features that maybe one day Apple will expose :crossed-fingers: :)
// IREE dev
https://timdettmers.com/2020/09/07/which-gpu-for-deep-learni...
It seems that their consumer GPUs also have a gap
It is, but they’re not using any of that. Only the M1 GPU at this time.
//part of nod.ai / SHARK team.