The die size is huge. This isn’t the kind of chip that would go into your MacBook, let alone an iPhone.
It’s for cloud based servers.
It’s for cloud based servers.
But give that time (e.g. microfluidics) - something interesting is that it would be extra hard to use all layers at once, but NN might be a good fit, imagining that computation will be sparse (subsets activating simultaneously)...
Will your comment age well? We'll see.
We might all be surprised if (somehow, ternary logic?) models come down drastically in size. It doesn't have to be the hardware getting more dense.