Just looked in the parts drawer at home and dont seem to have a $25,000 GPU for some inexplicable reason.
Just looked in the parts drawer at home and dont seem to have a $25,000 GPU for some inexplicable reason.
There should be a quicker way to differentiate between 'consumer-grade hardware that is mainly meant to be used for gaming and can also run LLMs inference in a limited way' and 'business-grade hardware whose main purpose is AI training or running inference for LLMs".
Defining GPU as "can output contemporary display connector signal and is more than just a ramdac/framebuffer-to-cable translator, starting with even just some 2D blitting acceleration.
I think it will also make sense to replace "H" with a brand number, sort of like they already do for customer GPUs.
So then maybe one day we'll have a math coprocessor called "Nvidia 80287".
"Accelerator card" makes a lot of sense to me.
Maybe renaming the device to an MPU, where the M stands for "matrix/math/mips" would make it more semantically correct?
I looked around briefly and could find no evidence that it's been renamed. Do you have a source?
Last consumer GPU with NVLink was the RTX 3090. Even the workstation-grade GPUs lost it.
https://forums.developer.nvidia.com/t/rtx-a6000-ada-no-more-...
Unless you’re running it 24/7 for multiple years, it’s not going to be cost effective to buy the GPU instead of renting a hosted one.
For personal use you wouldn’t get a recent generation data center card anyway. You’d get something like a Mac Studio or Strix Halo and deal with the slower speed.
So I wonder what I could be doing wrong. In the end I just use RTX 5080 as my models fit neatly in the available RAM.
* by not working at all, I mean the scripts worked, but results were wrong. As if H100 couldn't do maths properly.
It just means you CAN buy one if you want, as in they're in stock and "available", not that you can necessarily afford one.
adjective: available
able to be used or obtained; at someone's disposal