On another tangent - how do Google Cloud and EC2 attach GPUs to instances - given that you can choose CPU and RAM the GPUs must somehow be modularized away from a dedicated server?
On another tangent - how do Google Cloud and EC2 attach GPUs to instances - given that you can choose CPU and RAM the GPUs must somehow be modularized away from a dedicated server?
Disclaimer: I work at AWS on the team responsible for the Nitro System including EC2 Bare Metal Instances.
Rack A of servers has a base_server_x. Rack B of servers is base_server_x + GPU_Y.
You ask for no GPU, you get a server from rack A. You ask for a GPU, you get a server from rack B.
No magic monkey adding GPUs to instances ;)
It sounds a bit like MAAS [1], which allows you to throw images onto, and manage real servers easily, very much like you might spin up VMs on AWS.
[1] https://maas.io
https://aws.amazon.com/about-aws/whats-new/2017/10/announcin...