It's also used by Google Kubernetes Engine, OpenAI, and Cloudflare among others to run untrusted code.
It's also used by Google Kubernetes Engine, OpenAI, and Cloudflare among others to run untrusted code.
- You are using a container orchestrator like Kubernetes
- You are using gVisor as a container runtime
- Two applications from different users, containerized, are scheduled on the same node.
Then, which of the following are true?
(1) Both have shared access to an NVIDIA GPU
(2) Both share access to the NVIDIA GPU via CUDA MPS
(3) If there were 2 or more MIGs on the node with a MIG-supporting GPU, the NVIDIA container toolkit shim assigned a distinct MIG to each application
If you'd like to learn more, you can check out our docs here: https://modal.com/docs/guide/gpu
Re not using Kubernetes, we have our own custom container runtime in Rust with optimizations like lazy loading of content-addressed file systems. https://www.youtube.com/watch?v=SlkEW4C2kd4
(I work at Modal.)
E.g. it came up in this thread: https://news.ycombinator.com/item?id=41672168