> The biggest problem: developers don’t want GPUs. They don’t even want AI/ML models. They want LLMs.
I considered using a Fly GPU instance for a project and went with Hetzner instead. Fly.io’s GPU offering was just way too expensive to use for inference.