ParentFull threadfarseer·Yup, because LLM inference can be scaled by adding racks of hardware. Consulting can't be.View on HN