Ask HN: What's the point of training bigger language models?
What's the point of training bigger language models when you can't even serve (inference) them economically? What am I missing here?
They're built to impress people and push the boundaries of can be done with ML. Economies, accessibility, scale - those things come later.
Easy to inference/small model = poor results