ParentFull threadkraken12·They are using many chips and taking advantage of the way data flows in LLMs to make it work; so it would be cost-effective unlike CerebrasView on HN