OpenAI's Sol ultrafast (powered by Cerebras) is still in preview, presumably because they're overall capacity bound.
OpenAI's Sol ultrafast (powered by Cerebras) is still in preview, presumably because they're overall capacity bound.
Because they already have / had an okay coding subscription product for a bit and it gives them visibility and mindshare (in regards to their hardware, even if they don't compete with other providers that much). They could do what Kimi did - make a good subscription with good models, once you get enough customers to get some good PR and such, pause the signups so you don't have to spend more on running the service than you want/can. Do enough of that and people will talk about your offerings organically, make yourselves known to even devs as "That one company with their own hardware and the super fast subscription." experiencing which would do more than any marketing.
Cerebras is a B2B hardware company. It feels like a distraction: think of the opportunity cost, and resources/headcount not working on other things that would drive more impact.
Should NVIDIA do a coding subscription too? I'm sure they can make money off it, but I think it would be -EV.
Yes, obviously! Well maybe not a subscription but definitely an inference service.
https://resources.nvidia.com/en-us-inference-infrastructure/...
https://www.nvidia.com/en-us/data-center/dgx-cloud-lepton/
In their case not to gain mindshare or money or whatever, they're already a market leader, but to run something that validates the use case of their own hardware (across a bunch of 3rd party models) on a practical level and gain whatever insights or details might be relevant to pass on to other hardware and software teams.
They sort of do? They offer free access to various versions of nemotron via multiple routing services.
But it's not fully open to just anyone, I wasted time signing up to find out that I couldn't even sign up for it to test it out.