10 karma · joined September 21, 2026
There are already: - hundreds of western providers hosting open models, including large ones such as Kimi K3 which is near Opus/Sol sizes.
- people hosting large models such as Kimi K3 at home. Check on reddit, there are people with 20x DGX Spark setups etc.
If you signed up for a 16 core AWS server and they randomly kept changing it down to 8 cores, you'd sue them. How long until the same applies to these AI companies.
The amount of meddling they do to the harness, system prompt, model, quantization etc makes these products sometimes unbearable; you never know what you're going to get. A few more iterations of Qwen 27B and hopefully we won't have to deal with any of this malarky any more.
Opus and Sol are better for day to day dev work IMO in that they won't try to do too much.
At least Codex CLI in rust uses 80-100MB which is still not great but a big improvement.
These AI companies screw you over from both sides. They make the cost of RAM skyrocket and then they build terrible software that uses all the RAM you have.