It drove me to setup Qwen 3.8 this weekend. I couldn't see the value in just giving them money for a higher tier plan instead.
I've never run a local LLM model before. Certainly won't take as long to iterate on this.
I've never run a local LLM model before. Certainly won't take as long to iterate on this.
I have only 48gb of ram, so can fit only 80k context max, so good compaction is must.
I suspect the same will/is happening with AI. Either you will pay for it, or it will be so ad infested that it will become useless.