Picked a local model I was happy with, then figured out what is required to get it running at acceptable tok/s, where I settled on a amd "ai" variant NUC with 96gb vram avaliable to gpu. This arrived recently, and qwen3.8 27b runs fast enough for me on that, with full context, and plenty of paralell streams. That said i'm in a fortunate position, so also got an rtx6000 to continue rlhf based finetunes in data I amass over the years from myself and friendly highly knowledgable/skilled friends, since that is in essence what makes fronteir labs models better, so i figure, why shouldnt we benefit from our expertise and input direcerly, instead of having it sold back to me by some amoral company? If that eventuates in a model that genuinly beats current, obviously we'd give back to the community by releasing that.