Tomorrow my desktop computer hopefully
"Tomorrow" your desktop computer might be twice powerful but at the same time the "good model of tomorrow" will be four or ten times larger - I'd expect that the gap between what can be done locally versus what is offered as a service will grow, not shrink.
That doesn't mean you'll be able to run the best model, but I'm relatively optimistic about the gap not growing out of control.
git clone https://github.com/ggerganov/llama.cpp
cd llama.cpp
make
./server -m models/7B/ggml-model.gguf -c 2048
I don't think it'll take you the whole weekend :)