Buy 4–8 used 3090s (providing 96–192 GB of VRAM), depending on the model and weight quantization you want to run. Used 3090 costs around $800. Add more RAM to offload layers if needed. This setup currently offers the best value for performance.
https://www.reddit.com/r/LocalLLaMA/comments/1iqpzpk/8x_rtx_...
You can look for more rig examples on that subreddit.