Wow, I didn't know about "Hermes 2 Pro - Mistral 7B", cheers!
Q5 is minimum.
https://huggingface.co/NousResearch/Hermes-2-Pro-Mistral-7B-...
If you have the time, could you explain what you mean by "Q5 is minimum"? Did you determine that by trying the different models and finding this one is best, or did someone else do that evaluation, or is that just generally accepted knowledge? Sorry, I find this whole ecosystem quite confusing still, but I'm very new and that's not your problem.
If you're RAM constrained, you'll also have to make trade-offs about the context length. e.g. you could have 8 GB RAM and a Q5 quant with shorter context, vs Q3 with longer, etc.