You can probably get it to behave well with a fine tuning like this: https://arxiv.org/pdf/2202.08904.pdf
[1] https://huggingface.co/BAAI/bge-large-en
You can probably get it to behave well with a fine tuning like this: https://arxiv.org/pdf/2202.08904.pdf
[1] https://huggingface.co/BAAI/bge-large-en
That's a great list of existing embeddings models (in addition the SentenceBERT models https://www.sbert.net/docs/pretrained_models.html).
https://github.com/Dicklesworthstone/llama_embeddings_fastap...
So to add the base bge model, you could just add this URL to the list:
https://huggingface.co/maikaarda/bge-base-en-ggml/resolve/ma...
I will add that as an additional default.
When I talk with people about ChatGPT-esque things (ChatGPT is pretty much what most people know now), I say that it's crazy that you can run this on your own consumer-level hardware if you just want to mess with it. You don't need "prosumer"/enthusiast hardware (unless you want to train models, but then I see people are using Google Colab etc).
It's a crazy world we live in.