Could someone more knowledgeable suggest when it would make sense to use the SentenceTransformers library vs for instance relying on the OpenAI API to get embeddings for a sentence?
Could someone more knowledgeable suggest when it would make sense to use the SentenceTransformers library vs for instance relying on the OpenAI API to get embeddings for a sentence?
I presume if your customers are enterprise companies then you may opt to use this library vs sending their data to OpenAI etc.
And you can get more customisation/fine-tuning from this library too.
The main reason to use one of those providers is if you want something that performs well out of the box without doing any work and you don't mind paying for it. Those companies like OpenAI, Cohere and others, already did they work to make those models work well on various domains. They may also use larger models that are not as easy to deal with yourself. (although as I mentioned previously, a small embeddings model fine-tuned on your task is likely to perform as well as a much bigger general model)
There isn't a single usecase where they're better than the free models, and they're slower, needlessly large, and outrageously expensive for what they are.
Now it depends un specific usecase (domain, language, length of texts)