Show HN: ColBERT Build from Sentence Transformers
github.com
github.com
Assuming not much effort is required to make this work for similar models? (i.e. BGE)
Documents and queries embeddings can be obtained using .encode_documents and .encode_queries methods
I save most of my embeddings (python dictionnary with documents id as key and embeddings as values) using joblib in a Bucket in the cloud. I don't really know if it's a good pratice but it does scale fine to few millions documents for offline (no real-time) applications.
Do you have advice for how to measure the quality of the finetuning beyond seeing the loss drop?