ConstBERT: Efficient Constant-Space Multi-Vector Retrieval Research
pinecone.io
pinecone.io
What are y'alls thoughts on this approach? I would be curious on people's experience with multi-vector retrieval in production. Are you using multi-stage pipelines for retrieval? How do you currently balance the tradeoffs between speed, accuracy, and cost?