1. One of the reasons we created Khoj was being able to do natural language search with embeddings generated offline using open-source models!
2. We don't use any vector datastores (yet). You can do a lot in memory, it's faster and it does exact matches (no KNN, approx matching)
Feel free to ask if you were looking for something more specific?