Not mentioned on the website (because it’s targeted at general website owners rather than a technical audience) but we are using a 100% open source AI stack for this, with llamaindex, pgvector and llama3:instruct running on ollama hosted on a stack of GPUs we have mostly in our houses.