That "Ask the Bible" service that went viral the other day on Twitter had me thinking the same thing.
Especially for the set of problems with "Step 1: Hosting embeddings for some large corpus", it feels like there's a really useful role for offering static query/AKNN search atop popular datasets.