Anybody knows Cliqz's database stack? Curious to see what powers a large scale information retrieval index of this sort.
We will have a blog post tomorrow on this very topic, but in short, we use a combination of Keyvi, Granne (both in-house) along with Cassandra and RocksDB.
Though our approach mentioned in this blogpost significantly reduces the storage needed to host the index, we still have an index of around 50 TB of data.
FYI, there will be a bunch of articles regarding search in the next week.