CDB is one of my favorite data structures. When a student wants to learn about databases, I get them to implement cdb.
It's easy to implement and really demonstrates some good system engineering tradeoffs.
It's easy to implement and really demonstrates some good system engineering tradeoffs.
I wonder if there's any mileage in using a perfect hash function to build a database like this. It seems suited to the operating model of being slow to build but fast to access.
I think the theory is mostly that it's really simple and ends up requiring very few disk reads.