Show HN: Infreqdb – S3 backed key/value database for infrequent read access
github.com
github.com
We did this with S3 as the storage engine, 100M+ records for $10/day: https://www.youtube.com/watch?v=x_WqBuEA7s8
And discord had a very nice article on this as well: https://blog.discordapp.com/how-discord-stores-billions-of-m...
Great work, I think there is a lot of exciting stuff you can add to it!
Also why don't you dump the log data into a NoSQL like dynamoDB instead of S3 ?
> Also why don't you dump the log data into a NoSQL like dynamoDB instead of S3 ?
Price.
> Write Throughput: $0.0065 per hour for every 10 units of Write Capacity (enough capacity to do up to 36,000 writes per hour)*
> A unit of Write Capacity enables you to perform one write per second for items of up to 1KB in size
As I understand it, for $4.68/month I can only add 36 MB/hour, and thats assuming my objects are in exact multiple of 1KB.
That would probably turn out to be cheaper than using S3 - https://github.com/minio/minio
Is this basically equivalent to Datomic, then? (Not that that's a bad thing. The world needs an open-source, non-JVM-targeted Datomic.)
Couchdb?
I think flat JSON files wont be efficient. My goal is to have the cache on disk, and each cached file would be big with lots of keys on it. In order to use JSON files, I would either have to keep the whole parsed data in memory, or parse the whole JSON each time I want to lookup a key.
If the data fits in memory then sure JSON is more convenient.