Right on -- for the record, this data is ending up in mongodb at the moment, so I need to write logic for binning by time and so forth myself. Our use case is maybe odd in this space -- sampling a handful of values per item a couple times a day, but for millions of items.
I compared a dataset to postgres and the new mongo storage engine is very good -- dunno if wiredTiger is something you guys are looking at, but it's very compact and fast to query with the proper indexes.
Dumped some stats into a gist as well: https://gist.github.com/mattbillenstein/89969980025414e2bca8...