You can't really do that with distinct, as if you have 1 billion distint entries, you essentially have to store all of them to dedup.
Furthermore, PipelineDB has a special combine [1] aggregate that allows you to combine data structures such as HLL across multiple rows with no loss of information. A simpler example would be average: to get the actual average of multiple averages you obviously can't simply take the average of all the averages. Their weights must be taken into account, and combine handles that.
The capability to combine aggregate values in this way generalizes to all aggregates in PipelineDB.
[0] http://docs.pipelinedb.com/aggregates.html#hyperloglog-aggre...