Realtime Data Processing at Facebook
muratbuffalo.blogspot.com
muratbuffalo.blogspot.com
The founders were the guys who built Scuba, and we're taking a somewhat different approach (mostly driven by differences in scale). We're not quite at the second scale delivery times, and are based on more classical logfile rotation and aggregation mechanisms to get our raw data, and then an efficient sharding layer to get it into our columnstore.
AFAIK your tool cannot seem to identify trending events as they are streamed in (like moving standard deviation for example) and feed downstream to a pipeline unless I am mistaken
"We like JSON best." - lol
We also support CSV and apache logs, but JSON is what works for customers.