> Facebook just published a blog about moving petabytes per hour.
For the curious: https://engineering.fb.com/data-infrastructure/scribe/
Edit: HN thread: https://news.ycombinator.com/item?id=21181982
For the curious: https://engineering.fb.com/data-infrastructure/scribe/
Edit: HN thread: https://news.ycombinator.com/item?id=21181982
Basically, everything that needs logging and post-processing by both real-time systems (e.g. Puma) and batch processing (e.g. all of the data that's ingested and sent to the data warehouse) goes through Scribe.
(disclaimer: I work in Scribe)
Scuba: https://research.fb.com/publications/scuba-diving-into-data-...
Puma: https://research.fb.com/publications/realtime-data-processin...