Running Apache Kafka at Scale
engineering.linkedin.com
engineering.linkedin.com
We used graylog for a while but ran into some issues with back pressure on elastic search (hence kafka).
they had a major data corruption bug last year but have taken measures to correct it. I asked if elasticsearch could be a source of truth at elasticon and they didn't say yes, but they indicated that you "could do it" and it is a goal
1. Whatever mechanisim you're using to send data to Graylog, you send that data to S3 as well. You can then reload Graylog at anytime with S3 data.
2. Backup your Elasticsearch nodes to S3
I should've mentioned I run this in AWS. Sorry about that!
[1] http://www.quora.com/RabbitMQ-vs-Kafka-which-one-for-durable...