469 karma · joined July 25, 2009
[ my public key: https://keybase.io/josephruscio; my proof: https://keybase.io/josephruscio/sigs/w29vqp7iBRWN0reTQwXymfokT6d4I0rgcmB5cq1eisE ]
- Appreciate the thorough and balanced approach you took here, there's a lot of great feedback, most of which helps validate our current roadmap and 1-2 things we'll certainly consider.
- The tradeoff of providing a true utility service where you only pay for what you use is that the actual makeup of your bill can become complex (think of your AWS bill). We do pro-rate usage to the hour right now (matching AWS), so unless your metric "lifetimes" are sub-hour there's shouldn't be a huge issue.
- There's a similar power/complexity tradeoff with rollups. Unlike some solutions that only rollup via average we track multiple summary statistics e.g. min, max, sum, count, weighted average. Managing the differences between these can be confusing in some cases, so we're currently working on a simpler interface to toggle between these. More details can be found here: https://www.librato.com/docs/kb/visualize/faq/rollups_retent...
- Mobile support is something we've been incrementally adding and will continue until it's complete e.g. the actual alerts page you land on from a notification is now mobile-ready.
- Tagging. In the years since we first launched (when AFAIK only OpenTSDB supported this) it's become clear that this is becoming more and more a standard capability. Which is why we've been building out a next-generation data-store, which supports indexed tagging and a bunch of other new capabilities. We expect it to be in beta this summer.
Librato and Papertrail recently joined forces with Pingdom as part of the Solarwinds Cloud family. We're respectively industry leaders in metrics and log/event management/analysis-as-a-Service. We process billions of events in real-time every day for tens of thousands of users with a small team of ~20 engineers.
We're looking to expand our engineering teams to help build the next generation of real-time IT analytics products. We're looking for great frontend engineers, data pipeline engineers, designers, etc. We're a modern shop that practices chatops and continuous delivery using tools like Slack, Github, Asana, AWS, Salt, etc
You can find a list of current openings at http://solarwinds.jobs/jobs/?q=librato+papertrail
jobs _at_ librato.com
San Francisco, CA or REMOTE
Librato is changing the way teams monitor/manage their production infrastructure. We've built a world-class platform and need talented individuals to join us in taking it to the next level! While we're headquartered in downtown San Francisco fully half of our team is spread across the continental U.S, and we care more about what you can do than where you live. If you have a passion for bridging the gap between raw data and actionable insights and are interested in one of the positions listed below, we want to talk!
Front End Developer - Data visualization, Coffeescript, jQuery
Operations Engineer - AWS, Chef, ChatOps, Continuous Delivery
Support Engineer
Developer Evangelist
Director of Marketing
If you're interested in dashboards, our startup recently launched a service that makes it simple to hook in your metrics through OSS tools like StatsD and build real-time dashboards with a few clicks: https://metrics.librato.com You just provide the data, we do all the rest. Would love any feedback you might have!
In short a measurement is a single datapoint i.e. <key, value, timestamp>. So if you have a sense of how many metrics you want to track and at what frequency, it's relatively straight-forward to calculate the list price. There's an estimator on the pricing page to help you do that once you understand what constitutes a "measurement".
When metrics come into one of the API instances, they turn around and insert it into a Cassandra cluster that we're running spread across 3 different availability zones with an RF factor of 3, meaning your data is stored in 3 different availability zones.
The "GET" calls to pull data out of our service go through the same API, so as soon as data is written to the cluster, it can be pulled back out. Hence the marketing term "realtime".
1) Your time is the most expensive resource. The cost of hosting your own solutions is almost always dwarfed by the cost of time you spend configuring it, maintaining it, and recovering it in the face of failures. We provide the same value here as any SaaS team in any vertical. We care and manage for the infrastructure and are constantly developing and rolling out new features. Of course the time needed to invest depends on one's time and experience, but we intend to save a lot of people a lot of time.
2.) A small EC2 instance costs $61/month and has finite disk bandwidth (and CPU). You might get more than 50 metrics, but it's going to come up a lot short of "nearly unlimited". Most people I know running serious Graphite installations end up needing collocated physical hardware with SSD's. That's going to cost you more like $1K-$2K/month. You will still have a SPOF unless you double that cost. We handle all the scaling and reliability for you.
3) Our pricing is completely linear per the number of metrics, the steps in the slider are just to make it easier to chunk around different numbers. There are no step-wise increases that double your costs. I appreciate your comment here as it had not occurred to me that someone might (reasonably) infer a stepwise increase in pricing.
4) We also include other valuable tools like threshold-based alerting on all your data streams with GUI integration to 3rd party services like Campfire, PagerDuty, Email, Custom Webhooks, with more to come.
5) A lot of 'monitoring' companies are springing up these days because the market is clamoring for it ;-). While some teams would rather handle these things in-house, a lot of other teams would rather focus on building their core business value and out-source infrastructure head-aches. It's the same economics pushing teams to outsource version-control, logging, hosting, etc.
In the spirit of sharing, here's our campfire bot developed in Ruby on top of the Scamp (https://github.com/wjessop/Scamp) framework: https://github.com/josephruscio/twke
May not have started it if the hubbers hadn't taken so long with hubot ;-).
Librato - http://librato.com is looking to hire a fifth engineer into our small team. We've building a great product for infrastructure monitoring/management with enthusiastic early adopters. As a ground-floor member of the company you’ll receive a competitive salary and a meaningful equity position:
Here's an example of the visualizations we provide into what's going on in your server instances: http://support.silverline.librato.com/kb/monitoring-tags/app...
http://uec-images.ubuntu.com/query/lucid/server/released.txt