Linux server monitoring tools
aarvik.dk
aarvik.dk
* http://www.freshports.org/sysutils/htop/
* http://www.freshports.org/sysutils/py-glances/
* http://www.freshports.org/sysutils/apachetop/
Some others mentioned in the comments:
* http://www.freshports.org/net-mgmt/iftop/
* http://www.freshports.org/net-mgmt/bwm-ng/
* http://www.freshports.org/sysutils/xsysstats/
* http://www.freshports.org/sysutils/atop/
If you're running OS X / FreeBSD / Solaris, there are many useful DTrace scripts for system monitoring and profiling:
I went into the task thinking "the BSDs are OLD" yet there must be something great about this operating system. It's beloved in certain communities and used at popular successful startups such as Netflix. There must be something I've been missing that those who cherish it have come to understand.
What I found was the perfect blend of modern and old school. It's not old feeling at all!
It had all of the newest GNU software (and otherwise) that I've come to know and love on other operating systems but also a sense of stability about the core that you don't get with Linux. There was a sense of underlying structure that had actually been planned out instead of discovered over time and it made the whole process of learning about how it worked a pleasure.
In addition, the FreeBSD manual was actually helpful and gave me a sense of completeness rather than "the text in this wiki is just scratching the surface of a complex wrapper for x that used to be y".
It's simple yet powerful, up to date but solid and I'd highly recommend FreeBSD as a result of the experience.
Which is it? It would be difficult for documentation to be both complete and incomplete.
It can collect detailed memory usage profile of processes and when combined with some smart scripting it has a nice leak detection functionality [1]. Very useful when you run out of memory and want to find which daemon has used all of it.
[1] http://www.atoptool.nl/download/case_leakage.pdf
edit: formatting
$ du -sk * | sort -n | perl -ne '\''($s,$f)=split(m(\t));for (qw(K M G)) {if($s<1024) {printf("%.1f",$s);print "$_\t$f";last};$s=$s/1024}'\''While most sysadm books are years out of date, this one covers all the hot recent stuff like Dtrace and its equivalents on Linux, pidstat etc. Solid coverage and the author (from Joyent) knows his stuff. Available on Safari too.
ps. I can view new relic on my phone, so if I get a pagerduty, I can still see what's up if I'm at the beach.
We recently open source our Ruby instrumentation:
No server monitoring as of today though.
Or you can monitor your web-app using Appview Web too for synthetic monitoring. Plus you can monitor your network as well using Pathview.
And I've been getting way too may false positives on the systems alerts.
If you only want systems monitoring and not deep application performance insight, New Relic is way too pricey and not really that good.
munin-monitoring.org
> not really that good
New relic has such rich functionality that it is easy to overlook some of its utility. It took us a while to get it tuned to our needs, but now that we have it configured, I couldn't imagine running a high availabilty web service with anything else. Suppose I get a pagerduty for high memory usage on a server. I would then go look at that server in new relic, see what processes are using the memory, see what the memory usage for that process has been like for the last 6 months, perhaps notice a slow steady increase in memory consumption, realize there's a memory leak, etc.
If you're just starting out, the free tier is pretty good.
What the monitoring tools lack on specificity (and depending on your stack, it may provide varying levels of awesome -- server monitoring is weak but improving), it has massive win on zero-configuration installation.
Just sign up, instrument, and start monitoring.
If you find bits lacking, there are almost always local tools you can use to supplement.
Installing atop will also (depending on your distro etc.) set it up to snapshot the system state every 600 seconds. If you run "atop -r" you can review that legacy old data from today or an older day, and switch between the 10-minute snapshot with t and T.
Personally i like "sar" for quick text only overview (sysstat package). Once enabled you have a 10 minute snapshot of a huge amount of performance metrics (e.g. sar -r for memory, sar -b for disk). Of course, it's even better if you use something to collect them centrally (I signed up for DataDog which takes very little effort to integrate compared to rolling your own stuff).
http://stackoverflow.com/questions/7278326/how-to-monitor-a-...
We're always looking for feedback, and we're happy to give out discounted or free accounts to startups. Drop me a line -- steve@[company domain] -- if you're interested.
[1]: https://www.scalyr.com [2]: https://www.scalyr.com/appDashboard
$ man sarDemo here: http://afaq.dreamhosters.com/linux-dash/
Easily extensible if anyone wants to use it.
iftop
nload
htop
goaccess
...I'll add more as I remember them, currently gotta sleep.And OSSEC Web User Interface (ossec wui): https://scottlinux.com/wp-content/gallery/site/ossec_web.png
[1]: http://sebastien.godard.pagesperso-orange.fr/documentation.h...
Noting when things go titsup.com can be particularly useful.
Seriously, glances is nice, but I generally use tmux with a bunch of panels, showing stuff like `watch -c df -h +nr`
Couldn't come up with anything else to promote your useless blog?