Mean Shift Clustering
spin.atomicobject.com
spin.atomicobject.com
I don't understand the author's reasoning for preferring mean shift over k-means due to k-means needing the number of clusters as an input. It seems that in mean shift, choosing the kernel bandwidth parameter is just as arbitrary in that it requires domain knowledge. In k-means there are tests[0] for choosing an appropriate k. Is there a similar strategy in mean shift for choosing an appropriate kernel bandwidth parameter?
[0] http://en.wikipedia.org/wiki/Determining_the_number_of_clust...
Like most clustering problems, if you can't choose a reasonable set of parameter values based on some domain specific information, it is largely trial and error. Of course, there are several metrics that can be used to score certain clusterings (e.g., parameter values) over others, as described in your wikipedia link.
The other nice advantage that mean shift has over k-means is that it does not make any assumptions about the cluster shape. K-means assumes spherical clusters. Mean shift allows for clusters of any shape, since it is driven by density.
One recently invested method of clustering called QuickCluster doesn't require any parameters and inherently operates on loss function. It's pretty fast in practice: http://www.cs.yale.edu/homes/el327/papers/CorrelationCluster...