I just used this at work the other day to calculate similarities between different data models that had overlapping children models. One of our teams was going to go through manually to check these overlaps and consolidate, but by using this clustering algo based on Jaccard distance we were able to give them clusters to consolidate up front. Super cool stuff!