Tynt: What's being copied from your website right now?
tynt.com
tynt.com
What happens if the user just deletes the link, as I have with the above quote?
[1]After all, if someone copies your work and no once sees it, who really cares? The scope of your web crawler must be limited somehow. You could simply have it check up on your competition and/or popular links on news aggregating sites for categories related to your business. For example, if I was starting a technology blog, I could write up a Python crawler in about 30 minutes that parsed through blog articles on other popular tech blogs with related tags.
I think there is a need for a service like this (that watches the web for people republishing your content), but Tracer is not it. Tracer is just a silly JavaScript hack.
http://learninfreedom.org/colleges_4_hmsc.html
mentioning a college that doesn't actually exist, but has a name from the Greek word for "steal." That finally got one persistent thief to acknowledge that my site was his source.
There is one (somewhat) famous example of a trap, in the placement of two fictitious towns in Michigan, Goblu and Beatosu (this being back in the day of their rivalry).
http://en.wikipedia.org/wiki/Fictitious_entry
http://www.newyorker.com/archive/2005/08/29/050829ta_talk_al...
Similar technique I guess; and he took the hint after a couple of days.
For example, if you feed them www.tynt.com, they'd tell you that feedmyapp.com has borrowed a few sentences from them. (It makes sense, as the borrowed phrases are part of a listing for tynt.com)