Docracy Terms of Service Tracker
docracy.com
docracy.com
[1] The reality is far from this, of course.
[2]http://asiajin.com/blog/2010/05/21/english-please-rakuten-ba...
A very strict translation with no adaptation or nuance for the market they're entering.
1. https://www.docracy.com/doc/versions?docId=0b0kbmmpoon
version 1 looks pretty wrong (https://www.docracy.com/0b0kbmmpoon/local-com-privacy-policy...).
Wayback has versions from Jan 15th and Jan 16th (http://web.archive.org/web/*/http://www.local.com/privacy/) around the same time downloaded, and they look more normal.
2. Edit and download seems to pull the wrong version in some cases:
http://www.docracy.com/0xk2nizy6sk/fidelity-com-privacy-poli...
(It says version 1, displays version 1, Click edit and download, you get version 2)
3. The diff engine doesn't seem to try very hard in certain cases:
http://www.docracy.com/doc/diff?revisedId=0razhem25wh&or...
(The first paragraphs of these documents are a lot closer than it makes seem). It seems to have a bunch of stream alignment issues, which makes me think you are using a line based diff here, and post-processing the result.
Anyway, besides the above just found playing around, it looks otherwise nice.
Being able to schedule when you are sent notifications, batch them together (eg. if you're watching to web pages, you get one email each day at a specified time), as well as increasing the frequency of polling (goal is 1/hour) is on the list.
I'm working on a feature I see as unique in that it will allow you to choose which part of the page counts as a change, thus minimizing false-positives.
Can you think of anything else that would be useful? Something frustrating you of ChangeDetection?
If-this-then-that integration might be cool - but that's probably an edge-case not useful to most of your potential clients.
As a further feature request, it would be really great to be able to flag/vote a diff as alarming, so it could be highlighted for more people can notice it. For example, this diff by Geico is pretty questionable:
http://www.digitalpreservation.gov/formats/fdd/fdd000236.sht...
By default diff programs create a line-based output, but you can change it to minimum per-word highlighting via options (e.g. 'git diff --color-words').
The thing with PDF is that often even when you re-save the same PDF file in the same editor, you would probably get entirely different files. I'm not a PDF expert but from what I've learned, PDF is the type of file that saves kind of vector representation of glyphs and their placements and is often unaware of what that glyph represents (depends perhaps on the program used to create the PDF and options). Importing PDF back to e.g. OpenOffice is an ugly work for the plugings.
There are some exiting solutions for diffing PDFs [1] however I haven't played with them really.
[0] http://en.wikipedia.org/wiki/Diff [1] http://stackoverflow.com/questions/887186/java-pdf-diff-libr...
I really wish you guys would knock that off.
that said some of these terms look scary.
I have a striking suspicion that the lawyers (or webmasters) are just copying and pasting a lot of these terms from standard repositories or otherwise from other services.
"We do not save this data nor disclose it to any third parties."
(anything but comforting...)
As other have commented, a discussion area for each change would be very interesting, especially if there are multiple changes happening at the same time.
I can imagine not everyone want this focus on changed tos, but its very good the user can easily get the information.
I'm really curious to know what's so unique about this as compared to classic diff plus colours?
http://www.tosback.org/timeline.php
Source code is available at:
https://github.com/pde/tosback2
Historical crawl data is available at:
However, the timestamp on it is January 29. How often do you check for updates?