Delicious's Data Policy is Like Setting a Museum on Fire
readwriteweb.com
readwriteweb.com
I've long seen the no-crawling policy of Delicious plus the Roach Motel API that was all about getting people to put their data in but not about letting people get it out as the dark side of "Web 2.0"; often we hear about an API as if it were a gift, but it's often a self-serving effort to take our data and give nothing back in return.
Remember that IP addresses have a market price of about $3 /month, and that's what an honest proxy cost Honest proxy providers rent machines in data centers and have them bind to a wide range of addresses, all in the same netblock. If you're coming from 20 different addresses in a netblock (paying $60) a month, you still look suspicious. These guys might have machines in several data centers, but they can't put you into hundreds of different netblocks.
The economics might get better if you're sharing the proxies with other people, but those other people are up to the Devil's work, and are busting their asses 24-7 getting the IP addresses in everybody's block lists.
As for Tor, quite a few organizations block or limit Tor traffic... Databases of active Tor gateways are available, and sites like Wikipedia use them... Wikipedia won't let you make anonymous edits from Tor, because they don't like dealing with griefers who use Tor.
Now, some people will use hacked machines as proxy servers. A botnet can create a nearly indetectable cloud of IP addresses, but as far as I'm concerned, use of a botnet is an ethical line I won't cross.
http://ilpubs.stanford.edu/858/1/2008-2.pdf
The primary issue is that it does not seem terribly clear what the legal status of redistributing such data would be, and/or whether this is changed in any way by Y! shutting down delicious.
In my mind, delicious represents everything that was great and everything awful about 'Web 2.0'. Yes, it did something that everybody assumed was impossible because bookmarking sites failed so consistently in 'Web 1.0'. Although delicious provided a useful current awareness service, it never did any of the interesting things with the data that would have made it possible to move onto 'Web 3.0'. And it probably never will -- but maybe that's fine because it opens up an opportunity for the rest of us.
Perhaps "in some sense" you feel that that data belongs to all of us, but it's not really your place to decide. Delicious had terms of use, and you violated those terms. Your argument sounds a bit like the "the internet is public domain" cookbook lady from last month. It was wrong then, and it's wrong now.
He's free to put up whatever defenses he wants to put up. Other people are going to try to tear them down. That's the way of the world.
We're already drowning in data. We need to start making some executive decisions about what's important and what's not. If it turns out we're wrong - we'll deal with it.
Find one and ask them how many parts of their body they'd pay to spend even ten minutes in the town square listening to the mundanities you so causally consign to the bit bucket.
There's more information there than you think, more than you can even see, because you are a product of the time that generated it.
They offer iOS and Android apps
I dunno if Geocities had a similar robots.txt, but it didn't stop several groups from archiving it (which was the right thing to do in either scenario).
Then, put it on a website, and tell people "by staying on this page, you are donating bandwidth and helping archive delicious".
It's so no-hassle that I bet you could get a huge following.
I know delicious had active defenses because I ran afoul of them.
'...both the pagan historian Ammianus Marcellinus and the Christian historian Orosius wrote that the Bibliotheca Alexandrina had been destroyed by Caesar's fire. The anonymous author of the Alexandrian Wars writes that the fires Caesar's soldiers had set to burn the Egyptian navy in the port of Alexandria went as far as burning a store full of papyri located near the port. However, the geographical study of the location of the historical Bibliotheca Alexandrina in the neighborhood of Bruchion suggests that this store cannot have been the Great Library. It is most probable here that these historians confused the two Greek words bibliothekas, which means “set of books”, with bibliotheka, which means library. As a result, they thought that what had been recorded earlier concerning the burning of some books stored near the port constituted the burning of the famous Alexandrian Library.'
Really sad to see it go.... If yahoo had asked me to pay I'd happily have done so (I pay for Flickr, Spotify, Last.fm, RTM and many other oft-used services happily)
If you're going to shut it down anyway, what's the harm in trying? Maybe have a "stay of execution" for a quarter - tell users you're going to charge $10/month for the service, and see how many users sign up. If you can break even, why not keep it?
I was envisioning something like reddit gold. They have more users/subscribers, and seemed to have a lot of success with their monetization.
Just as well, since it looks like pinboard gets the things that bugged me about delicious right.
Uhmmm... this worked for me:
curl --user user:password -o DeliciousBackup.xml -O https://api.del.icio.us/v1/posts/all
I only have 280 links on there, so maybe it is limited somehow. I really hope it is, otherwise this would be REALLY a poor job on the part of readwriteweb.
http://www.michael-noll.com/projects/delicious-python-api/
Unfortunately Delicious will throttle you if you hit the service more often than once a second so you might not be able to get too much valuable information.
I'm sad to see delicious go as it's a great collaborative tool and has awesome powers when combined with instapaper.
(btw, if anyone wants a copy of my script you can get in touch with me through my site listed in my prof)
edit: Not completely sure why you all hate this comment, but fwiw I was being sincere not snarky.