Craigslist Suing Padmapper
gigaom.com
gigaom.com
http://www.quora.com/Why-hasnt-anyone-built-any-products-on-...
But now that 3Taps has found an ingenious way to get at the data with zero extra bandwidth cost to Craigslist (by retrieving it from the Google cache rather than CL itself), it's clear that what Craigslist really dislikes is competition.
While Craigslist is probably within their legal rights here, this case shows that for all their talk about their benevolent aims, Craigslist is no different from other companies.
I suppose we do have some estimates for their revenue, but not how much profit they actually make.
http://www.businessinsider.com/2011-digital-100#10-craigslis...
http://articles.businessinsider.com/2011-10-07/tech/30253426...
They're profitable alright, but let's not kid ourselves that they have by choice left a LOT of money on the table, which is why all these value-added services are trying to take a share of the CL-pie.
Hm, would you not say that what Craigslist really dislikes is their competition piggy-backing off their data? Craigslist seemed quite happy with the existing situation until Padmapper launched their own listing service.
Though I agree that it's not in the spirit of a .org, if such concept exists.
It's not their data. It's ours, as the users. Craigslist doesn't own anything I upload to them.
The reason PadMapper exists is because Craigslist refuses to not suck.
It says that you own everything you post, but you are granting them the rights to use it. The only unusual bit is that you are additionally granting them the right to go after people who scrape your content on your behalf.
> You also expressly grant and assign to CL all rights and causes of action to prohibit and enforce against any unauthorized copying, performance, display, distribution...
Seems to be vital to this case.
My personal interpretation/hope is that the right to sue for copyright infringement is nontransferable, which would give Craigslist no standing to sue. Individual posters could sue, however.
[1] https://www.eff.org/files/filenode/righthaven_v_dem/order6-1... (See also: http://en.wikipedia.org/wiki/Righthaven_LLC_v._Democratic_Un...)
I suspect this may have been one of the introduced terms.
Craigslist is delivering exactly what it its users signed up for (no more no less). People posting ads on Craigslist do not necessarily want or intend for it to be reposted on other sites.
I had to giggle a little bit at this in the grand scheme of the internet, sorry.
When I posted something on Facebook Marketplace, I started getting emails and comments from other "market" sites that Facebook had cross-posted my listing to.
While it was annoying to not know this up front, the fact that it was more visible and getting more bites because of it only helped me make the sale quicker. If a service wants to piggyback off of another to make my postings more buoyant, as a user and seller, I don't have a problem with it.
That argument appears to be invalidated by Craiglist's own terms of use, which say that when you upload a listing to Craigslist they can syndicate it wherever they want (http://news.ycombinator.com/item?id=4287519).
Still, I think there's something admirable in the simplicity and transparency of interacting with Craigslist. What you see is pretty much exactly what you get.
The point is Google respects publishers' desire not to be indexed. Padmapper does not. Robots.txt is simply a common method for conveying that message. It's not like Padmapper could argue they didn't know CL was unhappy; they got a certified letter!
This is just semantics. Google respects publishers who do not want their sites listed.
If this is illegal or otherwise objectionable without an explicit agreement from each Craigslist poster, I'm not really sure how a search engine or even descriptive hyperlinking is kosher without explicit agreement from each website indexed or referred to.
"Though the most successful founders are usually good people, they tend to have a piratical gleam in their eye. They're not Goody Two-Shoes type good. Morally, they care about getting the big questions right, but not about observing proprieties. That's why I'd use the word naughty rather than evil. They delight in breaking rules, but not rules that matter. This quality may be redundant though; it may be implied by imagination.
Sam Altman of Loopt is one of the most successful alumni, so we asked him what question we could put on the Y Combinator application that would help us discover more people like him. He said to ask about a time when they'd hacked something to their advantage—hacked in the sense of beating the system, not breaking into computers. It has become one of the questions we pay most attention to when judging applications."
Generally, I don't think you want your naughtiness to tie you up in a legal battle.
Can CL "win"? Can they stop everyone else, PadMapper is but the first of many, no doubt, from re-displaying facts (classified ads) in different formats?
The problem with a lot of startups in SV is that their only objective is to make money. Nothing else. Craigslist is very very rare exception.
Now lets downvoting starts...
Craigslist wastes several human lifetimes worth of time every month through maintaining a monopoly product with a shitty UI and refusing to let anyone innovate on top of it. They hold back progress and they are evil. They are the IE6 of classified ads.
That's a good one. I'm going to start using it. (Unless you make copyright objections.)
"folks, please remember, #craigslist community feedback massively against the use of their stuff for the profit of others."
I, for one, hate my competitors. Basic insticts perhaps?
Is not disliking your competitors something that is practiced by majority? OR is it at least very common? Common enough to point at a company that doesn't follow it?
Hmm, sounds familiar...
Yes, they got an early advantage into the market and has the critical mass that many company can seem to compete with but why take that away from them because they refuse to update/add new features.
Instead of piggybacking on them and relying on their data to earn money, shouldn't Padmapper focus on building their own content. Isn't that where innovation comes from? Beating an existing company by creating a better platform?
There is nothing forcing people to use Craiglist. They are not a monopoly, nor do they act like one.
The vast majority of people searching classified ads are searching craigslist, therefore if you're trying to list something in a classified ad, you're forced to use craigslist.
Sure you could use another service, but Netscape could have also just sold browsers only to Linux customers. It's all about the numbers.
People aren't really forced to use craigslist. That the vast majority of people choose to search CL isn't good enough. Another company could spend whatever it takes to get people to search their classifieds instead. As long as CL can't or doesn't block that (in contrast to stuff MS was doing that started the DOJ case against them), the competition is viable.
And its not about a matter of choice the users have. There are many competitors in this market, and yes Craigslist dominates every one of them because they had an early advantage on the internet. Small sites can't just leech off the contents on their website and slap ads on it to make money.
Lets say Craigslist was a print company that produce and distribute classified as. Will it be right for a small company to steal their content and slap their ads on it and distribute it themselves?
Your argument won't work here, because there are many major newspaper that do 1000x in revenue and distribution than independent newspapers.
I would really call this into question for the copyright claims. The key claim is copyright infringement, and the listings on Craigslist are almost certainly unprotected, much like telephone directory listings. They are statements of fact rather than creative works.
Craigslist may attempt to claim copyright over reproduction of their database as a whole (compilation), but the Supreme Court has ruled that in order to be eligible the compilation must be "original in its selection, coordination, and arrangement", and Craigslist can hardly claim that.
The breach of contract claims may be stronger though. The pages on the Google cache presumably still contain the ToS, and if they apply (who knows?) then PadMapper's use of Craigslist would likely constitute a breach. A ruling on this would be very interesting.
So there could be a entire listing for a SF apartment with lots of details and padmapper would not copy any of it- instead they would just make notes - 3 bedroom, location, price, pictures and then link back to CL if someone clicks the listing on PM
Would this work? Can a phonebook just put a section at the beginning of this page that by reading this book you are agreeing to the ToS which state you can't copy it?
I know most ToS contain a section about by using this service you are agreeing to the ToS.
But I think there is a good argument that Padmapper isn't actual using the service since they aren't getting the information from Craigslist's servers.
Am I using a service, and thus bound by contract if I look at a screengrab of a website on a third party website (that contains a ToS).
I've said before though that Craigslist could invent fictitious entries, and sue Padmapper for copying those creative works, just like mapmakers do with fake towns.
As for phone books, a shrink-wrap license on a CD-ROM phone book which prohibited copying was upheld by the US courts (ProCD), so that copying was a breach of contract even though the underlying data was unprotected.
> PadMapper isn't actually 'using' the service.
Using is a wonderfully subjective word, and I'd expect accesing, and making use of to be acceptable synonyms.
> I've said before though that Craigslist could invent fictitious entries, and sue Padmapper for copying those creative works, just like mapmakers do with fake towns.
These are known as "trap streets" and US federal court has ruled that they are not protectable. Map makers are able to sue because the rest of their map is protectable; the trap streets simply catch the infringer red-handed.
How many other websites use arguments like "bandwidth" to falsely portray competitors who access their publicly shared data as somehow in the wrong?
Many. Some here on HN. No need to name names.
No doubt even Google would complain about people "scraping" search results.
To me, it is a joke. Because the people who complain use automation to access, retrieve, organise and serve information and thereby establish their business. Only then to try to forbid others from using automation to do the same.
And all the while, it's NOT THEIR INFORMATION. This is not Craigslist's data. It's users' data.
It belongs to users, who are today's "publishers" and possess all those good ole publisher's rights. (Though they may naively license them out.)
They did give Craiglist permission to prevent others from using their data without permission. In this context, scraping any version of Craiglist's site (whether CL itself or a third party cache) falls within Craiglist's rights under the license they were given, and within the user's expectations of what Craiglist will do with their data.
Google is allowed in robots.txt. But that is not exactly what I would call an agreement.
The simple fact is this info is on the public web which, by its nature, copies and transfers data. That's what the web does. You upload something and it goes "viral". You have principles like the "Streisand effect" to contend with.
This goes back a long way. No doubt judges remember. The Ken Starr report on Ms. Lewinsky. Some random classified ad. Like it or not, information gets desseminated.
If you want to protect and restrict access to data, then you do not upload it to the public web. You put it behind access controls, e.g., a password. This is common sense.
If anyone has a claim here, it's users who do not want their ads on PadMapper (if there are any). CL has no standing and their motives are both pathetic and transparent.
It's easy to interpret your post to mean that stealing/reusing data just because you can is fair competitive practice. Please do clarify because your words mean a lot to people and your post might be interpreted that way.
Hilarious to expose their sanctimonious nonsense though.
If they go to court and get an opinion it should be really helpful as a guide for other entrepreneurs who see re-processing the information on the web in new ways as the foundation for their business.
You mean, like Blekko?
The simple, inconvenient truth is most folks who are making money from the web, like search engines, are not content creators (nor content owners), they are content publishers... who publish for free. "Are you a non-technical person who wants to get something onto the web? No problem. We'll help you with that, for free. Just give us some personal info about you so we can solicit money from advertisers."
(Placement, e.g., paid placement, where the eyeballs are more likely to see something, for a fee, is another matter.)
Now that use has generally not been highly contested, people want their pages to be found and so they tolerate search engines searching them. They can be explicit in what pages they want searched and which they don't using robots.txt. So that relationship is pretty well understood. People who ban Blekko (and presumably anyone else) from their robots.txt file are not crawled by us, we recognize and honor that it is there choice if they want to be in our index or not. On the copyright issue however it has been pretty clearly established that 'page rank', like someone's review rating on a movie or an application, constitutes an original work of the creator. There is a lot of experience with things like book reviews where the review, using snippets to illustrate the review, and a rating, are both fair use and the original work of the reviewer.
Google however got in trouble with their news aggregation service. And the bulk of much of the arguments there, were that the snippets were so complete on the news page as to exceed 'fair use' exemptions, and that by aggregating these pages they were 'stealing' traffic that might otherwise go to the news site. The results on those cases were mixed, with some newspapers being removed from Google's index, and others not. Generally everyone that was removed has since been replaced (at the request of the news source) because Google does drive more traffic to a web site than any other web service. So in this case while Google was found to violate the copyright of these news organizations by indexing their newspapers without their consent, the papers later found it in their best interest to give their consent.
Craigslist and Amazon and Ebay are a third kind of question. They are a collection of 'facts' (as many have pointed out) which are derived by a process (placing ads). And in the 'old' world the courts have generally sided with the person who had paid the economic cost for creating those collections. And as PadMapper and others before them have shown, is that there is a great temptation to use those same facts and re-package them into a new collection. This pretty naturally sets up a commercial tension between the original collector and the new user of those same facts. That seems to open another front in copyright litigation and policy. So if this court gets an opinion published it cannot help but be influential as there don't seem to be very many in this space. That could be because judges think the right answer is 'obvious' but I seriously doubt that to be the case.
Is the endgame for these companies to replace the content source, or to hope (or fight for) a legal precedent to 'open' CL data for third party usage?
>The case raises questions over whether Craigslist is stifling innovation or simply protecting its data
There is no question. Craigslist isn't stopping Padmapper from being built, its stopping Padmapper from using Craigslist data. The continuous attempts to reframe Craigslist's actions as an attempt to stifle innovation seem almost surreal. I just can't understand how a community of professionals could support Padmapper in this.
Edit: finished my thought that I left out mid-sentence.
As I have said elsewhere, I think Craigslist are within their rights here, but the situation still sucks. People want to browse apartment listings on a map, and Craigslist won't let us.
Does it suck for people on the other end - people with a property that they'd like to rent to others?
Why hasn't some other solution come onto the market to fix the problems of craigslist?
Are you asking why nobody has made a website that competes directly with Craiglist, but with a better interface? Padmapper allows you to post listing directly, but it's very difficult to get buyers or sellers to bother with a different marketplace when all the other buyers and sellers are already using Craigslist. You can't beat Craigslist on price, and many non-technical people are already comfortable with the existing system. Convincing them to switch to a different marketplace with fewer potential customers and a new interface is very difficult, no matter how great your features are.
Due to various factors (demographics, etc.), sellers also tend to be not the most tech savvy people. I can't believe that it is not standard procedure to post a video walk-through these days. A lot of listings lack even a photo.
Monopoly power.
It seems about time Craigslist re-evaluates creating an API. They have such vast amount of local data that others could use that they could easily charge for use and make some money there.
Is it Craigslist's data? Does Craigslist own the posts that caused women to be murdered and stolen property to be sold? If they own that data, then they must be culpable (to some degree) for these bad acts.
If they own it, they are now more than just a service provider (IMO) and can no longer be protected by the DCA. https://en.wikipedia.org/wiki/Communications_Decency_Act
For example, they do not own the fact that you are selling a toothbrush. The do own the records on their servers which catalog the fact that you are selling a toothbrush.
Why? Because they are the ones who went to the trouble of gathering and storing those records.
This is somewhat akin to recording a verbal note from a customer in a book, and storing that book in a giant library. Sure, the user owns the information, but that doesn't mean you are obligated to let every Tom, Dick, and Harry come in and use your library how they see fit- it's your library, even though you do not hold the copyright to all the information stored there.
/pedant
Note that these are free links, not paid-for API calls.
Who's free-riding whom?
It would be different if Craiglist embedded the maps from GM or MQ, but in such case CL would probably do so with permission from those companies (possibly even--gasp--paying for the right to embed maps).
I also noted that API usage would have been paid (or at least licensed), so in that sense, CL are (legitimately) freeloading on Google and Yahoo. While the argument could be made that CL are sending traffic to these sites, that same argument would apply to Padmapper.
Im sure that's why their try this lowball trick to defend themselves and show it to judge "hey we want to innovate, and Craiglist does not want to".
But sure its not their business what CL does...
Any copying, aggregation, display, distribution, performance or derivative use of craigslist or any content posted on craigslist whether done directly or through intermediaries (including but not limited to by means of spiders, robots, crawlers, scrapers, framing, iframes or RSS feeds) is prohibited. As a limited exception, general purpose Internet search engines and noncommercial public archives will be entitled to access craigslist without individual written agreements executed with CL that specifically authorize an exception to this prohibition if, in all cases and individual instances: (a) they provide a direct hyperlink to the relevant craigslist website, service, forum or content; (b) they access craigslist from a stable IP address using an easily identifiable agent; and (c) they comply with CL's robots.txt file; provided however, that CL may terminate this limited exception as to any search engine or public archive (or any person or entity relying on this provision to access craigslist without their own written agreement executed with CL), at any time and in its sole discretion, upon written notice, including, without limitation, by email notice.
When a group can interfere with how you hunt for a place to live something is up. Where you live is pretty fundamental.
What entitled bullshit. People/companies do that all the time. Padmapper could've run out of money and shut themselves down - would you pursue a court injunction on the basis that they shouldn't be able to interfere with how you hunt for an apartment? How about if they changed the UI in a way you didn't like?
Wanting the market to work isn't entitled BS. It's good civic thinking. Craigslist holding onto its incumbency, counter to the interests of the public is entitled BS.
> Padmapper could've run out of money and shut themselves down - would you pursue a court injunction on the basis that they shouldn't be able to interfere with how you hunt for an apartment? How about if they changed the UI in a way you didn't like?
The first would've been the market working as it should. Also, if Padmapper messes up its UI and goes out of business, then the market works as it should.
Craigslist holding onto its monopoly position is a broken market.
This sounds more like whining than anything based in reality. Craigslist has a monopoly on apartment rent listings?
http://www.apartmentguide.com/
http://www.apartmentfinder.com/
http://www.apartmentsearch.com/
Pedant posturing aside, to those who have "skin in the game," namely those renting and renting-out property and paying real money for leases, Craigslist is the 800 lb gorilla in most markets.
False. From the EFF: https://www.eff.org/wp/clicks-bind-ways-users-agree-online-t...
> Given the emphasis placed on a user’s assent, courts favor finding a binding agreement where the user engages in affirmative conduct acknowledging the terms of a TOS. For instance, a genuine clickwrap agreement, in which a service provider places a TOS just adjacent to or below a click-button (or check-box), has been held to be sufficient to indicate the user agreed to the listed terms. In these cases, requiring the user to click “I Agree,” after calling attention to the terms and affording the user an opportunity to review them, demonstrates the user agreed to the terms. However, courts generally do not require that you actually have read the terms, but just that you had reasonable notice and an opportunity to read them.
In other courts on the other hand in cases about copyright violations even if accuser provided IP address, logs clearly indicating defendants computer the case was dismissed because he still failed to indicate that it was in fact the defendant that downloaded and/or served the copyrighted file in question.
I'd say same way there's no possibility anyone could prove that it was I who checked the checkbox.
But please ignore me. I'm just venting.
But like padmapper, you would be completely screwed if it turned into a lawsuit.
Furthermore, the publicly-accessible data isn't being copied, it's being consumed. I may consume the data to find an apartment. Someone might consume the data to calculate the average market prices in a city. Padmapper consumes the data to generate locations on a map.
From what you've written above it sounds like you believe any contract between two parties is absurd in the extreme.
Well, if nothing else, the submitting users' copyright applies. CL has a license to use it however they like, you don't.
Since CL offers their data freely to "the public", by what theory can they prevent the public from using their own chosen client browser to view it? Why should there be a distinction between software installed in my computer and a cloud-hosted application like Padmapper?
I believe that Padmapper, as an agent acting on behalf of its users, has every right to reformat the information originating on CL as long as it does not purposefully seek to cause confusion about the origin of the data (which they do not as they link to the source).
http://www.bitlaw.com/copyright/database.html
Craigslist invested in the infrastructure to enter, store, and display this information. They should be able to set their terms as to how it is used in aggregate.
the U.S. Supreme Court ruled that a compilation work such as a database must contain a minimum level of creativity in order to be protectable under the Copyright Act.
The Supreme Court (...) held that Rural's white pages are not entitled to copyright protection, since the white pages did not meet the statutory requirement for originality under 17 U.S.C. §102(a).
Cragslist is completely automated. There's absolutely no creativity involved.
Craigslist invested in the infrastructure to enter, store, and display this information. They should be able to set their terms as to how it is used in aggregate.
Unfortunately for them, "sweat of the brow" doesn't afford copyright protection in the US.
In any case, Craigslist is completely unoriginal in its selection and arrangement of the data; the selection is "whatever people submit" and the arrangement is LIFO. I find it very hard to believe they'll be awarded copyright over it.
Whoa whoa whoa, the coding of the site didn't involve a species of creativity?
If you mean that the content of a posting didn't involve creativity on Craigslist's part, you'd be on firmer ground, but even there they made design decisions (however questionable...) about how to present it.
The design decision are irrelevant, since Padmapper isn't copying those.
So would I—phone books are not subject to compilation copyright, because there's no editorial choice going on. That was settled in the _Feist_ case.
When I upload a video to YouTube, I do expect to retain control of it solely through my interactions with YouTube, and expect that when I take it down it gets taken down across all Google properties. I would be pretty angry if I found that my videos were on another site which had scraped YouTube's data and now was hosting my video without any ability for me to remove it.
The facts themselves aren't subject to copyright.
Access through the CL site is governed by various computer use statutes. Though aggregation-via-proxy as PM are doing through Google cache raises some interesting issues.
Well, for one thing: different ideas about what constitutes fair game for property/control claims.
Not everything can be copyrighted or patented, and specifically, it's really not clear that a rental/for sale/wanted listing is actually a copyrightable work.
If that weren't enough, though, there's also not really a clear difference between what Padmapper does and a search engine -- it doesn't "steal" listings and put them wholesale on their site without attribution, it provides a geographic search that yields a limited digest and then points users to the original source.
But on the other hand, "users and developers are exasperated with Craigslist’s insistence on preserving an outdated interface and design."
I don't know the legal merits of either party's position, but from an ethical perspective, I tend to side with Craigslist here. Dissatisfaction with a commercial site's UI is not just cause for using their data without permission, particularly if they had made an effort to offer a licensing agreement, whatever the terms might have been.
In my (personal) opinion, I would argue that the content providers or aggregators and search engines benefit synergistically. CL and Belgian newspapers appear to disagree.
That's why 3taps is getting the data from google's cache without touching craigslist servers.
The two that jump to mind are Authority Labs and SEOmoz.
I guess: a shed load of proxies. :)
"I'm okay with people finding this listing through another service."
Perhaps because they don't want people to get to their listings that way even if they want to, and thus don't want your opinion? (Edit: and perhaps they really aren't as interested as they claim in making it easier for buyers and sellers to find each other?)
So, yes.
[edit for clarity, and for misspelling clarity]
While I agree the web is full of lowlifes engaged in web development, many of them in porn or some other area that appeals to base instincts, I find this comment perplexing. Because it is so subjective, yet it tries to seem objective by focusing on some random criteria.
Google employs a "bot army" to scrape the entire web. So what?
If the comment was something like "I don't like Company X." Or even "I don't like Company X because...", it would make sense to me.
But that is not how this common type of comment goes. Instead it suggests that bot=evil, i.e. any sort of automation or any sort of data collection by anyone other than [your favorite company] is "shady".
That's crazy. IMO.
It's what a company does with the data that matters.
Anyway, I'm not keen on 3Taps because they are not provinding bulk data, only API's that require "developer keys". Why?
Either you are going to democratise data, or you are just another schemer trying to find ways to collect infromation about people, in this case people using "your API".
I don't want API's I want the data. I can make my own interfaces thank you.
But robots.txt has no special legal authority, it's just a convention used to communicate a publisher's intent. I'm pretty sure the C&D letter made it 100% clear that CL did not want Padmapper crawling their site or using their data.
It does.
Look, it's obvious that even PadMapper believes Craigslist is THE data source they must have.
Google News could drop any given data source and not blink. PadMapper simply is not in the same position.
Lock-in is super common(and often unintentional or unavoidable), but it goes against the ideal of equal opportunity and competitive markets that rewards companies for a continuous commitment to quality and innovation.
I feel that craigslist has done nothing wrong except for standing in the doorway when other people who want to innovate and improve things are trying to get through (like PadMapper).
A Craigslist competitor doesn't need a huge market share to be viable. For me as a seller, it just has to expand my audience of buyers enough to justify the small amount of time it takes to post and manage a second listing.
It's not Craigslist's content. It's our content, the general public's. Craigslist is just a caretaker. Right now, it's a caretaker acting in its own interests against the public's.
Again, disingenuous. CL has the network effect. That's like Microsoft saying, Users are free to install their own browsers. Their current actions are only to preserve that, not to squash abuses. (Which is what they usually do with the power of their TOU.)
Big difference.
The real reason: Craigslist wants to hold onto their monopoly position.
Craigslist has no divine right to its users' information; it just happens to be the only viable option. If Craigslist didn't exist, someone else would, and would probably do it better.
If Craigslist wants to continue existing, they should take advantage of the fact that they are the go-to for online classifieds, create a modest subscription-based API and let the information flow.
I move to a new city every year, usually on short notice, and I often go into the process of apartment hunting with no knowledge of local neighborhoods, the public transportation system, or general geography. Without Padmapper, the process would be unbearable.
If you build it, doesn't automatically mean that they will come.
This is not to address the issue of ownership of listing data, or the case law, which latter isn't settled anyway.
The conundrum is that it is 'easy' to take an existing, quality data source, and re-skin it for broader market appeal. This is something that Craigslist should be doing but perhaps through poor management are not. It is 'hard' to take a conceptual model of skinning and building a quality data source behind it.
As I said elsewhere, the PadMapper guys might have taken this to Craigslist and said "Look what we can do with your data, lets make happy music together." and then debated the terms. Or they could take their UX to the venture capital world and say "Look at this cool product we could build if we had access to a database with Craig list's quality" and debate cap tables and dilution (assuming they got to that stage of a term sheet).
But they chose the third route, "Lets see how long we can get away with this..." and the timer on that just ran out.
Sadly, this third path makes the first two paths much harder. Craigslist already sees them as the 'enemy' and VCs will see them as a team that makes poor choices. Both views make it harder (but certainly not impossible) to consummate a deal. However, from the blog and from previous postings here it seems they made those choices with their eyes open.
But that's the thing. It isn't their data. They specifically say so in their TOS. The copyright belongs to the user.
You can't say the listing belongs to the user (to protect you from liability) on one hand, and then say the data belongs to craigslist on the other.
They could always do what map companies have done forever to prevent copyright. Facts can't be copyrighted so map makers insert fictional cities.
If someone copies the map they are copying the fictional city and thus violating their copyright.
Craigslist could add a fictional listing here and there that they do own the copyright to.
You can't say the listing belongs to the user on one hand, and then say the data belongs to craigslist on the other.
You can in fact make this argument in the case of compilations in copyright law - think a compilation of short stories or one of those "Songs of the 90s albums" - the initial authors still own their work, but the entity which creates the compilation still owns the right to the entire compiled work.
(I intentionally edited your statement - don't want to bring liability into this, I think its a separate issue and a bit of a red herring w/r/t this discussion).
Here is the relevant wikipedia section on compilation copyright
>Copyright Act allows for the protection of "compilations," provided there is a "creative" or "original" act involved in such a compilation, such as in the selection (deciding which things to include or exclude), and arrangement (how they are shown and in what order). The protection is limited only to the selection and arrangement, not to the facts themselves, which may be freely copied.
>The Supreme Court decision in Feist v. Rural..rejected what was known as the "sweat of the brow" doctrine, in ruling that no matter how much work was necessary to create a compilation, a non-selective collection of facts ordered in a non-creative way is not subject to copyright protection.
I think there is a good argument that there is no creativity on Craigslist's part in selecting postings, since they aren't selecting them--users are uploading them. If true selection wouldn't be covered.
The arrangement might be (if it can be said to be creative), but I'm assuming padmapper isn't copying their arrangement.
1) You are comparing physical media to a web service
2) You are confusing content with access to content
Since you mentioned Phonebook, let's take yellowpages.com This site has much of the same information that a phonebook does. I would argue that it would be illegal for company B to scrape yellowpages.com for this information in order to make money without express permission from yellowpages.com. However, if Company B found another way to access the same information (ex: scanning physical phonebooks or asking people to sign up to their site) then I'd say they're within the law.
Craigslist is 100% within their rights to control who accesses their service and how. If CL users want to register for Padmapper and post the same ads on both services then that's their prerogative.
Company B has found another way. They are getting the data from Google's cache of the craigslist.
Padmapper is not touching Craigslist's servers at all.
If PadMapper wanted to legally exploit the fact that the users own the listings, not Craigslist, they would have to establish a relationship with the user and get the information directly from them. They could contact the listing owners and suggest they list on PadListing. They can 'spiff' people who do so (meaning give them some benefit if they list on PadListing and the person discovers and rents through PadMapper). If they can prove that the listing owner asked them to list their property, Craigslist can't sue.
Alternatively PadMapper "need only"[1] create their own apartment listing service to have their own database where people listing apartments go to them directly. And that is a much harder thing to do than something which scrapes the listings from Craigslist and drops them on to a Google Map.
As a proof-of-concept, PadMapper is excellent. As a product, it has a data contamination issue.
[1] The scare quotes are there to acknowledge that its a challenge to get people to move from the known, to the new. I am trying to move people from Google to a new search service, its hard, its slow, but its the only way to do this and avoid this sort of litigation.
Care to share? I've been trying to move away from google, but DuckDuckGo's search results are so often massively inferior. Other suggestions would be nice to have!
Look at the Rightshaven case. The judge ruled that Rightshaven didn't have standing to sue on behalf just because the copyright holder granted it license to.
If that theory holds up, Craigslist wouldn't have standing to sue on behalf of the copyright holders.
Additionally copyright would only apply if the listings are considered creative works. Most of them are clearly simple statements of fact that wouldn't merit copyright protection in the first place.
I did a quick Blekko for the case with the Yellow Pages that was litigated this way (but alas did not find it) where a AT&T sued the maker of a competitive Yellow Pages over using the collection of businesses in their book. The defendant argument was similar to yours, that they could have walked down the street and collected the information so the information wasn't copyrightable, but the judge ruled in favor of AT&T because it was clear the defendant could have done that but they didn't do that. There was some errors and omissions that mirrored the Yellow pages and 'proved' the defendant took their data from the Yellow Pages rather than collect it themselves.
For PadMapper to escape liability they have to be able to prove they came to know about these listings in some other way than through Craigslist's collection of them. They argued in their blog that 3Taps did that for them because they got them from 3Taps they aren't liable for what ever 3Taps is doing. And my take on it is that given the case law it will be a very hard thing to prove. If Eric Goldman is reading he could probably whip out a definitive argument here.
If it goes to trial I'll definitely follow the case to see how it plays out.
In Feist Publications v. Rural Telephone Service Co, a phone directory was ruled to be protected by copyright only if the selection and arrangement of facts was an original creative act (listing numbers alphabetically was not).
I'm aware of a case after Feist where a yellow pages for chinese immigrants was copyrightable because of the creativity involved in selection of the facts and the arrangement into categories. But even then the facts themselves are not copyrightable.
Craigslist's selection is nonexistent, you send it they publish it. And the arrangement of the subset that padmapper is using is solely by geographic location and time. In addition padmapper is not copying the arrangement.
Righthaven did not have a copyright, it had a "right to sue" on behalf of the original copyright holder's copyright. The court deemed that insufficient to give Righthaven grounds to enforce the copyright, because Righthaven did not have any copyright or license therein. A right to sue is not considered a "copy right" because it involves no right to copy the material (i.e., by distribution or reproduction).
Craiglist does have a license to copyrighted content. It can actually "copy" the content. Ergo, it has the right to enforce its license against non-licensed users.
"Because the SAA (lawsuit contract) prevents Righthaven from obtaining any of the exclusive rights necessary to maintain standing in a copyright infringement action, the court finds that Righthaven lacks standing in this case,"
Craigslist ToS doesn't grant them exclusive rights, thus by that judge's definition they don't have standing to sue.
>the TOS also grants CL the right to prevent others from displaying it without CL's permission.
Just because the ToS says it doesn't mean it will work. Recently a judge said that a copyright troll called Rightshaven didn't have standing to sue on behalf of copyright holders for works that they licensed.
Regarding Righthaven, the original media company never gave Righthaven control of the copyright of their data, just the right to sue. This is why it was struck down.
As for Righthaven's lack of standing to sue over Review-Journal content, Hunt wrote the recently unsealed lawsuit contract between Righthaven and Stephens Media -- called the Strategic Alliance Agreement (SAA) -- clearly leaves Stephens Media in control of the copyrights and gives Righthaven only the right to sue.
In order to file lawsuits, copyright plaintiffs have to have actual control of the copyrights, not just the right to sue, Hunt found.
http://www.vegasinc.com/news/2011/jun/14/judge-rules-rightha...
Phone books are not protected by copyright unless there is something original about their selection or organization.
Assemblages of fact aren't protected, only the selection and arrangement of those facts. In addition the selection and arrangement has to be "creative."
The selection is definitely not a creative act on craigslist's part b/c they don't select anything, users post the information.
Craigslist posting selection is nonexistent. Therefore they are only left with "creative" arrangement for protection.
You could argue that the arrangement is a creative act. I don't think the arrangement counts because they are only using a subset of craigslist and that is merely arranged by geographical location, definitely not an original "creative" arrangement, but it doesn't matter because Padmapper isn't copying the arrangement.
See this for more information. http://www.copyright.gov/reports/dbase.html
>Regarding Righthaven, the original media company never gave Righthaven control of the copyright of their data, just the right to sue. This is why it was struck down.
The judge in the Righthaven case said this...
"Because the SAA (lawsuit contract) prevents Righthaven from obtaining any of the exclusive rights necessary to maintain standing in a copyright infringement action, the court finds that Righthaven lacks standing in this case,"
Craigslist ToS doesn't grant them exclusive rights, thus by that judge's definition they don't have standing to sue.
http://www.wired.com/entertainment/theweb/magazine/17-09/ff_...
An article in the same interest, "Why Craigslist Is Such a Mess" ( http://www.wired.com/entertainment/theweb/magazine/17-09/ff_... ).
CL just put them on the map in a big way as being "Craigslist with a good UI" -- so good, they had to sue. Stupid, stupid, stupid.
"The search tool is antiquated, the images are poor or nonexistent, locations of listings are hardly dependable, and you can forget about an integrated way to save anything for later reference. There is a litany of shortcomings that come with the Craigslist apartment search; they are many and they are painful..."
If you don't like the site, don't like the UI, and can't stand the UX, then don't use the site. Nobody is forcing you to search for an apartment on CL. Nobody is forcing owners to list apartments on CL.
And so now we get to the meat ...
"...Your site is chock-full of data I need..."
Ahh...data you NEED.
So here's the thing. If you truly need the data, then you need Craigslist and you have an obligation to use it the way they (and only they) want you to use it. If you don't like the way they want you to use it, there are several choices:
* Go work for them and convince them to improve it. * Build a better mousetrap.
The former probably won't succeed, so the latter seems to be the way to go.
Look, there's nothing preventing landlords from listing their apartments on Craigslist AND your new Craigslist replacement/improvement. They're not locked in. If they were, that would be a whole different discussion - then we can talk about anti-competitive, innovation-stifling behavior. But that's not the case - the only reason people list apartments on CL is because it works. Despite the bad UI. Hence the reason CL hasn't changed it.
So it's the job of some enterprising entrepreneur not only ro build a better mousetrap, but convince people to use it. Until that happens, as much as I too dislike the Craigslist UI, I can't say I support PadMapper on this one.
Yes they are being "forced" to list on CL, because as much as you hate the UX, UI, etc., that's where the buyers are. And as much as the buyers hate the site, that's where the sellers are. I hate CL, but I was forced to use it because lock-in makes it impossible for anyone else to compete.
Apartments aren't a fungible item. If I'm looking for an apartment in a certain area and 75% of the apartments in that area are only being listed on Craigslist, then how can I realistically "choose" to use another avenue for apartment listings?
Look, there's nothing preventing landlords from listing their apartments on Craigslist AND your new Craigslist replacement/improvement.
There are huge barriers preventing landlords from listing their apartments outside of craigslist. Nothing is as popular, so why waste the time? They don't have hours to scour the internet looking for alternatives, nor do they probably want to have to manage listings at many difference sites.
I think you have a flawed idea of what constitutes "huge barriers". The only thing stopping them from listing elsewhere is a cost-benefit analysis? I weep.
There is absolutely no reason that PadMapper cannot call them on the phone, and ask them to list with PadMapper. They can make it trivially easy supporting email, phone, or fax listings. They can sell them on using PadMapper.
Many apartments are owned by Real Estate Investment Trusts (REITs) there are dozens (not hundreds) of those. You can sell the REIT on the concept of easy listings that are so much better/cleaner/easier than Craigslist.
Craigslist was created when near every apartment was listed in the Classified Ads in the print newspaper. That wasn't "lock in" it was an opportunity, they sold these folks on lower costs (since Classifieds were a money fund for newspapers) Landlords hated the extortionate prices that the newspapers charged but they didn't have an alternative, Craigslist gave them that alternative, they moved.
So the 'answer' here is to actually build a classifieds business around rental space. That takes more than slinging some node.js and scraping other sites. Granted its 'easier' for a technical person to do it that way, but its not a 'sustainable' way of doing it.
But you are using an example in which there was a clear downside to using the existing listing model. The price. If you are a landlord and you have no problem renting out apartments in a reasonable amount of time on Craigslist, and it's free, what exactly can another site offer that is "better"?
They are already getting their apartments rented, there is minimal overhead to using Craigslist. You can't compete with Craigslist on price, unless you are actually giving landlords money for listing on your site. And trying to say "It's easier for users to find your properties." doesn't help if they aren't having a problem with renting out their properties.
This is what we call a monopoly. Many people (apartment hunters in this context) need to use Craigslist because there are no viable alternatives. Just like you have to pay your local utility for water or power -- there are alternatives, but not viable ones.
And a two-sided monopoly is a very real monopoly.
The difference is that other monopolies get regulated; Craigslist is not. Given their current behavior, I'd strongly support a law that explicitly denies craigslist any exclusive right to their listing data.
That's disingenuous. Craigslist is pretty much the only game in town in many places. Let's say you need to find a person with a specific need, willing to pay $2000 a month. If you don't find this person, you have to pay the $2000 a month. How do you feel about not using Craigslist now?
What people really "need" is the raw data.
If they want to make some UI that they like, then they can do it. If they want to offer this to others, they can do it.
If they want to load the data into some SQL database, they can do it.
If they want split the data into some other format using csplit and load into some other faster database, they can do it.
If they want to extract a portion of a raw file and just use agrep on that, they can do it.
The point is that UI is a personal decision.
Because some people do not like CL's bare bones UI doesn't give them the right to do anything. Because some people don't like bloated and clumsy web interfaces and prefer text commands doesn't give them the right to do anything either.
But people can't be stopped from making personal decisions about how to process data. Public data.
It's funny how some websites think they "own" data that is given to them. Do they "need" this data? Yes, they do.
"The company said it offered a license that would have allowed PadMapper to use its data on mobile applications but that the competitor did not accept the terms."
From the original: > They allow mobile apps to display their listings if you buy a license from them, but not websites.
It's a shame for the guy behind Padmapper, but hopefully there'll be a lesson learnt about knowing which fights are worth picking versus knowing when to move on.
Somehow Google managed to overcome Yahoo and Altavista without scraping their data. And look at something like GumTree in the UK, which has always been miles ahead of Craigslist - no first mover advantage taken up there. Plus, GumTree's design has been equally woeful over the years.
I don't know who's legally right, but I've got no sympathy for Craigslist. They've got a monopoly on the data (whether right or wrong) and refuse to either build a decent UI for consumers or let other people do so.
That sounds bad, but I think you can't know how infuriating it really is until you've tried to find housing in a town where CL is the only real option and where demand far exceeds supply (where you spend many, many hours on the effort because of intense demand-side competition).
Yahoo search would be a hell of a lot better if they could just scrape results from Google, but thankfully they lack the misguided sense of entitlement that Padmapper seems to suffer from.
But as you note, the important thing here is the "license." Craigslist offers paid licenses to app develoeprs, which PadMapper chose not to pay for.
No, it really doesn't. This isn't going to be a long, exciting case challenging the underpinnings of copyright law and server terms of services. Padmapper isn't that company.
Eric, everybody saw this coming.
Rupert Murdoch has threatened this in the past. Each time he's backed down, because he knows that being removed from Google's index is akin to being removed from the Internet.
http://www.craigslist.org/robots.txt
edit: oops! they're not actually blocking google from the apartment listings. thanks smackfu.
That's what Padmapper is doing right now. They're being sued for it.
It's not a new thing for a start-up to use Craigslist to gain an advantage (i.e. AirBnB), but they all seem to eventually be cut-off.
I believe ultimately, it is at the 3rd partys (in this case Craigslist) discretion who is allowed to appropriate their data and who not.
The case doesn't actually seem open and shut at all. CL seems to have a "moral high ground" in that they provide the original service and the platform for users to post information, but the legal ground is more shaky. It is relatively settled in copyright law (which appears to be the main thrust of CL's argument) that effort alone does not award copyright protection[1] (or indeed, any property rights in general[2]).
For example, phone listings are not copyrightable, nor is news. The general question is whether there was creative expression (which can include 'creative' presentation or organization). If I were Padmapper, I would argue that classified listings are just a collection of facts, in which case they are not copyrightable. Padmapper does not actually use the listings directly, but extracts the information and puts it in a different format (their interface). On this point I think their case is relatively strong.
Where CL might be able to make a better case is their TOU. The TOU forms a contract between CL and their users. In almost all such prominent cases, this type of contract has been found to be enforceable[3]. If this line of reasoning is followed, then Padmapper will probably lose on this point because they agreed not to "copy, aggregate, distribute", etc.[4]
If this is a contractual claim the damages would be what CL could have expected if Padmapper had not 'breached' the contract - in this case I'm not sure that amount is particularly large, because Padmapper redirects to CL's site for the actual listings.
TLDR: copyright claim seems weak, but the contractual claim might succeed. The trademark stuff seems like a red herring and is CL throwing stuff at the wall to see what sticks.
[EDIT: see below, I overlooked the fact that Padmapper is no longer using CL data directly, so they might not be bound by the TOU at all. Same goes for 3Taps since they use Google's cache. Looking stronger for Padmapper]
[1] Feist v. Rural, http://en.wikipedia.org/wiki/Feist_v._Rural
[2] INS v. AP, http://en.wikipedia.org/wiki/International_News_Service_v._A...
[3] http://en.wikipedia.org/wiki/Clickwrap
[4] CL Terms of Use (see section 3), http://www.craigslist.org/about/terms.of.use
My problem is with section 3. CONTENT AND CONDUCT, part a, where they first state "CL does not control, is not responsible for and makes no representations or warranties with respect to any user content." Then it goes on to state that posters assign a bevy of licensing rights to CL for the use of the content.
How can they claim copyright to content and at the same time aver that they don't control it?
So the first part of that statement is akin to waiver, so that someone can't sue CL for what the listings say. The second part is a claim on copyright and licensing, which they may or may not have.
But craiggers provides a very valuable service, namely showing pictures of items as you search. It made furniture searching so much easier.
I have zero sympathy for craigslist. They didn't get into their position through "years of hard work". They won the lottery (someone had to win it). They have done zero innovation (other than maybe anti-spam) for years and prevent others from innovating on top. Society wins if they are forced to open their data.
Name one startup that began knowing they were going to be sued and went on to a successful exit.
Parenthetically, the original Napster, the one that began knowing it would be sued, was sold about 3 years after its founding for less than $2.5 million as part of bankruptcy proceedings -- and I doubt that the owners took any significant dividends out of its before its sale: http://en.wikipedia.org/wiki/Napster#Current_status
Some do it to prove a point, some do it to prove that the market exists.
Craigslist is one of the poorest user experiences around and succeeds only because people insist on using it. The alternatives somehow manage to be even more spectacularly useless by over-designing their apps and cluttering them up with junk.
Padmapper is one of the few that does what they're supposed to do, and it's not even an optimal implementation of this sort of thing.
Padmapper is (indirectly) screen-scraping Craigslist's data in an attempt to unseat Craigslist.
Unseat Craigslist as the undisputed champion of what, exactly? Cartographical apartment hunting?
Craigslist didn't write a single one of those posts. They try and assert control over this content through their terms of service:
"You automatically grant and assign to CL, and you represent and warrant that you have the right to grant and assign to CL, a perpetual, irrevocable, unlimited, fully paid, fully sub-licensable (through multiple tiers), worldwide license to copy, perform, display, distribute, prepare derivative works from (including, without limitation, incorporating into other works) and otherwise use any content that you post. You also expressly grant and assign to CL all rights and causes of action to prohibit and enforce against any unauthorized copying, performance, display, distribution, use or exploitation of, or creation of derivative works from, any content that you post (including but not limited to any unauthorized downloading, extraction, harvesting, collection or aggregation of content that you post)."
The last part of that may not be enforceable. They're asserting that they are singularly responsible for "authorizing" reproduction despite not having copyright for the material in question.
http://arstechnica.com/tech-policy/2011/09/righthaven-copyri...
Craiglist does have a copyright use license, and thus has the right to enforce the license it has been granted.
So, to paraphrase, it only succeeds because something about it attracts customers. As opposed to other sites, which succeed by fairy dust and rainbows.
If padmapper has a great user experience, that's wonderful -- They should be able to get customers to post their listings on it, and generate data without relying on Craigslist. Or, if they rely on Craigslist, and what the article said is true, they could have negociated a license to the data, instead of scraping it.
It's a difficult nut to crack. Kajiji seems to be gaining some traction, but it's trading one set of UX nightmares for another.
Craigslist controls a two-sided market. If I build a new site, I can't get buyers without there already being sellers. And vice-versa. That's an incredibly difficult business problem.
Github seems to have had it easier. I can just host my new project on github and be done with it (my website links to it after all). I'm sure sourceforge had some value just by being the go-to place, but how strong was the network effect from a developer perspective?
SourceForge and its related properties were the backbone of the early web, supporting a number of important efforts to which people felt a strong allegiance. It was like a benevolent force at the time.
SourceForge had, at the time, a fairly formal process for registration. They considered themselves more like a library where getting shelf space was a privilege not doled out lightly. This is not unlike how getting into the Yahoo! directory required a lot of begging and pleading.
While this meant that most of the projects hosted by it had a lot of merit, those lesser efforts were left out in the cold. They failed to switch to a more casual model as the "Web 2.0" philosophy started being the dominant mind-set, where expectations shifted dramatically from carefully curated content to emergent user-driven communities.
Also worth noting, GitHub's pace of innovation is so far beyond nearly anything else in the industry that it was only a matter of time before they became the superior platform in terms of technology.
Additionally they were able to ride the surge of popularity that git was gaining, something that SourceForge didn't support at the time, and persuaded a number of high-profile projects to move to them. The real coup was Ruby on Rails, which once hosted there, solidified their position.
Any Craiglist displacer would need to swing a few important, strategic deals to cement it in the minds of people as a reasonable alternative. The rest would be a case of just driving harder than Craiglist is willing to keep up with.
If a lot of people INSIST on using it, then that means it has good UX. Complaints with their UX should be qualified.
For instance, many payroll companies have truly awful sites, yet people use them since they have no other choice.
As a specific example, searching for an apartment should not mean reading every single listing in order to discern where the property is, what features are available, and what conditions apply. That people post the same listing repeatedly until it is rented is not helping matters.
eBay, which is arguably still quite primitive, has much better filtering options, alerts, and a single listing that persists until the auction is completed. Craiglist is a series of random, repetitive posts with only a minimum of categorization.
Padmapper provided a superior way to navigate listing data and present it in a much more meaningful presentation. Being able to "favorite" listings helps when filtering.
Also, don't confuse my opinions about Craigslist's design/UX with an opinion on the case. I've used Padmapper extensively. While I believe that Craigslist has a right to do what they please with their content, I don't see what Padmapper was doing as wholly legally wrong.
It may look butt ugly, but it is as close to a functionally efficient classifieds board as you can get. Craigslist isn't trying to be anything more than that.
This doesn't mean that PadMapper is right, from a legal standpoint, but I think it's fair to say that craigslist's UX is terrible, for apartment/house searches (where small variations in location matter a LOT) at least.
Please use the proper legal word - "Monopoly". I wonder how the Department of Justice decides who to prosecute?
Padmapper is the best apartment rental interface out there. The author has been single handedly working on this for 3-4 years and constantly improving his system.
Legally he might not have been in the right but we should be helping him out not uselessly debating and casting judgement.
If theres anyone who exmplifies the true spirit of Hacker News it would be the author of PadMapper.
If he's not "legally not in the right," why should I be helping him (doesn't that imply casting a judgement - that PadMapper is elevated above CraigsList because it's "the best apartment rental interface our there")?
What if the tables were turned and CraigsList was violating PadMapper's TOS? Would you be willing to help CraigsList (they serve more than only apartment listings)?
Should somebody be rewarded, even if (legally) not in the right, just because he/she is a hard worker/smarty pants/exemplifies the True Hacker Spirit?
I often wonder if there is some sort of unwritten clause that You-Must-Pull-For-The-Underdog to join Hacker News? (In the name of Divine Disruption, of course.)
Because legal right and moral right sometimes intersect, but not nearly always.
Because once in a blue moon somebody comes along and creates something with such enormous public good that it forces us to reconsider the context of both legal and moral righteousness.
Personally, while I'm not a big fan of Craiglist's UI, it's stretching it to think that PadMapper would be classified as "such enormous public good" simply because a small subset of total Craigslist (and Internet) users (most likely in a small subset of American cities) are fond of it.
In particular, read the comments by Craig himself. Funny how these things change.
I guess that wasn't true.
Are there any Craigslist employees reading this? Would any of you like a full time job sorting through the cesspool that is your apartment listings page? Oh, you wouldn't? I don't blame you, because that's my job right now, and PadMapper is the only thing helping me preserve what's left of my sanity.
Do a better job than them or shut the fuck up. Stop hurting people by trying to shut PadMapper down.
The point is that simply defending against the lawsuit could kill Padmapper. Lots of time, money, and energy go into lawsuits. Conveniently, those are three things startups don't typically have in excess. The outcome doesn't matter for Padmapper if Padmapper dies during the defense.
STEALING
transitive verb
1a : to take or appropriate without right or leave
and with intent to keep or make use of wrongfully
http://www.merriam-webster.com/dictionary/stealingIf you're going to downvote me, that's fine, but if you do, please show how copyright infringement does not fit within the definition of stealing. I'd like to settle this semantic debate once and for all.
edit: Ok, I won't be so cryptic. Padmapper doesn't copy the posts on Craigslist, they link to them. That's not a semantic quibble. All search engines do so. If I blogged that I was renting my apartment and linked to the Craigslist post describing it, I would be doing exactly what Padmapper is doing.
However,
If I blogged that I was renting my apartment and linked to the Craigslist post describing it, I would be doing exactly what Padmapper is doing.
This doesn't line up in a few ways. Padmapper is a third-party, not the original poster. There is also a distinction between doing it as a one-off and doing it in an automated fashion.
Padmapper needs to make it trivial for people to "upload" their listings from CL to PM. It could be a single bookmarklet.
Now, suppose that Alice tells Padmapper about Bob's apartment for rent on CL. That's getting even weirder, but PM might be okay if they follow DMCA safe harbour procedures.
It doesn't seem, however, that the link is the crux of CL's complaint. Rather it's the fact that PM has copied CL's data in order to generate a mapping of physical locations to CL listings.
Padmapper does link to images, but Padmapper would be useless if it didn't also re-publish the content jacked from Craigslist.
That was my main point; it's fairly useless for non-attorneys to argue legal cases on HN (my apologies if you're a lawyer and I'm just wasting your time).
AFAICT, Padmapper is doing something very much in line with what journalists and search engines have traditionally done. That point of view may not prevail in court, but I think it's completely defensible on ethical grounds.
So the dictionary agrees with the common use of steal in the realm of intellectual property. It's really the proponents of copyright violation who have redefined stealing to exclude copyrighted digital media.
The people who will whip out dictionaries over "steal" are crudely trying to inject an emotional response into the debate and trying to override rational discourse with rhetoric.
The people who are constantly whipping out "you still have it" are naively pretending that, just because it doesn't fit within all the criteria of physical theft, it also means that the other moral concerns associated with it are equally moot.
tl;dr: Both sides of the "stealing!" debate are being incredibly disingenuous.
Does the data belong to Craigslist or to those posting on Craigslist?
If so, is this in their TOS? I don't remember ever clicking a checkbox that says "we own your listing" whenever I have sold stuff on Craigslist or posted jobs/gigs. Maybe real estate is different.
Why would I want to give Craigslist ownership of my listing? If I were posting a property for rent I would probably want other services to pick-up the listing.
Similarly, why do they own the Copyright on something I have written?
Again, I don't remember ever giving that away either.
So, I haven't verified it, but the idea seems to be the user owns their post, and could post it elsewhere if they wanted - but Craigslist has the right to stop other sites copying posts direct from Craigslist.
That's what makes this a little less open and shut, at least to me.
"Craigslist is not only gigantic in scale and totally resistant to business cooperation, it is also mostly free. The only things that cost money to post on the site are job ads in some cities ($25 to $75), apartment listings by brokers in New York ($10), and—in a special case born of recent legal trouble—advertisements in categories commonly used by prostitutes, because authorities encourage vendors to maintain a record that would aid investigators. There is no banner advertising. They won't let you join them, and at this price you can't beat them either."
http://www.wired.com/entertainment/theweb/magazine/17-09/ff_...
Usually, when companies get mad about scraping, it's because it's either bypassing the ads that make them money, or in rare cases bypassing a premium API. Here, PadMapper is doing neither.
In fact, if Craiglist is so committed to not updating their ugly website, why don't they just license out a public API, killing two birds (new monetization strategy and having a better interface for free) with one stone?
Though to be honest that makes no sense, because that's exactly what they're getting from the third party site, and they're paying for that...
Edit: Not sure why this comment is getting down-voted, but if I offended some one that was not my intention.
PadMapper is basically a pointer to a listing on Craigslist (when Craigslist listings are the reference), where you can search the pointers with filtering that you can't on Craigslist. In order to get to PadMapper, you have to visit the site directly.
Once you have a web site that attracts buyers, you can replace craigslist data with your own listing service and stop paying for their API.
They clearly don't fear this kind of competition in the mobile space, hence offering licenses for mobile-only API use.
Also, I think a lot of complainers are not aware of what it takes to roll out UI design changes to hundreds of millions of users. Frankly, I'm not very knowlegeable myself, but I can stop and think about it for a second and imagine there are huge logistical issues (performance, browsers, etc.) and huge psychological issues (non-technical users, familiarity, etc.)
And then there's the whole question of whether CL's UI should be improved. There are clear benefits to having a simple UI that does not do everything for the user. When I used it to find housing near SF, I copy-pasted addresses to map them, and I did a little extra work, but I never found that CL was actually in my way. For neighborhoods and recommendations, I asked real people. When I looked for a car, I made a little spreadsheet to give weighted scores to the features I wanted--I did not go around whining that CL didn't do everything for me.
My impression is that CL wants to be a basic classified service. For whatever reason they don't want to morph into a specialized listing service, and I can understand and respect that. After all, the CL interface is not that much different from HN's.
PS: I've had this item open in my browser for maybe an hour or two--why does the Add Comment button turn into a "dead link" OMG such a shitty UI, if they won't fix it, I am justified in rehosting all the content on a better site.
edit: especially one that doesn't allow you to use their data in the first place.
How about Padmapper develop a plugin for chrome/firefox that once you go to a craigslist appartment list page, it does what it does now: loops through them all and displays them on a google map. Padmapper could even develop on API on their end which handled all the hard work of actually finding the correct address of the posts and sends it back to the plugin in a nice format and all it has to do is plot in on a map that the plugin places on the craigslist page.
This sounds legal to me, as Padmapper isn't actually storing anything, and it puts a UI on top of the apartment search page.
RECAP is on a better legal footing because the filings are public domain instead of owned by someone.
Second, it's not infromation that's protectable by copyright (at least not in the US). It's like the telephone book as others have said.
There's heaps of websites that mistakenly believe they have some copyrights to assert against users that harvest data. But it's not just websites. A longstanding, classic example of this sort of silly argument is WHOIS. The registries can't stand the thought of others getting the data. They claim bandwidth is an issue. Yet they won't provide (mirrored) bulk data, which would easily resolve any such bandwidth claims. They purport to "license" the right to do things with the data, yet they have no rights to grant to begin with. It is public data.
Then there's the "spam issue". As if email is somehow different from direct mail, and from telemarketing. Why don't we outlaw those practices?
Anyhow, again, it's public data: addresses and phone numbers. We call it the telephone book. If, e.g., a brokerage wants to have its summer interns do cold calling to solicit new institutional investors, the interns do not need a "license to use the telephone book". If someone listed in the phone directory does not want to be called, there's no law that says they have to be listed; they can use a nonlisted phone number.
People who list with CL presumably want their information to be found and to be presented in an appealing way. Maybe not exclusively on CL. Maybe CL should ask them what they would prefer.
Whether or not Padmapper is in the right here, CL's long-standing stance in not evolving their UI or creating a decent API is going to bite them in the ass. Helping people sort your content is now more important for them than the idea that change is bad or time-consuming. As many people have said in the past, it wouldn't take much to really bring CL up-to-speed, but if they don't want to do it themselves, they should at least make it an option for others to pay a fee to use the data.
Edit: Didn't realize http://www.padlister.com/ already existed. Score.
The more people that use PadMapper and find CraigsList entries on there, the more people that are using CraigsList.
Why wouldn't they want that? If I were a real estate listing company, I'd want my data to be found on as many real estate search apps as possible.
Most programmers and designers have made peace with the idea that what's public on the internet is public for the whole internet, but I think most people perceive this kind of information leakage as a lack of control. I think that Craigslist, more than most companies, has a real understanding of this, and a lot of their restrictions that may seem arbitrary to entrepreneurs come out of it.
https://www.google.com/search?btnG=1&pws=0&q=iphone+...
They even have cached thumbnails and you can likely use their custom search API to get results, snippets, etc.
Is what 3taps does that different?
Boo on Craigslist! They claim that they're protecting their users, but what harm is padmapper causing to people listing properties on Craigslist? Getting them more potential renters?
- The Craigslist apartments listing interface really, really sucks.
- Padmapper's search interface is much, much better.
- CL almost certainly have legal recourse to control use of their data and trademarks, or at least make things miserable for PadMapper should they choose to do so. Which would put a crimp on PM's future growth prospects as a lawsuit magnet.
- CL undoubtedly add value to their listings by filtering (through community input) for misclassified and spam listings.
- I really wish the two organizations could reach an equitable compromise which preserves the PM interface concept. Licensed use. Buyout. Whatever.
- Barring that, CL are ripe for disruption. My read is that PM have enough mojo at this point to make it on their own by licensing listings from other sources and soliciting them on their own.
Its no secret that the CL community have frequently requested an update to the UI as well as many additional features to be added in order to make using the site much more friendly; although with no changes being implemented at this point. I look at this litigation as a foreshadowing of what may be in store for the CL community.
With that said, shutting off access of those parties utilizing their services without proper consent may be the first in a strategic battle plan for the CL team in order to ensure they have the most optimal acceptance of their revamp.
Anyone else see it this way, or am I indulging wishful thinking?
That attitude is horribly wrong. Agreed, CL is a monopoly. Agreed, they are holding back innovation. But, you don't disrupt a business (and build yours) by copying their data, especially illegally. Although, the legality of issue at hand is up in the air right now, the general attitude is a bit appalling.
What's required is a way for users to post their listings simultaneously at other properties. The content of the postings belongs to the user and they should be free to re-post it anywhere.
If people had a secondary Craigslist portal they could use to access the same data, while avoiding the shitty interface, more and more might migrate over to the new 'design'. The current Craislist design has been dated for some time, and I just can't see a startup or even a big player not moving into their space with a better design and interface.
Padmapper is probably getting enough publicity out of this to get over the critical mass problem and will be able to rely on their own listings.
Craigslist content is user generated, and the copyrights remain property of the users. While facts aren't copyrightable anything that padmapper displays that isn't just a fact, photos or a realtor's description for example, is a copyright violation.
This is why sites like stack overflow are awesome, because they licence "their" content, aka your work cc-wiki and claim no ownership.
The bummer here is not that craigslist is suing over unauthorized data access, it's that your work is too restrictively licenced on their site to be of maximum benefit to you.
The relevant part of the act would be: "Intentionally accessing a computer without authorization to obtain ... information from any protected computer". A "protected computer" is a computer "which is used in or affecting interstate or foreign commerce or communication", which fits Craisglist pretty well.
Violation of the clause I quoted is a criminal offense with potential jail time.
All of this anti-Craigslist sentiment makes me sad.
Pad mapper was sent a c&d and decided to not comply. Hence the case.
Edit: and as has been pointed out, google isn't competing with Craig's list.
Padmapper's gameplan is to disintermediate Craigslist away.
Well then go use / collect your own apt listings, lobby congress to pass compulsory use licensing for all entities that collect data, or STFU.
Become a for-pay API.
Thoughts?
They already offer an API for mobile.