Bypass Paywalls – A Firefox extension to bypass paywalls of many news sites
addons.mozilla.org
addons.mozilla.org
I imagine search engine could offer publishers an API to send content for indexing without having to publish it. Google and others could even charge publishers for that and then published could ask readers for subscriptions.
Publishers, if you want to have people pay for your content, make honest paid subscriptions and deal with it that you vanish from the openly accessible web.
You're imagining that it can't work because it can't work in a generic way with all search engines, including future ones that don't exist yet. But that doesn't have to be the case. Most content providers only care about Google.
"Cloaking refers to the practice of presenting different content or URLs to human users and search engines. Cloaking is considered a violation of Google’s Webmaster Guidelines because it provides our users with different results than they expected.
Some examples of cloaking include:
...
- Inserting text or keywords into a page only when the User-agent requesting the page is a search engine, not a human visitor
" [0]
Google is happily showing LinkedIn, FB, pinterest and news sites content. But when I, Joe User, go to the page, I see nothing but some login/register/pay now form. How is this not a violation of the cloaking guidelines? Clearly google is getting different content than what I am!
(Presumably this is how article's extension works... by masquerading as GoogleBot -- again proving that these sites are serving up different content)
[0] https://support.google.com/webmasters/answer/66355?hl=en
edit: formatting
Do you have an example?
This is a pretty good model I think. It’s accessible and sustainable for both publishers and readers.
I know it's adding complexity and most people don't care, but I wish they used a system of crypto vouchers (not like cryptocurrencies, more like Mozilla Persona, which allowed identity providers - in this case, Blendle - to vouch for the user without knowing to whom they were vouching).
Is there an ethical journalistic entity that does not run on ad revenue, but only paid for by readers?
Your statement comes off as really entitled.
Paid journalism (by the consumer) is getting harder and harder. Do you have a good example of journalism today that is paid for by the consumer (i.e. monetarily) that is good journalism?
(And I consume a wide variety of sources to specifically try to pick up on bias, and oh boy!)
I developed a system for subscription payments that (1) is resistant to falsifying visits and (2) preserves user privacy. I got a US patent late last year[0], so that I might enjoy being a bit of a lightning rod for downvotes.
The gist is that the publisher can ask "is this person a member of the subscription pool?" and get a yes or no. And that's all they can get. I can't easily prevent them from using other mechanisms but I can refuse to make the problem worse.
As much as I'd like 'information to be free' we do not yet live in a 'Star Trek' world where money no longer has any meaning, and where people produce 'stuff' purely for the benefit of themselves and others.
Until we reach that utopia (or dystopia depending on your view) you can't have your cake and eat it; It's Ads or Paywalls, pick one.
But... bots (google, etc) are allowed to get through. I understand the reasons, but then, is it really a paywall? To be honest, I won't bother with cookies and other tricks: logged in, allowed, logged out, walk away.
In the end, I agree: don't bypass paywalls, the whole adblocking argument becomes invalid, if we do. Some publishers are trying to find a new way of income, and bypassing that invalidates the efforts.
(EDIT: treating bots differently is wrong on so many levels. I changed the User Agent of my miniflux RSS reader to "miniflux-legacy (Googlebot for Tumblr)" from "Miniflux (https://miniflux.net)" because otherwise Tumblr shows GDPR cookie consent interstitial for RSS feeds. There really, really shouldn't be a separate version for bots.)
The existing model is to treat the Internet as a magazine rack. It doesn't work that way.
A user might read one article from the WSJ a month. Obviously they won't pay 5 USD for that, they'll either attempt to get around the paywall or not bother.
Giving users a few free articles means that they'll just rotate around sites to get what they want. You need to charge a small amount from hit 1.
Not withstanding that it's not hard to build a paywall that actually functions, just don't send the content unless you've paid. This addon relies entirely on the fact that content which has not been paid for gets sent anyway.
Realistically though, the answer is that paid journalism disappears, or the Internet as we know it disappears. Increasingly lately it's looking like both will happen.
Seems a touch hyperbolic. My question would be: what part of the internet depends on paid journalism and how would the disappearance of paid journalism effect e.g. buy stuff online, looking at pornography, or accessing social media?
I know this is unacceptable to those who have a moral opposition to accessing content without accepting the rules, but to those who are just concerned about funding journalism, this approach is essentially equivalent to not bypassing paywalls.
Also, there's a problem with paywalls, which is the enhanced tracking (by subscribing, you're giving them an excellent profile to attach to their analytics). Some people might therefore prefer to use a Bypasser even if they are subscribed.
But news flash: you are not entitled to something if you don't want to pay for it.
In both options, the outcome for the site is exactly the same. So, not being a masochist, why shouldn't I choose the outcome that benefits me without harming anyone else?
It's the same argument for piracy, "if I make a copy, I'm not taking away the original, so I'm not stealing", which is just a lie to tell ourselves to feel a bit better at the end of the day.
Regarding the bandwidth costs, those are absolutely negligible, since the methods for bypassing the paywalls either only fetch the text (e.g. http://outline.com) or cache it on other servers (e.g. http://archive.is).
It is definitively not what they wish, but not complying with someone's wishes is not harming them.
--
On that last point, while I don't take my moral code from the law, I note that it agrees with me; for example, even if you have a contract with a guy to build your house, and you say your want a certain brand of pipes, but he ends up using another, you're not entitled to anything. Unfulfilled wishes are not harms.
Source: https://github.com/iamadamdev/bypass-paywalls-firefox/blob/m...
The "Referer" header is an interesting question. Hypothetically, if there's be some kind of renumeration for the site (eg: "I grant access to all of your users to my site and in exchange, for every referred visitor, you pay me 0.01 cent"), then spoofing the Referer would incur a financial cost to Google/Facebook. The incurred cost would be negligible for the individual user, but a case could be made against the extension publisher for the aggregated costs.
What?
Wallace and Gromit in a case of... the Wrong Bits!?
Also likely to be punishable under DMCA as circumventing access controls.
Of course, the ads revenue model for the Internet only happened because people basically want free labor, being unwilling to pay for the content they consume, screw the publishers the world will survive.
It's sad because what we'll get is legislature and/or more DRM.
And this makes me think that the existence of DRM is justified.
If there is just a banner hovering over the actual text, and the extension merely removes that banner, than one could question whether there even was an access control in the first place.
As an extreme example, a Finnish court ruled that CSS (as used by DVDs a long time ago) was ineffective.
https://www.turre.com/finnish-court-rules-css-protection-use...
[0] https://dejure.org/gesetze/UrhG/95a.html
[1] https://www.wbs-law.de/it-recht/verbreitung-einer-anleitung-...
In the case of this addon, the paywalls are often just overlays that you can also remove manually with a few clicks.
Ie, the anti-anti-adblock is actively interfering with the site's function on the client side.
Cookies and referer are handled on the server-side and outside the user's control.
I would question if the last sentences of 95a apply to setting a referer and clearing cookies. They are more closer to tampering with computer system (263a).
It's very easy to argue that this protection that is trivial to circumvent is not "wirksam".