Twitter to Sell 50% of All Tweets for $360,000 a Year Through Gnip
readwriteweb.com
readwriteweb.com
[1] That's a hard drive manufacturer terabyte, not a real one
That, or it will be like email, where everyone has @example.com appended to their username.
Neither situation strikes me as desirable.
Processor speed, disk capacity, network bandwidth, and available software are all growing much more rapidly than online populations.
In some years' time I might be able to run an operation like Google, Facebook, or Twitter from my bedroom.
Edit: wiki says they already support this: "Supports Federation, which provides the ability to subscribe to notices by users on a remote service through the OpenMicroBlogging protocol."
Federated twitter clone, powered by status.net
"What's happening right now?"
This question is worth a lot of money, and something that doesn't have a good algorithmic solution(e.g. Google News.) Twitter is probably the only company that has a privacy-compliant solution to this, hence making it a very monetizable product.
The value would be back into the API and not into weird sponsored trendic topics. Twitter seems to go back towards Alex Payne's vision (data hose platform) and away from Biz Stone's (twitter as a media with celebrities etc...). They could also set up separated Twitters Hoses: like there could be automatic sensors data input for the Internet of things for instance, separated from human input. Any link where Twitter guys are speaking of this?
Twitter would notice a major scraping operation, but if it's done correctly they wouldn't be able to distinguish between user IPs and bot IPs.
edit: Barracuda already did more than 10% of users just for a white paper: http://www.barracudanetworks.com/ns/news_and_events/index.ph...
150,000,000 registered users only takes 170 days at 10 users a second for a first pass. Focus on frequent tweeters for subsequent scrapes. Even among the ~20% of twitter accounts that are active, most don't need to be scraped daily, and the most active accounts are likely spam.
translates to "Scrape every user" unless you know of some magical way to get list of "active" users.
Guess howmany active users there are? Guess how many servers you need running to get through those in 1 hour. My guess is something much more than $360,000 worth / yr.
look for users who tweeted in last X days? also look for their repliers+buddies, since they too are likely to be active. doesn't seem hugely complicated to me?
Requires you to check aka scrape every user to see which ones tweeted in last X days.
Let's do some math, these are all based on numbers from this past June which have likely only gone up since then [1]:
65 million tweets per day / 20 tweets per page = 3.25 million page views per day
Just to keep up with the stream, you'd need to do about 3.25 million page views per day or a little over 1.5 million to get half of it. Again, I'd find it very hard to believe that nobody at Twitter would catch on.[1] http://techcrunch.com/2010/06/08/twitter-190-million-users/
Anyone here from people genuinely ready to spend that kind of cash for 50% of Twitter?
In fact, I find it kind of refreshing that Twitter is just flat out saying all your base belong to us, and anybody that wants it gets access for $360,000 a year.
This in contrast to Google, Facebook, Yahoo etc that muddy the waters whenever it comes to how they're actually using / sharing your data.
There's a difference between the perception of your feed being public, and Twitter selling your feed data to a corporation to utilize for targeting, advertising, reselling to employment agencies in the future long after your stupid teenage profile is deleted, whatever.
When a solid 10% of the users are kids (and a much much larger percentage is entirely clueless) it's worth questioning and the people that do know what's actually going on have a responsibility to ask questions.
That's my best attempt to get fired up about privacy on twitter, you forced my hand - oh won't you please think of the children?
Ref twitter age - http://royal.pingdom.com/2010/02/16/study-ages-of-social-net...
2. What about the people whose Twitter accounts are private?
3. The people 'racing' with each other for followers or retweets are by definition more public than most other people. Unless you are going to claim that all or most Twitter users fall into that category. Trying to use them to categorize the user base of Twitter as a whole seems a bit off.
4. Twitter is a broadcast medium, but what we are talking about are the perceptions of the people using it, not the reality of the situation. There are plenty of people that broadcast stuff publicly that they wouldn't want their parents to read. Why would they do so? "My parents aren't on Twitter." I'm sure the same thing applies to bosses and the workplace.
In most cases it's not at all, here's a good page on the topic: http://www.canyoucopyrightatweet.com
And if Twitter wants to make an extra (maybe morally gray) dime off of any "privacy outrage", they can offer certain users to pay a fee to have their tweets NOT included in these dump.
By comparison, Google's data of my searches and Facebook's data of me and my friends is much more intimate than Twitter's database of my tweets.
Ryan Singer nails the distinction here:
http://37signals.com/svn/posts/2618-twitters-ux-separate-the...
"Public by default is better than public-by-surprise."
I also think that a research effort with the ability to process 100% of Tweets can most likely afford to pay. Something like the 5% though, I think a great deal of research can be done on the Twitter platform may be prevented because of the cost. For research maybe you could charge the bandwidth required to deliver the stream, no idea what kind of ballpark this would be in.
Processing 100% of all tweets is not actually very hard or expensive: 2000 messages/sec is small potatoes if you want to do it in realtime, and even easier if you are just doing batch analysis queries. You could do it (with reasonable performance) for much, much less than $360,000/year (let alone 2x or 3x that for the 100% feed).
Previously some serious time would be spent scraping content from the feeds but given that it's only 2% of the content and would take months (at least) to gather a significant amount makes it less than ideal. Although I would assume for a vast majority of cases, months are worth less than $360k.
P.S. if this just a "display" does it mean there is no any value to the stale links?
> Customers will only be allowed to analyze the messages, not display them