Twitter just updated its robots.txt to exclude all scrapers
twitter.com
twitter.com
http://webmarketingschool.com/no-twitter-did-not-just-de-ind...
https://twitter.com/robots.txt
They blocked robots on their marketing pages.
https://web.archive.org/web/20150715164726/https://twitter.c...
What they did was some perfectly legitimate duplicate content protection.
Will write it up in a bit more detail...
https://www.twitter.com/robots.txt https://twitter.com/robots.txt
This is likely just to prevent content duplication/nudge users to visit without the "www".