That blog post doesn't state that _all_ collected data through Crowdsource will get published.
Another negative would be release of Bad Words or illegal content submitted by malicious users. Depends on the task.
But the actual raw data would be of more use to researchers than one cleaned from an output of algorithms. Perhaps there could be a program for educational researchers?