Also, minor detail, but the images should also be rotated to their proper orientation. Crowdsourcing data collection has to be as frictionless as possible, and this is an easy fix.
Depending on how many actual documents there are (i.e. how many pages are in those 200 folders), it might be worth it to go the route of ProPublica's "Free the Files" project, in which they built a mini-app that let people voluntarily transcribe the important fields in each document:
https://www.propublica.org/series/free-the-files
Their Al Shaw wrote a piece about designing for efficient crowd-sourcing:
http://www.propublica.org/nerds/item/casino-driven-design
They even open-sourced the Rails plugin for it:
http://www.propublica.org/nerds/item/transcribable-free-the-...