http://metmuseum.org/api/collection/collectionlisting?offset...
and increase 'offset' by 100. The JSON output contains image URLs. The total number of results is 441048, so finding another endpoint that doesn't enforce a limit of 100 on the 'perPage' argument would be great.
---
EDIT: thanks to spitfare, I updated the perPage argument to 100. The site doesn't allow larger values, but that's definitely a great start! (about 4k requests to get everything)
The repo doesn't include the images, which is understandable; however, the CSV file doesn't even include links to the images.
This Perl script should get most of the images (but it would probably be better to save them in the original sub-directories):
use strict;
use warnings;
use LWP::Simple;
use JSON qw(from_json);
my $url = "http://metmuseum.org/api/collection/collectionlisting?offset=";
my $args = "&pageSize=0&perPage=100&sortBy=Relevance&sortOrder=asc";
my $offset = 0;
while ($offset < 5000) {
my $decoded = from_json(get($url.$offset.$args));
$offset = $offset + 100;
my @results = @{ $decoded->{'results'} };
foreach my $i ( @results ) {
my $filename = $i->{"largeImage"};
my $title = $i->{"title"};
print "Status: ".getstore("http://images.metmuseum.org/CRDImages/".$filename,$title.".jpg")." ".$filename. "\n";
}
}I'm particularly interested in the paintings for use as wallpapers, but there are lots of printed materials dating back to the 17th century (and probably earlier).