Genuine question from a non-programmer: why? Is it because the volume of requests increases load on the servers/costs?
That's the great thing about HtmlAgilityPack, extracting data from HTML is really easy. I might even say even easier than if I had the page in some table-based data system.
Not quite. Many Wikipedia infoboxes (and some other templates) use standardised class names from microformats such as hCard:
Bulk downloads (database dumps) are much cheaper to serve for someone crawling millions of pages.
It gets even more significant if generation of reply is resource intensive (not sure is Wikipedia qualifying for that but complex templates may cause this).