Good resource - I've been using BeautifulSoup[1] for the scraper I set up for my needs, and it's probably worth checking out as well!
>>> from lxml import etree
>>> etree.LIBXML_VERSION
If you find a page in the wild that it can't handle, I'd love to know the URL.