Screenshotting or PDFing of a website is an increasingly important archiving tool, to supplement wget. I've come across a lot of websites that won't render any content if not connected to a live server.
https://medium.com/@dschnr/using-headless-chrome-as-an-autom...
- shift+f2
- screenshot urcodiaz.png --fullpage
A better option is to render a page with JS turned on and save the resulting HTML.
However, as you point out, PDFs are designed around the printed page, not the flowing arbitrary-page-size documents of the web.
(execute JavaScript on the page is most likely possible?)
It came out of a screenshot/archiver I've been working at at Mozilla, but I've split it up as the screenshotting is shipping and DOM archiving is still way outside Mozilla's comfort zone.
I.e. DOM copy > screenshot > wget?
Well, I took a screenshot, better than nothing.