Archiveis: simple Python wrapper for the archive.is capturing service
github.com
github.com
domain = "http://178.62.195.5"
save_url = urljoin(domain, "/submit/")
Huh? What's this IP? Why does it send stuff in cleartext?[edit] adding screenshot from passive dns lookup https://m.imgur.com/a/uW9eL2h
I recently open-sourced an archive.is package and command-line client written in Go:
https://jaytaylor.com/archive.is (aliased to https://github.com/jaytaylor/archive.is)
go get jaytaylor.com/archive.is
So far I've been using it and it's worked reasonably well :)One cool thing- I actually used this python package as a starting place for how to automate archive.is submissions.
---
RE: archive.is: The person running archive.is deserves a lot of credit, it is a remarkable system. It may not be immediately clear how how challenging it is to capture and bottle up (safely and _reliably_) the contents of arbitrary URLs until you actually try to make such a thing. Archive.is person has implemented it at scale, and plans to keep the content available indefinitely [0], all on their own dime.
Mad props.
Why go through so much trouble and financially support it, while it doesn't seem to bring anything to the maintainer to financially support it over time?
I have a fairly large bookmark collection and I want to archive them. Burdening archive.is with that seems a bit of a waste tbh, I'd rather host a screenshot and selfcontained HTML file myself without costing someone else money.
No way!
web.archive.org is defacto standard for snapshotting[0] sites[1]!
[0] http://web.archive.org/save/https://news.ycombinator.com/ite...
[1] http://web.archive.org/web/*/https://news.ycombinator.com/it...
See also http://www.gwern.net/Archiving-URLs for a description of their usage of it.
This is archiving that we're talking about, after all.