Unlike web scraping where spaghetti logic is required to follow abstracted JS links, archival of nostr events can be as simple as running a relay and mirroring blog content.[1][2] nostr does a lot of what NNTP did but with additional flexibility.
[1] https://logperiodic.com/rbsr.html#nostr
[2] https://github.com/hoytech/strfry/blob/master/docs/negentrop...