While clicking through some random Lemmy instances, I found one that’s due to be shut down in about a week — https://dmv.social. I’m trying to archive what I can onto the Wayback Machine, but I’m not sure what the most efficient way to go about it is.
At the moment, what I’ve been doing is going through each community and archiving each sort type (except the ones under a month, since the instance was locked a month ago) with capture outlinks enabled. But is there a more efficient way to do it? I know of the Internet Archives save from spreadsheet tool, which would probably work well, but I don’t know how I’d go about crawling all the links into a sitemap or csv or something similar. I don’t have the know-how to setup a web crawler/spider.
Any suggestions?
Maybe a plug-in for Lemmy server could be developed to automatically back up and / or restore instances from Arweave. Some protocol could be used to turn the instances into Json, which could then be uploaded as documents and parsed, or something like that. And then the Json could then be potentially restored. There might be many pages for a large instance, but they could perhaps be organized in a thoughtful and functional way.