

Discussion summary
Discussion revolves around the preservation and loss of digital history, with examples like halfbakery.com and technical links. Opinions vary on the morality and practicality of archiving everything, with concerns about resource allocation and rewriting history.
What the discussion says
- Some see digital preservation as valuable for history.
- Others argue indiscriminate archiving is inefficient.
- Concerns about the morality of keeping or deleting data.
- The role of AI and resource limits in archiving decisions.
“History is not only what 'official' resources want you to believe.”
“Those who cannot remember the past are condemned to repeat it.”
Join the discussion
Write your take first — we'll ask for email only when you're ready to publish.
- Hacker News
- I often run a linkchecker on my blog and substitute broken URL with links to the Wayback machine. Unfortunately, this is becoming quite difficult to detect broken URL as everybody is fighting bots. I am using linkchecker <https://github.com/linkchecker/linkchecker/> and it respects robots.txt but many sites are now serving 503 or various other codes.by vbernat
- This brought up an inadvertent benchmark I accidentally made between the big three AIs. I had them all research an old BBS i used to use, hunting for a DOOR game i played on it. I gave details about the bbs that i could remember and a pretty well defined description of the game and threw it at the deep researchers to see what they could fine.
ChatGPT gave me about a ten page report on who ran the bbs and the name of the game. When I looked into it the game was totally different and the guy named had nothing to do with the bbs. “Since these were both popular items at the time, I just inferred.” But it had fabricated the entire report. Nothing in it was true.
Gemini did the same thing but the report was about twenty pages. 100% hallucinated.
Claude said it couldn’t find any information.
Best advertisement I’ve ever lived.
I still hunt for the door game today….
by conception - Anyone in their late 30s or early 40s should be grateful. We got to be stupid teenagers on the internet without any of it going on our permanent record.by Venn1
- In modern times, archive.org is an international treasure.
Which of course means it's facing major opposition from capital interests.
Apparently no one ever thought an incoming presidential administration would literally wipe gigabytes of government funded research results off the web.
Now we see in bold type how precarious is our democracy...
by johnea - Long ago in Seattle there was a network of BBSs and the head board was called Rat City. They had a lot of work by local artists (mostly tracker files and digital artwork IIRC).
I have not been able to find a single hint of their existence. Everything about what was once a collection of artistic works, wiped from the earth.
We really need to do a better job managing our historical legacy.
by com2kid - "Do you do backups too, for example to guard against corrupt data getting mirrored across both copies, or accidental deletion?"
John Gonzalez, Internet Archive infrastructure lead, replied:
"We have done experiments to confirm that we can back up large portions of our corpus... but this is not a regular practice for us at this time."
https://blog.archive.org/2016/10/25/20000-hard-drives-on-a-m...
by badlibrarian - There was a website that I had quoted a long time ago. The author said something like "when the robots are taking over the world, don't panic. Buy a robot." I loved it. So I linked to it on my old blog. Then years later, I went to the source only to find that the page returned a 404. So I linked to the wayback machine instead. But then, it was removed from the archive.org. I can't even remember the name of the website at this point, just that it had the word "café" in it.
Anyway, all this to say that since there are no sources for this quote, then I'm the new original source. You can quote me on that.
https://imgflip.com/memegenerator/117370206/You-made-thisI-m...
by firefoxd