The Internet Archive and the Authors Alliance are producing a six-part series from the Future Knowledge podcast. The series, Vanishing Culture explores what happens when our shared cultural heritage disappears, and what we can do to preserve it. The first episode was published July 1st, 2026 starts out discussing the growing threat of cultural loss in the digital age:
From disappearing websites and deleted social media archives to lost films, books, music, and video games, Luca explores why culture vanishes and why it matters. He explains how copyright law, corporate control of digital platforms, and the shift from ownership to licensing are making it harder for libraries, archives, and communities to preserve cultural memory. Along the way, he shares stories that illustrate both the fragility of our digital heritage and the importance of preserving it: from a favorite YouTube recipe rescued by the Wayback Machine to the role cultural artifacts play in memory and identity. The conversation wraps on a positive note, with a look toward solutions and what creators, libraries, and everyday citizens can do to help ensure culture remains accessible for generations to come.
There is also a open access ebook with the same title, Vanishing Culture from this year by Luca Messarra, Chris Freeland, and Juliya Ziskina. It's available from the Internet Archive as EPub or PDF. See also another open access book, Walled Culture: A Journey Behind the Copyright Bricks, from 2022 by Glyn Moody.
In addition to all the other problems, digital goes away by default once active support stops. Contrast that with paper, microfilm, and microfiche which have to have someone expend time and effort (aka money) to be disposed of or, better, you can keep the copy you have regardless of other factors. Physical media are also decentralized, where as with digital information, there is often only a single copy in the world. Yes, the physical media disappear and rot if neglected, but it is a far cry from the out-like-a-light loss that digital information is afflicted by — unless proactive efforts are taken, such as the heroic efforts by the Internet Archive and other archiving services.
Previously:
(2024) Internet Archive Responds to Appellate Opinion in Hachette v. Internet Archive
(2024) It's the End of the Web as We Know It
(2022) Digitization Wars, Redux
(2020) On the Disappearance of Open Access Journals Over Time
(2014) The Importance of Information Preservation
Related Stories
As the world slowly moves towards a 100% digital existence, and increasingly consumes their information online, we run the risk of destroying our own legacy. Consider this hypothetical future narrative:
Historians are at a loss to explain the demise of the first pan-human civilisation, as although they agree that the populous dwindled and went almost extinct at around AD 3500, there seems to be no surviving written historical records that can be dated any later than circa AD 2000.
It can only be assumed that around this time, that there was a sudden uptake of illiteracy, maybe caused by a new religion or global-governmental policy. There are surviving references to an organization or group known as the Inter Nets. We can only guess at what this actually was, but the commonly accepted theory is that it was actually some type of wearable mesh harness that prevented humans of this era from actually writing anything down.
Sound ridiculous? I'm not so sure. As information is continually and fully migrated from the printed page and on to the Internet we lose the permanency that a book or ancient scroll brings. Paper and parchment when stored correctly can survive for thousands of years, and if not, the information held within can be transcribed in to replacement volumes when required. If it wasn't for the (well documented) fire that destroyed the Library of Alexandria we'd still have knowledge of the information that was contained there today.
David Rosenthal discusses the last 25 years of digital preservation efforts in regards to academic journals. It's a long-standing problem and discontinued journals continue to disappear from the Internet. Paper, microfilm, and microfiche are slow to degrade and are decentralized and distributed. Digital media are quick to disappear and the digital publications are usually only in a single physical place leading to single point of failure. It takes continuous, unbroken effort and money to keep digital publications accessible even if only one person or institution wishes to retain acccess. He goes into the last few decades of academic publishing and how we got here and then brings up 4 points abuot preservation, especially in regards to Open Access publishing.
Lesson 1: libraries won't pay enough to preserve even subscription content, let alone open-access content.
[...] Lesson 2: No-one, not even librarians, knows where most of the at-risk open-access journals are.
[...] Lesson 3: The production preservation pipeline must be completely automated.
[...] Lesson 4: Don't make the best be the enemy of the good. I.e. get as much as possible with the available funds, don't expect to get everything.
He posits that focus should be on the preservation of the individual articles, not the journals as units.
Previously:
(2020) Internet Archive Files Answer and Affirmative Defenses to Publisher Copyright Infringement Lawsuit
(2018) Vint Cerf: Internet is Losing its Memory
(2014) The Importance of Information Preservation
Digital librarian, Karen Coyle, has written about controlled digital lending (warning for PDF), where an artificial scarcity is applied to digital artifacts to limit concurrent access similar to the limitations that a finite number of objects exhibit in libraries' physical collections. This concept raises a lot of questions about not just copyright and digital versus physical, but also about reading in general. Some authors and publisher associations have already begun to object to controlled digital lending. However, few set aside misinformation and misdirection to allow for a proper, in-depth discussion of the issues.
We now have another question about book digitization: can books be digitized for the purpose of substituting remote lending in the place of the lending of a physical copy? This has been referred to as "Controlled Digital Lending (CDL)," a term developed by the Internet Archive for its online book lending services. The Archive has considerable experience with both digitization and providing online access to materials in various formats, and its Open Library site has been providing digital downloads of out of copyright books for more than a decade. Controlled digital lending applies solely to works that are presumed to be in copyright.
Controlled digital lending works like this: the Archive obtains and retains a physical copy of a book. The book is digitized and added to the Open Library catalog of works. Users can borrow the book for a limited time (2 weeks) after which the book "returns" to the Open Library. While the book is checked out to a user no other user can borrow that "copy." The digital copy is linked one-to-one with a physical copy, so if more than one copy of the physical book is owned then there is one digital loan available for each physical copy.
A great public resource is at risk of being destroyed:
The web has become so interwoven with everyday life that it is easy to forget what an extraordinary accomplishment and treasure it is. In just a few decades, much of human knowledge has been collectively written up and made available to anyone with an internet connection.
But all of this is coming to an end. The advent of AI threatens to destroy the complex online ecosystem that allows writers, artists, and other creators to reach human audiences.
To understand why, you must understand publishing. Its core task is to connect writers to an audience. Publishers work as gatekeepers, filtering candidates and then amplifying the chosen ones. Hoping to be selected, writers shape their work in various ways. This article might be written very differently in an academic publication, for example, and publishing it here entailed pitching an editor, revising multiple drafts for style and focus, and so on.
The internet initially promised to change this process. Anyone could publish anything! But so much was published that finding anything useful grew challenging. It quickly became apparent that the deluge of media made many of the functions that traditional publishers supplied even more necessary.
[...] The arrival of generative-AI tools has introduced a voracious new consumer of writing. Large language models, or LLMs, are trained on massive troves of material—nearly the entire internet in some cases. They digest these data into an immeasurably complex network of probabilities, which enables them to synthesize seemingly new and intelligently created material; to write code, summarize documents, and answer direct questions in ways that can appear human.
These LLMs have begun to disrupt the traditional relationship between writer and reader. Type how to fix broken headlight into a search engine, and it returns a list of links to websites and videos that explain the process. Ask an LLM the same thing and it will just tell you how to do it. Some consumers may see this as an improvement: Why wade through the process of following multiple links to find the answer you seek, when an LLM will neatly summarize the various relevant answers to your query? Tech companies have proposed that these conversational, personalized answers are the future of information-seeking. But this supposed convenience will ultimately come at a huge cost for all of us web users.
[...] If we continue in this direction, the web—that extraordinary ecosystem of knowledge production—will cease to exist in any useful form. Just as there is an entire industry of scammy SEO-optimized websites trying to entice search engines to recommend them so you click on them, there will be a similar industry of AI-written, LLMO-optimized sites. And as audiences dwindle, those sites will drive good writing out of the market. This will ultimately degrade future LLMs too: They will not have the human-written training material they need to learn how to repair the headlights of the future.
Originally spotted on Schneier on Security.
Related: Responsible Technology Use in the AI Age
Internet Archive Responds to Appellate Opinion in Hachette v. Internet Archive:
We are disappointed in today's opinion about the Internet Archive's digital lending of books that are available electronically elsewhere. We are reviewing the court's opinion and will continue to defend the rights of libraries to own, lend, and preserve books.
Take Action
Sign the open letter to publishers, asking them to restore access to the 500,000 books removed from our library: https://change.org/LetReadersRead
The Internet Archive Loses Its Appeal of a Major Copyright Case:
The Internet Archive has lost a major legal battle—in a decision that could have a significant impact on the future of internet history. Today, the US Court of Appeals for the Second Circuit ruled against the long-running digital archive, upholding an earlier ruling in Hachette v. Internet Archive that found that one of the Internet Archive's book digitization projects violated copyright law.
Notably, the appeals court's ruling rejects the Internet Archive's argument that its lending practices were shielded by the fair use doctrine, which permits for copyright infringement in certain circumstances, calling it "unpersuasive."
In March 2020, the Internet Archive, a San Francisco-based nonprofit, launched a program called the National Emergency Library, or NEL. Library closures caused by the pandemic had left students, researchers, and readers unable to access millions of books, and the Internet Archive has said it was responding to calls from regular people and other librarians to help those at home get access to the books they needed.
The NEL was an offshoot of an ongoing digital lending project called the Open Library, in which the Internet Archive scans physical copies of library books and lets people check out the digital copies as though they're regular reading material instead of ebooks. The Open Library lent the books to one person at a time—but the NEL removed this ratio rule, instead letting large numbers of people borrow each scanned book at once.
The NEL was the subject of backlash soon after its launch, with some authors arguing that it was tantamount to piracy. In response, the Internet Archive within two months scuttled its emergency approach and reinstated the lending caps. But the damage was done. In June 2020, major publishing houses, including Hachette, HarperCollins, Penguin Random House, and Wiley, filed the lawsuit.
(Score: 1, Interesting) by Anonymous Coward on Friday July 03, @07:29AM (6 children)
I've seen some sites go down because they got hacked and the owners just went f*ck it.
Example:
https://web.archive.org/web/20070430075431/http://www.rahoi.com/2006/03/may-i-take-your-order/ [archive.org]
Some other sites have gone down because of the lack of $$$. Possibly this one:
https://web.archive.org/web/20240527012127/https://www.sadanduseless.com/american-breakfast-recipes/ [archive.org]
Maybe that's why you don't go around teaching everyone especially random strangers how to block ads... Some has got to view the ads. If it ain't me or you or our loved ones, it should be others right?
(Score: 3, Insightful) by turgid on Friday July 03, @10:25AM (1 child)
Maybe that's why you don't go around teaching everyone especially random strangers how to block ads... Some has got to view the ads. If it ain't me or you or our loved ones, it should be others right?
There is clearly a market for websites with ads. People will look at them. I generally block ads. I hate them. They're distracting. If a website won't let me view it with ads blocked, it would have to be a pretty good website for me to disable my ad-blocker (and even then their stupid website still can't detect I've disabled it often). It's self-selecting. If your website really is that valuable to me I will unblock the ads or maybe subscribe. Otherwise I just don't care. I'm not your customer. It's a market.
I refuse to engage in a battle of wits with an unarmed opponent [wikipedia.org].
(Score: 2) by sgleysti on Friday July 03, @11:22PM
I absolutely hate advertisements. I would rather pay for something than be forced to view ads.
(Score: 4, Insightful) by aafcac on Friday July 03, @08:02PM (3 children)
I wouldn't personally bother blocking ads if they weren't a vector for malware, delaying the loading of the page and spying on me.
(Score: 2) by The Vocal Minority on Saturday July 04, @03:22AM (2 children)
Ads are almost always a vector for malware targeted at your wetware.
(Score: 2) by aafcac on Saturday July 04, @03:37AM (1 child)
They are, that's their point in existing, however the degree to which that's an issue depends in part on how well regulated they are. But, there's likely to always going to be more sites that are providing things worth seeing without paying an explicit fee. And ads address that fairly well.
The main issues tend to be the ones that I mentioned, on top of just how many there are and the lack of proper standards over what they're advertising.
(Score: 0) by Anonymous Coward on Friday July 10, @06:58AM
But AFAIK most ads run tons of hard to trust javascript. Technically most don't have to - they can do the ad bidding, redirect to winner's static ad etc all on server side without involving JS on the client.
(Score: 3, Interesting) by jb on Friday July 03, @08:19AM (2 children)
Paper, yes clearly. Microfilm & microfiche, not so much. In both cases the readers are getting few and far between these days.
The only thing I still have on microfiche is a copy of an old VAX/VMS manual (well, the half dozen or so most useful volumes anyway; never had the whole thing) and that only because printing it out would require far too much paper for something I hardly ever need to refer to. It's been a long time since I saw anything else on microfiche, anywhere, let alone a reader. Pretty safe to say that it was a format that didn't really survive beyond the 20th century.
Microfilm fared only slightly better. To the best of my knowledge the State Library is the only place in my state that has any microfilm readers still available for public use ... and that only because they invested a large amount of money towards the end of the 20th century to preserve on microfilm every newspaper published in the state in the preceding 150 or so years (mostly to save space by getting rid of the paper originals).
(Score: 4, Insightful) by turgid on Friday July 03, @10:32AM (1 child)
Microfilm? It's optical. You just read it with a light and some sort of microscope? You can put it through a hi-res digital optical scanner too. I suppose film degrades over time, so I'd get started on that digitisation sooner rather than later.
I refuse to engage in a battle of wits with an unarmed opponent [wikipedia.org].
(Score: 4, Interesting) by jb on Saturday July 04, @09:39AM
True, I've been known to repurpose old slide projectors (of the old 6cm open slot type) as makeshift microfiche readers before, which is pretty much the same as what you suggest.
But a proper reader makes the medium much more usable. Especially for microfilm, e.g.: some readers will let you skip forward/back an arbitrary number of frames far faster than you could possibly count them by hand; some have a key to print the current frame; some let you zoom/pan arbitrarily within a frame then automatically revert to frame-fills-the-screen when you change frames, etc. etc. Even the plain old rewind key (which almost every model had) saves a huge amount of time & effort.