Back to News
Advertisement
Advertisement

⚡ Community Insights

Discussion Sentiment

47% Positive

Analyzed from 4292 words in the discussion.

Trending Topics

#books#book#companies#copyright#copies#copy#more#destroy#rare#archive

Discussion (137 Comments)Read Original on HackerNews

ezfeabout 2 hours ago
I dislike these AI companies but let's be clear here: the copyright holders are the ones locking these books up. If they don't want to print more copies, then they could release the copyright on them.

Instead, they enforce the copyright and force AI companies to shred books they want to ingest.

edit: Also, an AI company would only ever care to purchase, scan, destroy a book once. Presumably many books have more than one copy.

RajT88about 2 hours ago
The articles I've read on this are not clear, but I strongly suspect "rare" is not the definition you and I probably use for the level of rarity of books actually being destroyed.

These are not going to be the kinds of books "The Ninth Gate" resolved around - truly one of a kind. It's not good they are destroying books, but they are books which do have other copies. Just perhaps not many.

card_zeroabout 2 hours ago
Quite possibly not many, and no copy held in any form by the copyright owner either. Say a few hundred copies of some obscure book from 40 years ago. They probably won't be erased from the face of the earth by the judicious and proportionate actions of, of a few, AI companies? Hmm.
scarmigabout 1 hour ago
The hypothetical "heroic figure goes and buys last copy of a 1962 guide to Ford cars to carefully maintain it in an appropriately climate controlled library" is vanishingly unlikely. A ten or a hundred or a thousand times to one, it just goes to the trash. At least here it gets scanned by the AI company.
bulbarabout 1 hour ago
> They probably won't be erased from the face of the earth by the judicious and proportionate actions of, of a few, AI companies?

I don't see why not. Pretty sure it's gonna happen. Doesn't matter if a hundred copies still exist somewhere, if access or discoverbility falls below a certain threshold, it doesn't matter, because those books become practically inaccessible to the world.

willy_kabout 2 hours ago
Is there a specific book from 40 years ago you have in mind? Asking out of curiosity.
alightsoulabout 2 hours ago
There's just a few copies in a single library worldwide which is probably a national or a university library
tptacekabout 2 hours ago
The 404 story suggested that these are largely vanity press books and instruction manuals for things no longer sold. Implying that these books were almost certainly headed for the recycling center had the AI companies not snatched them up.
enraged_camelabout 2 hours ago
Also, a lot of these "rare books" are stuff like TV programming magazines from October 1994.
card_zeroabout 2 hours ago
So, have you tried finding out what the programming was in October 1994? Or what cultural ephemera appeared in the TV guides of that era alongside the schedules? Either there's a copy for the week you want in an archive, or somebody's got one for sale, or most often neither. This can piss you off, if as it happened you had a reason to care.
unleadedabout 1 hour ago
Whose job is it to dictate what is and isn't worth saving?
runarbergabout 2 hours ago
At this scale, there are no guarantees of anything. There very likely will be unique copies in there. If these were expert archivists a lot of damage could be prevented, but given the malice and indifference of AI companies, there very likely will not be an expert archivist involved, and unique copies will be destroyed unceremoniously.
eruabout 2 hours ago
My personal wastebook at home is so rare, it's unique. That doesn't mean it needs preservation.
zmmmmmabout 2 hours ago
It is also the case that the copyright holders are often putting restrictions around use of electronic forms that are driving the desire to use physical copies. I doubt AI companies would use a single physical book if they could avoid it - absent the legal cloud over electronic rights.

I have no evidence but I can't help suspecting in part the publicity around this is driven in part by rights holders that want to force AI companies back to e-books where they can force them into licensing deals.

hn_throwaway_99about 2 hours ago
There is a whole legal saga here that is often misunderstood. Googling "Project Panama" should give more information.

The legal ruling from Judge William Alsup declared that if AI companies purchased the books legally and then copied them to their servers, it was fair use as a "transformative" operation, but the originals had to be destroyed in that case, because then there was only one copy still in existence (the one on Anthropic's servers):

From https://www.theguardian.com/commentisfree/2026/aug/05/anthro...

> Under US copyright law, the “fair use” doctrine allows you to make “transformative” use of copyrighted works without the owner’s permission. Anthropic took printed books and scanned them, “transforming” or remediating them into a new, electronic format. They then disposed of the original printed copy: the “destructive” part of destructive scanning. Along the way, Anthropic’s vendors had already sliced the spines and edges of the books, to scan them more easily before destroying them. “One replaced the other,” as Judge William Alsup wrote, noting: “There is no evidence that the new, digital copy was shown, shared, or sold outside the company.”

ivell5 minutes ago
Can they keep backup of the digital copy?
freejazzabout 2 hours ago
> I doubt AI companies would use a single physical book if they could avoid it

They just don't want to pay what the copyright holders want to charge

alightsoulabout 2 hours ago
Ai companies don't use ebooks, because they are more expensive than second hand books
breezybottomabout 2 hours ago
They absolutely do. Meta torrented 81 terabytes of ebooks. They just have no incentive to pay when the law looks the other way.
bawolffabout 2 hours ago
i imagine its because the doctrine of first sale does not apply to ebooks.
cm2012about 2 hours ago
Yes. I dont understand at all what AA is worried about. One copy of a book is no big deal? good will and used book stores throw out a lot more than that.
customguy10 minutes ago
> force AI companies to shred books they want to ingest.

Nothing forces them to shred books, they do it because it's slightly cheaper that way.

hparadiz7 minutes ago
There was a court case where they said that if they copied the books it's not fair use because they didn't own it but if they bought physical copies and then destroyed them then somehow it was fair use because it fell into the niche of personal backups. I forget the details but basically they buy one time prints and destroy them immediately.
postepowanieadm8 minutes ago
By destroying them they don't copy only convert them into another format.
jscdabout 1 hour ago
Sorry, is your stance seriously that authors and publishers should digitize and freely distribute their work, at their own expense?

Also, who’s forcing AI companies to “ingest” books in such a destructive way?

Also also, if there’s one thing I’ve learned from AI scrapers, it’s that they’d never scan the exact same thing multiple times at the expense of public access to the resource.

scarmigabout 1 hour ago
Relinquishing copyright does not imply any of the labor you're suggesting. It's the opposite: you're just committing not to perform the labor of pursuing legal action against someone who does digitize and freely distribute the work.

Anna's Archive, for one, would be more than happy to host at no cost to the author.

breezybottomabout 2 hours ago
They don't "force" anything. Trillion dollar AI companies and their owners have as much agency as book publishers.
parineumabout 2 hours ago
They are "forced" to do this because that's what they have to do to abide by copyright law. They can't create a digital duplicate without destroying the original.
freejazzabout 2 hours ago
If that's true, then why did they pirate so many books?
breezybottomabout 2 hours ago
Since when do AI companies care about copyright law? They're destroying them so their competitors can't use them.
tptacekabout 2 hours ago
To do what?
runarbergabout 2 hours ago
To not destroy rare books.
HedonicEscal8rabout 2 hours ago
If only this complaint was being posted by an organization ideologically opposed to copyright itself!
raincoleabout 2 hours ago
> Instead, they enforce the copyright and force AI companies to shred books they want to ingest.

What? Even if there are no copyright holders, the AI companies will still do scan'n'destroy because it's just cheap.

Are you expecting the authors/publishers to send digital copies to AI companies directly? Or expecting AI companies to preserve the physical copies indefinitely? Both are not gonna happen, copyrighted or not.

sophaclesabout 2 hours ago
If you're buying second hadn books by the lot, you'll get a lot of duplicates and its eaiser to scan wholesale and dedupe in the computers than it is to try to run a sorting operataion on "things".
jacobo37about 2 hours ago
this is plainly stupid ... many of these books are likely to have no current publisher nor any way to "reprint" the book. "ai" companies are simply burning our cultural context ...
rpdillonabout 2 hours ago
Wait: the entire premise of copyright is to prevent someone from publishing a book, and a competitor buys a copy, clones it, and sells copies way cheaper because they don't have to pay the author.

Now, in 2026, we're acting like cloning a published book is not technically feasible? That doesn't track. With publishing on-demand, it's easy to imagine a business with digital copies of all these works that they make available for print-on-demand.

The uncomfortable reality is that most of these books are nothing anyone cares about. Even the book sellers in the 404 story call them dead inventory.

Can we get some actual book titles into the discussion so we can focus on facts rather than speculation?

alightsoulabout 2 hours ago
This is not a technical problem at all. This is a copyright problem. Anthropic thought it was just a technical problem until they had to pay 1.5 billion after they lost a copyright court case

Op means a lot of those books were made before computers were used for that purpose and the publishers and probably authors no longer exist, so there is no digital copy to just reprint, unless someone scans it themselves and publishes it, risking copyright violation when done at large scale due to possible exceptions to this rule

sieve11 minutes ago
Physical books and digital content is special in that you can mostly archive their content almost permanently for cheap. Buildings, paintings, idols, living things, natural features of the environment ... not so much.

So the solution is:

- mandatory copyright registration and renewal with links to where the work can be acquired

- a blanket carve out for any non-commercial trust-style org so that they can scan books etc and keep the data on their servers. They should be able to issue digital membership cards for a fee so that patrons can access the archives. Any work that is "live" based on the registration database will be locked. All "dead" material can be shared with members.

In this way, a hundred digital preservation societies can bloom.

HedonicEscal8rabout 2 hours ago
The piracy organizations are playing 4D chess while everyone else is playing checkers. The irony of this entire situation - AI companies being legally required to shred books due to kafkaesque copyright laws, then used as a marketing tactic by Anna's Archive - is a work of art.

I support Anna's Archive, by the way. Information wants to be free.

Cider9986about 2 hours ago
You can donate with over 20 different payment methods after making an anonymous account.

https://annas-archive.gl/donate

Levitzabout 2 hours ago
It's an excellent play by them, using moral outrage to the benefit of the project. When life gives you lemons...
emtel32 minutes ago
“Rare books” usually refers to rare editions of books. Any books out there where there are only a few extent copies of the text itself, are probably not of very much interest or social value, since almost no one is able to read them, by definition.

If you think there is priceless knowledge locked up in books so rare that it is on the verge of being lost forever, then AI labs are not really the problem!

xvxvxabout 2 hours ago
Pretty funny that they just took Anna’s archive and ingested it.

As for the story: they make it sound like AI companies are buying up all existing copies of rare books and stealing the knowledge, which isn’t the case, as far as I know.

glimsheabout 2 hours ago
The very first paragraph is fascinating: "Several AI companies are acquiring large quantities of secondhand books through intermediaries, scanning and destroying them, all to obtain training data “untouched by machines” from before 2022."

Is the corpus of human knowledge useful for high quality AI training now essentially frozen in time? Also, how useful old books really are for AI training besides helping AI acquire knowledge about history?

npn13 minutes ago
No but with 100% clean data you can easily train a model to filter ai generated content.
eruabout 2 hours ago
> Is the corpus of human knowledge useful for high quality AI training now essentially frozen in time?

No. They also use lots of other methods to get training data.

jupp0rabout 2 hours ago
I highly doubt they destroy digital copies of the books after scanning. They will want to train their future models on the same content. So what prevents them from making these digital copies available to the public? Copyright!
tptacekabout 2 hours ago
These stories are weird, because actual professional specialized book dealers pulp books by the millions. People keep pointing out, and it doesn't seem to sink in, that model trainers only have use for a single copy of a book. Even if they were literally burning these books to spite you, they'd be destroying an infinitesimal fraction of the books the book trade already destroys.

It is not natural in the industry to preserve books! It's tricky to even give most books away. Our library has big donation boxes, and my understanding is: most of those books are destroyed.

The copyright thing I get, sort of (I mean, it's galling, because it's such a total special pleading argument from a cohort of people who otherwise have absolute contempt for copyright on anything other than code). The model trainers are getting away with something other people haven't gotten away with. OK, sure.

But this seems like the AI water use story, where the reality is that existing industries do whatever the bad thing is at scales cosmically larger than AI ever could, and we're zeroing in on this weird little slice of it that AI does. Like, let me know when we stop growing pecans in the California desert, and then we can talk?

imperfect_light2 minutes ago
I don't know how it works today, but 20 years ago bookstores wouldn't return unsold books (too expensive to ship) but would simply tear the covers off and throw them in the garbage.
Levitzabout 2 hours ago
The outraged people don't care. They hate AI, and so anything surrounding AI that can be evil is evil. Books are good, AI destroys books, AI is bad.

Furthermore, they like that AI is bad. Because they think it's bad, and being right feels good.

tptacekabout 2 hours ago
I think people genuinely don't get that book destruction is like a pretty natural part of the book lifecycle.
eruabout 2 hours ago
And the solution to the water issues can be found in any introductory textbook on the subject: a water price.

> People keep pointing out, and it doesn't seem to sink in, that model trainers only have use for a single copy of a book.

Please pardon the tangent: that's what always bothered me about the Borg in Star Trek. Why do they need to assimilate whole species? I'm sure there are enough volunteers in the federation that would join the Borg collective. Even a handful should be enough.

globular-toast14 minutes ago
They pulp books that have many copies surplus to requirement, not the last few copies in second hand book shops.

What if people like food more than AI? Have you considered that?

fenomasabout 1 hour ago
More and more I feel like anti-AI is a bigger bubble than AI. It seems like every week it expands into a new dimension - anti-Flock protesters tearing down years-old traffic cameras that were used for research into auto accidents, etc.

Like, the current thing in the news cycle is a poll that young people are now more worried than hopeful about AI. Which sounds scary, but my first thought is that one could find similar polls from the 80s and 90s about satanic cults or alien abduction..

silcoonabout 2 hours ago
As much as I hate piracy in a sector in financial crisis like book publishing (because Anna’s project is piracy), I hate even more what these large AI companies are doing: privatizing human knowledge.

On one side, there’s copyright law, which exists to support the work of creative people. “Information wants to be free” is bullshit spread by people who have never spent a minute in their lives trying to create something themselves. Artists need some form of reward.

On the other side, buying and destroying copies of rare books is quite scary. We would lose access to those books if they weren’t digitized. They are creating walls around knowledge that they acquired because there are no laws in place to protect authors.

This is scary, and it reminds me of Fahrenheit 451.

Do not believe Anna’s claims, since physical book sales are plummeting — the main source of income for writers — and shadow libraries are killing the incentive to write. But even more importantly, do not believe AI companies will help you discover and access knowledge.

We might end up with all of humanity’s books digitized and accessible for free, and LLMs capable of writing entire books for us. But there would be no human writers left.

In a world like that, what motivation would we still have to read?

Cider9986about 1 hour ago
>In a world like that, what motivation would we still have to read?

Why would a reduction in human writers cause a complete reduction in motivation to read? There's millions of books already written and it makes zero sense that people would stop writing. People write for hundreds of reasons other than to make money and they created literature before copyright was a thing.

msftgreedabout 1 hour ago
The human tradition is storytelling. The idea that storytelling was something a company could own and other's weren't allowed to tell is very, very new in our history.

People write without any profit motive today. It's weird of the OP to think of writing in such a narrow space as commercialization.

silcoonabout 1 hour ago
> Why would a reduction in human writers cause a complete reduction in motivation to read?

Because there would not be human written books about the present. All books would be about the past. But literature is not stuck in time. Today writers talk about topics and feelings that writers of the last century might never know or experienced. Many people read books to better understand the today world (non-fiction) and to better understand their today feelings (fiction).

> People write for hundreds of reasons other than to make money.

Agree, but most of the writing that we have from the past still came with some form of financial incentives. Shakespeare didn't write all of the compositions just because he wanted to express himself. He was making money with theater performances. Many religious writing got patronage by the church. Dante Alighieri had a career as politician, Plato came from an aristocratic family. Writing was reserved to elites because education was expensive and people had to work for food.

Today we are lucky because education is accessible and printing is cheap.

> they created literature before copyright was a thing.

Copyright wasn't a thing because replicating content was hard. Try to manually copy a book...

protocolture20 minutes ago
>On one side, there’s copyright law, which exists to support the work of creative people.

Which exists to enrich Disney and other large corps, while they hide behind artists as a human shield.

>Artists need some form of reward.

Right, as do artists who use the work of other artists as their starting point. Copyright holders aren't bill and bob artist, they are massive corporate trolls throwing around the weight of almost 100 years of our cultural heritage, sucking the marrow from its bones.

>On the other side, buying and destroying copies of rare books is quite scary. We would lose access to those books if they weren’t digitized. They are creating walls around knowledge that they acquired because there are no laws in place to protect authors.

The books are getting digitised into a permanent record of all our cultural heritage. It just sucks we don't have control over it. If only there was a way we could get them digitised AND control our cultural heritage. HMMMMMMM.

>Do not believe Anna’s claims, since physical book sales are plummeting — the main source of income for writers — and shadow libraries are killing the incentive to write.

The incentive to write is being killed by slop groups like 20Booksto50K and Kindle which predate AI by at least a decade. AI just lets them work faster.

>We might end up with all of humanity’s books digitized and accessible for free

Excellent

>But there would be no human writers left.

Unlikely, but there would definitely be no Disneys or Conde Nasts left, which is a massively pro social outcome.

>In a world like that, what motivation would we still have to read?

In a world with all books digitised and accessible to read? A huge huge huge incentive. I already partake if books are too expensive where I am. It would take me the rest of my life to read all the books I already want to read. What kind of inane dribble is the idea that copyright makes it interesting to read? I havent even read all of Howard and he's in the public domain (in cool countries at least)

Cider9986about 2 hours ago
The AI companies should work with the Internet Archive to release the digitized copies once the copyright expires.

Unrelated: So with this one copy BS are you not allowed to have backups of the data?

0x0000F8_about 2 hours ago
I entirely believe the litigation brought against Internet Archive was secretly sponsored by these exact organizations, because they want to monopolize information to train models.

No data => No models => No competition.

QuantumNomad_about 1 hour ago
IA was in hot water already even before ChatGPT came out.

> ChatGPT […] originally released on November 30, 2022

https://en.wikipedia.org/wiki/ChatGPT

> On March 24, 2020, following shutdowns caused by the COVID-19 pandemic, the Internet Archive opened the National Emergency Library, removing the waitlists used in Open Library and expanding access to these books for all readers. More than one user could borrow a book at the same time. Two months later, on June 1, the National Emergency Library (NEL) was met with a lawsuit from four book publishers. Two weeks after that, on June 16, the Internet Archive closed the NEL, and the prior Open Library CDL system resumed after the 12 weeks of NEL usage.

https://en.wikipedia.org/wiki/Hachette_v._Internet_Archive

visargaabout 1 hour ago
Yeah, great logic, that way they are sure there are no extra copies around.
thuruvabout 2 hours ago
I am baffled at these practices and somewhere confused on what's the end game here? monopoly on information? altering data? exclusive subscription based knowledge? Feels like we have welcomed the AI era with open hands hoping( at-least assuming) that data democracy will be there, yet feels like its a long road!
Advertisement
shaknaabout 2 hours ago
I wholeheartedly believe the AI controversy on destroying books is being stirred up by the companies themselves.

Copyright law requires you destroy a book, if you format shift it. If you digitise, you need to ensure its not a "copy" but that your one license went with the book.

So... If enough people complain, they get to pressure for copyright changes. Which will just so happen to have massive carveouts to let them do whatever they want.

altcognitoabout 2 hours ago
"You own a particular physical copy; you don't possess an abstract transferable 'one-copy license."

There is nothing that says you have to destroy something because you scanned it. This argument has been confusing me since I've seen this pop up.

Edit; despite the above, looking at the court documents from the Anthropic case, this is pretty close to what they were arguing: “we are just transferring the physical form we purchased, therefore it is legal.”

I still dont think there is a requirement to destroy the book, but since there isn’t a reason to store the book and they can’t sell it, they probably just took the cheapest route. It might be worth an argument that they only purchased the right to use the digital copies while the physical copies exist, but I’m in over my head from a copyright standpint

breezybottomabout 2 hours ago
Since when do AI companies care about the law? Most of their training data is pirated.
shaknaabout 2 hours ago
The headlines about destruction came not soon after they got rapped on the knuckles and told "no more pirating".
blooalienabout 1 hour ago
> Since when do AI companies care about the law? Most of their training data is pirated.

I imagine since the law recently cost one of them truckloads of money for their violations of it?

altcognitoabout 2 hours ago
What evidence do we have that they are "destroying" books?

I'm not saying this in their defense, but as someone who has worked at companies who has scanned books at scale, and generally speaking, I wasn't on site there, but I knew we/they were pretty delicate with the books. And while the kneejerk reaction might be "hey, why would they go through the effort?" -- my guess is that they are following or even hiring people that have done this process in the past (out of laziness) and just follow what works easiest. The literal machinery is not designed to destroy the books for various practical reasons. Books that are bound are easier to be kept in order and work with. Getting a flat scan is done with specialized tools, you don't need to put it on a plate (it would be too slow that way anyway)

All of the above is just to justify my question: Who knows that the books are being destroyed? (I also agree with the general sentiment that there's a good chance these books are just cheap and bulk, they aren't pulling one of a kind rare books.)

knowaveragejoeabout 1 hour ago
Yes, we know they're destroying the books. Whether this is actually a problem or another convenient "AI bad" trope remains to be seen.

https://www.techbrew.com/stories/2026/01/28/anthropic-ai-boo...

qwertytyyuuabout 2 hours ago
I'm sure the AI companies will retain scans of the books for training on newer models
jsphweidabout 2 hours ago
To clarify: Are they scanning and destroying a single copy of Book X or are they buying up all copies of book X, scanning it once, then destroying all copies of book X they can get their hand on?
Ekarosabout 2 hours ago
They are ordering books with ISBN. So I take that they are tracking what books they have scanned or pirated already and only picking up what they are missing. As just ordering mass bulk and getting 20 of the same encyclopaedia would be waste.

And I guess something like encyclopaedia would be good example of book they scan. At one point popular, but with most copies destroyed as no one actually wants them anymore.

ColdStreamabout 2 hours ago
The question I have is, do these companies keep copies of the scans after they have finished training on them? If so, then it isn't the worst outcome. Not great but at least the information is not completely destroyed forever just the original physical being of it.

Deeper thought however, eventually this will all be lost to time and I suspect that about 99% of all printed materials probably would never be read again simply due to the huge volume of it and sheer obscurity. Ernest Becker and his work 'The Denial of Death' might have some thoughts on this.

go to any second hand book store and just pick out something at random from the 1950's for instance, something about pottery or bird watching or whatever. The history of Bisbee Arizona, I don't know. Look up the author, see if they even left a trace of their work and the vast majority of the time they have already been forgotten to the great void of the universe. In the end, it all goes away. Clinging only creates pain.

I'm not saying that we should let them just do this, I am just saying that long term it is a tough battle to fight only to lose the war.

landgenootabout 2 hours ago
Isn't this a matter of regulation? I'm not sure about US, but in EU you have old houses/buildings that are protected. Sure, you can buy them, but you can't modify or destroy them (being cultural heritage).
eruabout 2 hours ago
Most old books aren't worth protecting, and the publishing industry destroys oodles of them as waste that no one wants.
ohthanksabout 1 hour ago
Being purchased and juiced for model weights is about as noble of an end as any book could hope for.
bawolffabout 2 hours ago
This whole situation is such a disgusting consequence of copyright law. The most frustrating part is that its so artificial. It is 100% the consequence of stupid laws.
derektank17 minutes ago
I mean, everything about intellectual property is kind of inherently artificial tbf.
blooalienabout 1 hour ago
> It is 100% the consequence of stupid laws.

More like the consequence of being unwilling to change stupid laws once the stupidity of them is discovered. Nope. Gotta double down on the stupidity instead...

luciana1uabout 2 hours ago
Someone should build the digital equivalent of a fire department. Train a model on the books, then if the originals get destroyed you still have the smoke.
globnomulousabout 2 hours ago
Is this a reference to Fahrenheit 451?
protocolture27 minutes ago
>It’s outrageous is that it’s legally permissible

No its not.

>but ethically, it’s an extremely serious crime against humanity.

Its only a crime if they dont also upload the scans to the internet.

>After AI companies massively scan and destroy physical books, they become the only ones in the world with digital copies. Knowledge is permanently monopolized on private servers.

This Law on the other hand is a crime against humanity.

>Anna’s Archive needs a plan to combat the destruction of physical books by AI companies.

No it doesnt.

>If every person scans a book, and there are 10 million volunteers worldwide, we can obtain 10 million pieces of invaluable wealth.

This however is an unvarnished good.

Look, piracy is the only realistic media archive we have.

We should be inviting, and working to eliminate opposition to, AI companies to assist in piracy.

This US v Them mentality is weird. If Anthropic has 10 million books scanned, get a copy. Thank them for the copy. Spread the copy.

Advertisement
WillAdamsabout 2 hours ago
"Whoever destroys a book destroys a link in the chain of human knowledge"

-- Thos. Jefferson

pkayeabout 1 hour ago
Public libraries destroy unsold book donations all the time. I often tried to give away some old books I have online and nobody wants them. Some of these books have some nostalgic value to me so I hate to see them just get destroyed so they just lie in my shed.
winridabout 1 hour ago
I have a copy of Michael Abrash's Graphics Programming Black Book (it's like 1k+ pages) with DESTROY written in red on the sides. I appreciate that someone saved it and sold it to me for cheap :)
userbinator8 minutes ago
That's an example of a "very much NOT rare" book, as you can easily find dozens of sources of scans online.
pkaye20 minutes ago
I have one of those I got second hand also. :)
whycombinetorabout 2 hours ago
It's giving Vishnu, but the world cannot exist without Shiva.
imperio59about 2 hours ago
Getting 10 million people to do anything is really, really hard. Getting 10 million people to spend hours scanning a book (which takes a really long time with a home scanner) sounds impossible :(
userbinator6 minutes ago
In the context of books, "scanning" is now more commonly something that should be called "camming" --- you simply point a camera at the book, and take a picture of every page.
alightsoulabout 2 hours ago
mycallabout 2 hours ago
Aren't AI companies all about the rare book auctions now?
c0lpan1cabout 2 hours ago
that's ironic, the url annas-archive.gl is blocked by my local DNS category for AI Threat Detection.
alightsoulabout 2 hours ago
Use a vpn
mplewisabout 2 hours ago
Can someone name a rare book that was destroyed as part of AI scanning? I want to know what kind of thing we're losing.
SanjayMehtaabout 2 hours ago
Google Books was a great resource until the lawyers got involved. I was able to find and download (one screenshot at a time) a rare family history. The author died 100 years ago. The published disappeared 80 years ago. But now Google has locked it behind a limited preview.

Google probably has the best collection of high quality scans, followed by the Hathi Trust. None of which are useable by anyone outside of those systems.

userbinator5 minutes ago
Does Anna's Archive have it now? They scrape tons of sources, Google Books and HathiTrust included.
tamimioabout 2 hours ago
I can imagine 100y from now, most if not all books and knowledge are in electronic format or even just as part of an AI, then a wild solar flare wipes out all electronics in a minute..
BrenBarnabout 2 hours ago
It's not a bad idea but we need a multi-pronged approach, with at least one other prong being "destroy the companies that are doing this".
partiallyproabout 2 hours ago
The hysteria around AI and data centers has hit a precipice. It's actually a bit embarrassing now. I am pretty sure there are foreign adversaries that are trying to stop the US, but I also really blame the AI companies for doing the most horrendous job imaginable in pitching AI to the public. Not a shock that people are against something that tech bros have claimed will destroy everyone's lives in the next 5 years. These books were probably going into a landfill without AI companies getting them, regardless. Tons and tons of books go into the garbage every day.
Advertisement