I’m still waiting to see what an example of a “Rare Book” is supposed to be. Are these first edition copies of To Kill A Mockingbird and Ulysses? Or are we just talking about books by new authors that were never widely distributed or reprinted, because their sales numbers were no good.
There are so many niche instructional books that are now out of print. If these books were scanned (yes, destroyed in the process) and then added to a free and open book archive for all to access, I wouldn’t have a problem with it. I have a problem with these AI training centers because all that knowledge is just going into a back hole.
There are music theory books and penmanship books I’m actively hunting for through piracy and ebay because they are out of print and rapidly getting lost to time. If I find them, I plan to scan and share them, not for piracy but because the original authors works do not deserve to disappear just because they no longer generated profit.
That’s not even touching in the price. Sometimes they sit there untouched simply because whoever happens to have one of the few copies put a ridiculous price on it.
I’m not saying it needs to be free (I mean I do but that’s a separate argument). I’m saying many of these sellers want as much money as physically possibly with no real world basis for the prices, which is why they sit there.
There’s several old Welsh poetry books I’ve gone looking for only to find copies marked for a thousand bucks pricing out any regular person and leaving only companies with stupid money to burn or maybe one day a collector.
Well, the value in the books is that they (a) haven’t been digitized yet and (b) aren’t contaminated with AI autowriting.
The buyers that would find the most value in these books are, themselves, AI training companies.
But let me be clear, I’m against their destruction for private use.
shrug That’s how books are digitized. You break the spine, split out the individual pages, and run them through an industrial scanner. I guess we can go back to reading from scrolls to alleviate this step. Past that, Idk what the problem is.
The digitized copies should be made freely available, if it all possible.
I don’t hate this idea. But I might argue that the Library of Congress should be digitizing published works as part of the copywriting process anyway. And, in fairness, the LoC currently hosts 21 petabytes of digitally archived data across 91 million unique works in 470 languages.
This keeps getting floated as some kind of scandal. I see it compared to “The burning of the Library of Alexandra” over and over again. But it appears to be nothing more than another, more primitive form of data harvesting of documents barely more valuable than Reddit shitposts. Less Alexandra and more the graffiti scribbled across Pompeii.
It would likely be stuff under copyright or stuff out of copyright that wasn’t considered a priority in other book digitization efforts. 1930 is the current copyright-free publish date.
So probably not a prestige edition of a famous work, but when buying in bulk like this who knows.
I’m still waiting to see what an example of a “Rare Book” is supposed to be. Are these first edition copies of To Kill A Mockingbird and Ulysses? Or are we just talking about books by new authors that were never widely distributed or reprinted, because their sales numbers were no good.
I guess I’ll never know.
There are so many niche instructional books that are now out of print. If these books were scanned (yes, destroyed in the process) and then added to a free and open book archive for all to access, I wouldn’t have a problem with it. I have a problem with these AI training centers because all that knowledge is just going into a back hole.
There are music theory books and penmanship books I’m actively hunting for through piracy and ebay because they are out of print and rapidly getting lost to time. If I find them, I plan to scan and share them, not for piracy but because the original authors works do not deserve to disappear just because they no longer generated profit.
That’s not even touching in the price. Sometimes they sit there untouched simply because whoever happens to have one of the few copies put a ridiculous price on it.
I’m not saying it needs to be free (I mean I do but that’s a separate argument). I’m saying many of these sellers want as much money as physically possibly with no real world basis for the prices, which is why they sit there.
There’s several old Welsh poetry books I’ve gone looking for only to find copies marked for a thousand bucks pricing out any regular person and leaving only companies with stupid money to burn or maybe one day a collector.
One of the books in my hunt list is the one below. It originally sold for $16 but is now out of print. Notice the reseller’s price.
Yup, and it’s just gonna sit there and go up in price as no one buys it because it’s “rare” and out of print. I feel ya
The big tech companies will drive up the price of second-hand books, just as they drove up the price of computer hardware and electricity.
Rare meaning uncommon, hard to find, not many copies are available.
But that doesn’t necessarily mean valuable. These have been sitting on shelves in warehouses unsold for a long time. No one else wanted them.
But let me be clear, I’m against their destruction for private use. The digitized copies should be made freely available, if it all possible.
Well, the value in the books is that they (a) haven’t been digitized yet and (b) aren’t contaminated with AI autowriting.
The buyers that would find the most value in these books are, themselves, AI training companies.
shrug That’s how books are digitized. You break the spine, split out the individual pages, and run them through an industrial scanner. I guess we can go back to reading from scrolls to alleviate this step. Past that, Idk what the problem is.
I don’t hate this idea. But I might argue that the Library of Congress should be digitizing published works as part of the copywriting process anyway. And, in fairness, the LoC currently hosts 21 petabytes of digitally archived data across 91 million unique works in 470 languages.
This keeps getting floated as some kind of scandal. I see it compared to “The burning of the Library of Alexandra” over and over again. But it appears to be nothing more than another, more primitive form of data harvesting of documents barely more valuable than Reddit shitposts. Less Alexandra and more the graffiti scribbled across Pompeii.
It would likely be stuff under copyright or stuff out of copyright that wasn’t considered a priority in other book digitization efforts. 1930 is the current copyright-free publish date.
So probably not a prestige edition of a famous work, but when buying in bulk like this who knows.