🦔AI companies are bulk-buying rare books, scanning them through high-speed machines that cut the spines off, and shredding the originals. A service called ISBNdb facilitates orders of up to a million books and keeps buyers anonymous. Pre-2022 books are premium because they're free of AI-generated text. A federal judge ruled the practice is fair use because eliminating the original means only one copy exists at a time. Anthropic hired the former head of Google Books partnerships to obtain "all the books in the world."
My Take
This got to me. A bookseller told 404 Media that rare books with almost no surviving copies are being fed into this pipeline. Books that survived wars, fires, and centuries of handling are being shredded so an AI can learn to write a better marketing email.
ISBNdb's website literally says "'AI company destroys two million books' is not a headline that generates sympathy," and they still built an entire business around making it happen quietly. They offer NDAs as a feature. They coach clients to call it "digital preservation."
I've covered AI companies scraping the internet, torrenting libraries, and stealing music. This is worse because it's irreversible. You can re-upload a website. You can reprint a bestseller. You can't replace the last three copies of an 18th-century botanical text once someone shreds them for training data. And the judge said it's legal. So it's going to accelerate.
"We shred rare books and offer NDAs so nobody finds out" is a legitimate business model in 2026. What a timeline.
Hedgie🤗
"So you're telling me that an OS model that needs 2TB of RAM to run locally and is much more expensive to run than their previous model (aka needs far more compute for inference) is bearish for memory and semis?!?!""
"Yes, Dave."
"And that it also proves that scaling laws is alive and well, so the more memory and compute we use, the better the models get... And THAT is also bad for memory and semis?!?!"
"Right, Dave."
"So basically going forward we are not dependent on a few hyperscallers capex anymore for memory/semis demand and basically every company/person can have super intelligence running locally, with no risk of IP theft, as long as it spends a hefty amount into memory. And that's also not good for semis?""
"That's correct, Dave."
ChatGPT Work enabled me, for the first time, to automate repetitive but important tasks, as an solo founder. Important things I couldn't afford to hire someone to do but also couldn't automate with Excel and all those SaaS subscriptions... Now I have more time to focus on what matters! 🙏🏼🙌🏻
"So you're telling me that an OS model that needs 2TB of RAM to run locally and is much more expensive to run than their previous model (aka needs far more compute for inference) is bearish for memory and semis?!?!""
"Yes, Dave."
"And that it also proves that scaling laws is alive and well, so the more memory and compute we use, the better the models get... And THAT is also bad for memory and semis?!?!"
"Right, Dave."
"So basically going forward we are not dependent on a few hyperscallers capex anymore for memory/semis demand and basically every company/person can have super intelligence running locally, with no risk of IP theft, as long as it spends a hefty amount into memory. And that's also not good for semis?""
"That's correct, Dave."
"So you're telling me that an OS model that needs 2TB of RAM to run locally and is much more expensive to run than their previous model (aka needs far more compute for inference) is bearish for memory and semis?!?!""
"Yes, Dave."
"And that it also proves that scaling laws is alive and well, so the more memory and compute we use, the better the models get... And THAT is also bad for memory and semis?!?!"
"Right, Dave."
"So basically going forward we are not dependent on a few hyperscallers capex anymore for memory/semis demand and basically every company/person can have super intelligence running locally, with no risk of IP theft, as long as it spends a hefty amount into memory. And that's also not good for semis?""
"That's correct, Dave."
"So you're telling me that an OS model that needs 2TB of RAM to run locally and is much more expensive to run than their previous model (aka needs far more compute for inference) is bearish for memory and semis?!?!""
"Yes, Dave."
"And that it also proves that scaling laws is alive and well, so the more memory and compute we use, the better the models get... And THAT is also bad for memory and semis?!?!"
"Right, Dave."
"So basically going forward we are not dependent on a few hyperscallers capex anymore for memory/semis demand and basically every company/person can have super intelligence running locally, with no risk of IP theft, as long as it spends a hefty amount into memory. And that's also not good for semis?""
"That's correct, Dave."
Exactly!
But I think Kimi (and open source) goes one step further: it makes super intelligence access available to everyone with enough compute and that, by consequence, makes semis demand for diffuse and less dependent on hyperscalers (which is very healthy for semis).
https://t.co/GsdwZMoS4i
"So you're telling me that an OS model that needs 2TB of RAM to run locally and is much more expensive to run than their previous model (aka needs far more compute for inference) is bearish for memory and semis?!?!""
"Yes, Dave."
"And that it also proves that scaling laws is alive and well, so the more memory and compute we use, the better the models get... And THAT is also bad for memory and semis?!?!"
"Right, Dave."
"So basically going forward we are not dependent on a few hyperscallers capex anymore for memory/semis demand and basically every company/person can have super intelligence running locally, with no risk of IP theft, as long as it spends a hefty amount into memory. And that's also not good for semis?""
"That's correct, Dave."
"So you're telling me that an OS model that needs 2TB of RAM to run locally and is much more expensive to run than their previous model (aka needs far more compute for inference) is bearish for memory and semis?!?!""
"Yes, Dave."
"And that it also proves that scaling laws is alive and well, so the more memory and compute we use, the better the models get... And THAT is also bad for memory and semis?!?!"
"Right, Dave."
"So basically going forward we are not dependent on a few hyperscallers capex anymore for memory/semis demand and basically every company/person can have super intelligence running locally, with no risk of IP theft, as long as it spends a hefty amount into memory. And that's also not good for semis?""
"That's correct, Dave."